Fetching the paper…
Reading the bibliography…
Labelled data is the foundation of most natural language processing tasks.
Relevant Linguistics , 2nd edition
Paul Justice. 2006 · 2006
Earlier work this paper cites.
Mark-up barking up the wrong tree
Annie Zaenen. 2006 · 2006
Earlier work this paper cites.
Subjective natural language problems: Motivations, applications, characterizations, and implications
Cecilia Ovesdotter Alm. 2011 · 2011
Earlier work this paper cites.
Learning whom to trust with MACE
Dirk Hovy, Taylor Berg-Kirkpatrick, Ashish Vaswani, and Eduard Hovy. 2013 · 2013
Earlier work this paper cites.
Truth is a lie: Crowd truth and the seven myths of human annotation
Lora Aroyo and Chris Welty. 2015 · 2015
Earlier work this paper cites.
Noise or additional information? leveraging crowdsource annotation item agreement for natural language tasks
Emily Jamison and Iryna Gurevych. 2015 · 2015
Earlier work this paper cites.
Ethics: Theory and Practice , 11th edition
Jacques P Thiroux and Keith W Krasemann. 2015 · 2015
Earlier work this paper cites.
Defining hate speech
Andrew Sellars. 2016 · 2016
Earlier work this paper cites.
Are you a racist or am I seeing things? Annotator influence on hate speech detection on Twitter
Zeerak Talat. 2016 · 2016
Earlier work this paper cites.
Hateful symbols or hateful people? Predictive features for hate speech detection on Twitter
Zeerak Talat and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Improving crowdsourced label quality using noise correction
Jing Zhang, Victor S Sheng, Tao Li, and Xindong Wu. 2017 · 2017
Earlier work this paper cites.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M. Bender and Batya Friedman. 2018 · 2018
Earlier work this paper cites.
Addressing Age-Related Bias in Sentiment Analysis , page 1–14. Association for Computing Machinery, New York, NY, USA
Mark Diaz, Isaac Johnson, Amanda Lazar, Anne Marie Piper, and Darren Gergle. 2018 · 2018
Earlier work this paper cites.
Large scale crowdsourcing and characterization of Twitter abusive behavior
Antigoni Maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Earlier work this paper cites.
Sentiment analysis: It’s complicated!
Kian Kenyon-Dean, Eisha Ahmed, Scott Fujimoto, Jeremy Georges-Filteau, Christopher Glasz, Barleen Kaur, Auguste Lalande, Shruti Bhanderi, Robert Belfer, Nirmal Kanagasabai, Roman Sarrazingendron, Rohit Verma, and Derek Ruths. 2018 · 2018
Earlier work this paper cites.
Comparing Bayesian models of annotation
Silviu Paun, Bob Carpenter, Jon Chamberlain, Dirk Hovy, Udo Kruschwitz, and Massimo Poesio. 2018 · 2018
Earlier work this paper cites.
Semeval-2019 task 5: Multilingual detection of hate speech against immigrants and women in Twitter
Valerio Basile, Cristina Bosco, Elisabetta Fersini, Debora Nozza, Viviana Patti, Francisco Manuel Rangel Pardo, Paolo Rosso, and Manuela Sanguinetti. 2019 · 2019
Cited alongside, same era.
Are we modeling the task or the annotator? an investigation of annotator bias in natural language understanding datasets
Mor Geva, Yoav Goldberg, and Jonathan Berant. 2019 · 2019
Cited alongside, same era.
Inherent disagreements in human textual inferences
Ellie Pavlick and Tom Kwiatkowski. 2019 · 2019
Cited alongside, same era.
Online hate ratings vary by extremes: A statistical analysis
Joni Salminen, Hind Almerekhi, Ahmed Mohamed Kamel, Soon-gyo Jung, and Bernard J. Jansen. 2019 · 2019
Cited alongside, same era.
The risk of racial bias in hate speech detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A. Smith. 2019 · 2019
Cited alongside, same era.
Toward a perspectivist turn in ground truthing for predictive computing
Valerio Basile, Federico Cabitza, Andrea Campagner, and Michael Fell. 2021a · 2021
Closest in time.
ConvAbuse: Data, analysis, and benchmarks for nuanced detection in conversational AI
Amanda Cercas Curry, Gavin Abercrombie, and Verena Rieser. 2021 · 2021
Closest in time.
Beyond black & white: Leveraging annotator disagreement via soft-label multi-task learning
Tommaso Fornaciari, Alexandra Uma, Silviu Paun, Barbara Plank, Dirk Hovy, and Massimo Poesio. 2021 · 2021
Closest in time.
The disagreement deconvolution: Bringing machine learning performance metrics in line with reality
Mitchell L Gordon, Kaitlyn Zhou, Kayur Patel, Tatsunori Hashimoto, and Michael S Bernstein. 2021 · 2021
Closest in time.
Decoding emojis and defining ’support’: Facebook’s rules for content revealed
Alex Hern. 2021 · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Predicting the type and target of offensive posts in social media
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019 · 2019
Cited alongside, same era.
Modeling annotator perspective and polarized opinions to improve hate speech detection
Sohail Akhtar, Valerio Basile, and Viviana Patti. 2020 · 2020
Cited alongside, same era.
Identifying and measuring annotator bias based on annotators’ demographic characteristics
Hala Al Kuwatly, Maximilian Wich, and Georg Groh. 2020 · 2020
Cited alongside, same era.
A novel methodology for developing automatic harassment classifiers for Twitter
Ishaan Arora, Julia Guo, Sarah Ita Levitan, Susan McGregor, and Julia Hirschberg. 2020 · 2020
Cited alongside, same era.
I feel offended, don’t be abusive! implicit/explicit messages in offensive and abusive language
Tommaso Caselli, Valerio Basile, Jelena Mitrović, Inga Kartoziya, and Michael Granitzer. 2020 · 2020
Cited alongside, same era.
Detecting stance in media on global warming
Yiwei Luo, Dallas Card, and Dan Jurafsky. 2020 · 2020
Cited alongside, same era.
Beneath the tip of the iceberg: Current challenges and new directions in sentiment analysis research
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder, and Rada Mihalcea. 2020 · 2020
Cited alongside, same era.
Understanding international perceptions of the severity of harmful content online
Jialun Aaron Jiang, Morgan Klaus Scheuerman, Casey Fiesler, and Jed R Brubaker. 2021 · 2021
Closest in time.
Designing toxic content classification for a diversity of perspectives
Deepak Kumar, Patrick Gage Kelley, Sunny Consolvo, Joshua Mason, Elie Bursztein, Zakir Durumeric, Kurt Thomas, and Michael Bailey. 2021 · 2021
Closest in time.
Agreeing to disagree: Annotating offensive language datasets with annotators’ disagreement
Elisa Leonardelli, Stefano Menini, Alessio Palmero Aprosio, Marco Guerini, and Sara Tonelli. 2021 · 2021
Closest in time.
Confident learning: Estimating uncertainty in dataset labels
Curtis Northcutt, Lu Jiang, and Isaac Chuang. 2021 · 2021
Closest in time.
On releasing annotator-level labels and information in datasets
Vinodkumar Prabhakaran, Aida Mostafazadeh Davani, and Mark Diaz. 2021 · 2021
Closest in time.
HateCheck: Functional tests for hate speech detection models
Paul Röttger, Bertie Vidgen, Dong Nguyen, Zeerak Talat, Helen Margetts, and Janet Pierrehumbert. 2021 · 2021
Closest in time.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2021 · 2021
Closest in time.
SemEval-2021 task 12: Learning with disagreements
Alexandra Uma, Tommaso Fornaciari, Anca Dumitrache, Tristan Miller, Jon Chamberlain, Barbara Plank, Edwin Simpson, and Massimo Poesio. 2021 · 2021
Closest in time.
Introducing CAD: the contextual abuse dataset
Bertie Vidgen, Dong Nguyen, Helen Margetts, Patricia Rossini, and Rebekah Tromble. 2021a · 2021
Closest in time.
Annotating online misogyny
Philine Zeinert, Nanna Inie, and Leon Derczynski. 2021 · 2021
Closest in time.
Jury learning: Integrating dissenting voices into machine learning models
Mitchell L Gordon, Michelle S Lam, Joon Sung Park, Kayur Patel, Jeffrey T Hancock, Tatsunori Hashimoto, and Michael S Bernstein. 2022 · 2022
Closest in time.