Fetching the paper…
Reading the bibliography…
Since state-of-the-art approaches to offensive language detection rely on supervised learning, it is crucial to quickly adapt them to the continuously evolving scenario of social media.
Squibs and discussions: The kappa statistic: A second look
Barbara Di Eugenio and Michael Glass. 2004 · 2004
Earlier work this paper cites.
Computing reliability for coreference annotation
Rebecca J. Passonneau. 2004 · 2004
Earlier work this paper cites.
Toward a paradigm shift in social computing: The acp approach
Fei-Yue Wang. 2007 · 2007
Earlier work this paper cites.
Survey article: Inter-coder agreement for computational linguistics
Ron Artstein and Massimo Poesio. 2008 · 2008
Earlier work this paper cites.
Exploiting ‘subjective’ annotations
Dennis Reidsma and Rieks op den Akker. 2008 · 2008
Earlier work this paper cites.
Get another label? improving data quality and data mining using multiple, noisy labelers
Victor S. Sheng, Foster Provost, and Panagiotis G. Ipeirotis. 2008 · 2008
Earlier work this paper cites.
From annotator agreement to noise models
Beata Beigman Klebanov and Eyal Beigman. 2009 · 2009
Earlier work this paper cites.
Data quality from crowdsourcing: A study of annotation selection criteria
Pei-Yun Hsueh, Prem Melville, and Vikas Sindhwani. 2009 · 2009
Earlier work this paper cites.
Part-of-speech tagging for twitter: Annotation, features, and experiments
Kevin Gimpel, Nathan Schneider, Brendan O’Connor, Dipanjan Das, Daniel Mills, Jacob Eisenstein, Michael Heilman, Dani Yogatama, Jeffrey Flanigan, and Noah A Smith. 2010 · 2010
Earlier work this paper cites.
Evaluating the impact of coder errors on active learning
Ines Rehbein and Josef Ruppenhofer. 2011 · 2011
Earlier work this paper cites.
Learning whom to trust with MACE
Dirk Hovy, Taylor Berg-Kirkpatrick, Ashish Vaswani, and Eduard Hovy. 2013 · 2013
Earlier work this paper cites.
Difficult cases: From data to learning, and back
Beata Beigman Klebanov and Eyal Beigman. 2014 · 2014
Cited alongside, same era.
Learning part-of-speech taggers with inter-annotator agreement loss
Barbara Plank, Dirk Hovy, and Anders Søgaard. 2014 · 2014
Cited alongside, same era.
Truth is a lie: Crowd truth and the seven myths of human annotation
Lora Aroyo and Chris Welty. 2015 · 2015
Cited alongside, same era.
Noise or additional information? leveraging crowdsource annotation item agreement for natural language tasks
Emily Jamison and Iryna Gurevych. 2015 · 2015
Cited alongside, same era.
The social impact of natural language processing
Dirk Hovy and Shannon L. Spruit. 2016 · 2016
Cited alongside, same era.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Human uncertainty makes classification more robust
J. Peterson, R. Battleday, T. Griffiths, and O. Russakovsky. 2019 · 2019
Later among the works it cites.
The risk of racial bias in hate speech detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A. Smith. 2019 · 2019
Later among the works it cites.
Predicting the Type and Target of Offensive Posts in Social Media
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019 · 2019
Later among the works it cites.
Modeling annotator perspective and polarized opinions to improve hate speech detection
Sohail Akhtar, Valerio Basile, and Viviana Patti. 2020 · 2020
Later among the works it cites.
It’s the end of the gold standard as we know it. on the impact of pre-aggregation on the evaluation of highly subjective tasks
Valerio Basile. 2020 · 2020
Later among the works it cites.
Harmonization sometimes harms
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Large scale crowdsourcing and characterization of twitter abusive behavior
Antigoni-Maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Cited alongside, same era.
Sentence encoders on stilts: Supplementary training on intermediate labeled-data tasks
Jason Phang, Thibault Févry, and Samuel R Bowman. 2018 · 2018
Cited alongside, same era.
Deep learning from crowds
Filipe Rodrigues and Francisco C. Pereira. 2018 · 2018
Cited alongside, same era.
Crowdsourcing subjective tasks: The case study of understanding toxicity in online discussions
Lora Aroyo, Lucas Dixon, Nithum Thain, Olivia Redfield, and Rachel Rosen. 2019 · 2019
Cited alongside, same era.
A crowdsourced frame disambiguation corpus with ambiguity
Anca Dumitrache, Lora Aroyo, and Chris Welty. 2019 · 2019
Cited alongside, same era.
Manfred Klenner, Anne Göhring, and Michael Amsler. 2020 · 2020
Later among the works it cites.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A. Smith, and Yejin Choi. 2020 · 2020
Later among the works it cites.
Low resource sequence tagging with weak labels
Edwin Simpson, Jonas Pfeiffer, and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
SemEval-2020 Task 12: Multilingual Offensive Language Identification in Social Media (OffensEval 2020)
Marcos Zampieri, Preslav Nakov, Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Hamdy Mubarak, Leon Derczynski, Zeses Pitenis, and Çağrı Çöltekin. 2020 · 2020
Later among the works it cites.
The disagreement deconvolution: Bringing machine learning performance metrics in line with reality
Mitchell L. Gordon, Kaitlyn Zhou, Kayur Patel, Tatsunori Hashimoto, and Michael S. Bernstein. 2021 · 2021
Closest in time.