Fetching the paper…
Reading the bibliography…
Detecting online hate is a complex task, and low-performing models have harmful consequences when used for sensitive applications such as content moderation.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
HuggingFace’s Transformers: State-of-the-art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 1910
Earlier work this paper cites.
The Measurement of Observer Agreement for Categorical Data
J. Richard Landis and Gary G. Koch. 1977 · 1977
Earlier work this paper cites.
Grounded theory research: Procedures, canons, and evaluative criteria
Juliet M. Corbin and Anselm Strauss. 1990 · 1990
Earlier work this paper cites.
Black-Box Testing: Techniques for Functional Testing of Software and Systems
Boris Beizer. 1995 · 1995
Earlier work this paper cites.
Free-marginal multirater kappa: An alternative to fleiss
Justas Randolph. 2005 · 2005
Earlier work this paper cites.
DeBERTa: Decoding-enhanced BERT with Disentangled Attention
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2021 · 2006
Earlier work this paper cites.
Yukun Zhu, Ryan Kiros, Richard Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Earlier work this paper cites.
How Cosmopolitan Are Emojis? Exploring Emojis Usage and Meaning over Different Languages with Distributional Semantics
Francesco Barbieri, German Kruszewski, Francesco Ronzano, and Horacio Saggion. 2016 · 2016
Earlier work this paper cites.
Evidencing the harms of hate speech
Katharine Gelber and Luke McNamara. 2016 · 2016
Earlier work this paper cites.
Equality of Opportunity in Supervised Learning
Moritz Hardt, Eric Price, and Nati Srebro. 2016 · 2016
Earlier work this paper cites.
A global analysis of emoji usage
Nikola Ljubešić and Darja Fišer. 2016 · 2016
Earlier work this paper cites.
Analyzing the targets of hate in online social media
Leandro Silva, Mainack Mondal, Denzil Correa, Fabricio Benevenuto, and Ingmar Weber. 2016 · 2016
Earlier work this paper cites.
Automated Hate Speech Detection and the Problem of Offensive Language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm
Bjarke Felbo, Alan Mislove, Anders Søgaard, Iyad Rahwan, and Sune Lehmann. 2017 · 2017
Earlier work this paper cites.
Understanding Abuse: A Typology of Abusive Language Detection Subtasks
Zeerak Talat, Thomas Davidson, Dana Warmsley, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
A Reductions Approach to Fair Classification
Alekh Agarwal, Alina Beygelzimer, Miroslav Dudík, John Langford, and Hanna Wallach. 2018 · 2018
Earlier work this paper cites.
“The C-Word” Meets “the N-Word”: The Slur-Once-Removed and the Discursive Construction of “Reverse Racism”
Anna Bax. 2018 · 2018
Cited alongside, same era.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M. Bender and Batya Friedman. 2018 · 2018
Cited alongside, same era.
The Emoji Report
Brandwatch. 2018 · 2018
Cited alongside, same era.
Measuring and Mitigating Unintended Bias in Text Classification
Lucas Dixon, John Li, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman. 2018 · 2018
Cited alongside, same era.
Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior
Antigoni-Maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Cited alongside, same era.
Annotating emoticons and emojis in a german-danish social media corpus for hate speech research
Eckhard Bick. 2020 · 2020
Later among the works it cites.
Hybrid Emoji-Based Masked Language Models for Zero-Shot Abusive Language Detection
Michele Corazza, Stefano Menini, Elena Cabrio, Sara Tonelli, and Serena Villata. 2020 · 2020
Later among the works it cites.
Evaluating Models’ Local Decision Boundaries via Contrast Sets
Matt Gardner, Yoav Artzi, Victoria Basmov, Jonathan Berant, Ben Bogin, Sihao Chen, Pradeep Dasigi, Dheeru Dua, Yanai Elazar, Ananth Gottumukkala, Nitish Gupta, Hannaneh Hajishirzi, Gabriel Ilharco, Daniel Khashabi, Kevin Lin, Jiangming Liu, Nelson F. Liu, Phoebe Mulcaire, Qiang Ning, Sameer Singh, Noah A. Smith, Sanjay Subramanian, Reut Tsarfaty, Eric Wallace, Ally Zhang, and Ben Zhou. 2020 · 2020
Later among the works it cites.
Content moderation, AI, and the question of scale
Tarleton Gillespie. 2020 · 2020
Later among the works it cites.
Racist comments after Euro 2020: Saka, Sancho and Rashford racially abused online after England defeat
Alastair Jamieson. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lisbon Emoji and Emoticon Database (LEED): Norms for emoji and emoticons in seven evaluative dimensions
David Rodrigues, Marília Prada, Rui Gaspar, Margarida V. Garrido, and Diniz Lopes. 2018 · 2018
Cited alongside, same era.
Fair Regression: Quantitative Definitions and Reduction-Based Algorithms
Alekh Agarwal, Miroslav Dudik, and Zhiwei Steven Wu. 2019 · 2019
Cited alongside, same era.
New Modality: Emoji Challenges in Prediction, Anticipation, and Retrieval
Spencer Cappallo, Stacey Svetlichnaya, Pierre Garrigues, Thomas Mensink, and Cees G. M. Snoek. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Build it Break it Fix it for Dialogue Safety: Robustness from Adversarial Human Attack
Emily Dinan, Samuel Humeau, Bharath Chintagunta, and Jason Weston. 2019 · 2019
Cited alongside, same era.
Identification of hate speech and abusive language on indonesian Twitter using the Word2vec, part of speech and emoji features
Muhammad Okky Ibrohim, Muhammad Akbar Setiadi, and Indra Budi. 2019 · 2019
Cited alongside, same era.
Attending the emotions to detect online abusive language
Niloofar Safi Samghabadi, Afsheen Hatami, Mahsa Shafaei, Sudipta Kar, and Thamar Solorio. 2019 · 2019
Cited alongside, same era.
BERTweet: A pre-trained language model for English Tweets
Dat Quoc Nguyen, Thanh Vu, and Anh Tuan Nguyen. 2020 · 2020
Later among the works it cites.
Beyond Accuracy: Behavioral Testing of NLP Models with CheckList
Marco Tulio Ribeiro, Tongshuang Wu, Carlos Guestrin, and Sameer Singh. 2020 · 2020
Later among the works it cites.
The Curse of Performance Instability in Analysis Datasets: Consequences, Source, and Suggestions
Xiang Zhou, Yixin Nie, Hao Tan, and Mohit Bansal. 2020 · 2020
Later among the works it cites.
Dynabench: Rethinking Benchmarking in NLP
Douwe Kiela, Max Bartolo, Yixin Nie, Divyansh Kaushik, Atticus Geiger, Zhengxuan Wu, Bertie Vidgen, Grusha Prasad, Amanpreet Singh, Pratik Ringshia, Zhiyi Ma, Tristan Thrush, Sebastian Riedel, Zeerak Waseem, Pontus Stenetorp, Robin Jia, Mohit Bansal, Christopher Potts, and Adina Williams. 2021 · 2021
Closest in time.
Dynaboard: An Evaluation-As-A-Service Platform for Holistic Next-Generation Benchmarking
Zhiyi Ma, Kawin Ethayarajh, Tristan Thrush, Somya Jain, Ledell Wu, Robin Jia, Christopher Potts, Adina Williams, and Douwe Kiela. 2021 · 2021
Closest in time.
DynaSent: A Dynamic Benchmark for Sentiment Analysis
Christopher Potts, Zhengxuan Wu, Atticus Geiger, and Douwe Kiela. 2021 · 2021
Closest in time.
HateCheck: Functional Tests for Hate Speech Detection Models
Paul Röttger, Bertie Vidgen, Dong Nguyen, Zeerak Waseem, Helen Margetts, and Janet Pierrehumbert. 2021 · 2021
Closest in time.
Introducing CAD: The Contextual Abuse Dataset
Bertie Vidgen, Dong Nguyen, Helen Margetts, Patricia Rossini, and Rebekah Tromble. 2021a · 2021
Closest in time.
Exploiting Emojis for Abusive Language Detection
Michael Wiegand and Josef Ruppenhofer. 2021 · 2021
Closest in time.
Handling and Presenting Harmful Text
Leon Derczynski, Hannah Rose Kirk, Abeba Birhane, and Bertie Vidgen. 2022 · 2022
Closest in time.
Two Contrasting Data Annotation Paradigms for Subjective NLP Tasks
Paul Röttger, Bertie Vidgen, Dirk Hovy, and Janet B. Pierrehumbert. 2022 · 2022
Closest in time.