Fetching the paper…
Reading the bibliography…
Online discussions, panels, talk page edits, etc., often contain harmful conversational content i.e., hate speech, death threats and offensive language, especially towards certain demographic groups.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Offensive language detection on video live streaming chat
Zhiwei Gao, Shuntaro Yada, Shoko Wakamiya, and Eiji Aramaki. 2020 · 1940
Earlier work this paper cites.
Logistic regression
Raymond E Wright. 1995 · 1995
Earlier work this paper cites.
Logistic regression
David G Kleinbaum, K Dietz, M Gail, Mitchel Klein, and Mitchell Klein. 2002 · 2002
Earlier work this paper cites.
NLTK: The natural language toolkit
Edward Loper and Steven Bird. 2002 · 2002
Earlier work this paper cites.
Applied logistic regression analysis , volume 106
Scott Menard. 2002 · 2002
Earlier work this paper cites.
What is a support vector machine?
William S Noble. 2006 · 2006
Earlier work this paper cites.
Hatebert: Retraining BERT for abusive language detection in english
Tommaso Caselli, Valerio Basile, Jelena Mitrovic, and Michael Granitzer. 2020 · 2010
Earlier work this paper cites.
Searching for a mate: The rise of the internet as a social intermediary
Michael J. Rosenfeld and Reuben J. Thomas. 2012 · 2012
Earlier work this paper cites.
Hateful symbols or hateful people? predictive features for hate speech detection on Twitter
Zeerak Waseem and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Su Lin Blodgett and Brendan O’Connor. 2017 · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Understanding abuse: A typology of abusive language detection subtasks
Zeerak Waseem, Thomas Davidson, Dana Warmsley, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Large scale crowdsourcing and characterization of twitter abusive behavior
Antigoni Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Cited alongside, same era.
Gender bias in coreference resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard, and Benjamin Van Durme. 2018 · 2018
Cited alongside, same era.
Racial bias in hate speech and abusive language detection datasets
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
SemEval-2019 task 6: Identifying and categorizing offensive language in social media (OffensEval)
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019 · 2019
Cited alongside, same era.
Detect all abuse! toward universal abusive language detection models
Kunze Wang, Dong Lu, Caren Han, Siqu Long, and Josiah Poon. 2020 · 2020
Later among the works it cites.
Demoting racial bias in hate speech detection
Mengzhou Xia, Anjalie Field, and Yulia Tsvetkov. 2020 · 2020
Later among the works it cites.
Differential tweetment: Mitigating racial dialect bias in harmful tweet detection
Ari Ball-Burack, Michelle Seng Ah Lee, Jennifer Cobbe, and Jatinder Singh. 2021 · 2021
Later among the works it cites.
RedditBias: A real-world resource for bias evaluation and debiasing of conversational language models
Soumya Barikeri, Anne Lauscher, Ivan Vulić, and Goran Glavaš. 2021 · 2021
Later among the works it cites.
On the gap between adoption and understanding in NLP
Federico Bianchi and Dirk Hovy. 2021 · 2021
Later among the works it cites.
Dataset for identification of homophobia and transophobia in multilingual youtube comments
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Counterfactual data augmentation for mitigating gender stereotypes in languages with rich morphology
Ran Zmigrod, Sabrina J. Mielke, Hanna Wallach, and Ryan Cotterell. 2019 · 2019
Cited alongside, same era.
Language (technology) is power: A critical survey of “bias” in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé III, and Hanna Wallach. 2020 · 2020
Cited alongside, same era.
Jamell Dacon and Jiliang Tang. 2021 · 2020
Cited alongside, same era.
Toxic, hateful, offensive or abusive? what are we really classifying? an empirical analysis of hate speech datasets
Paula Fortuna, Juan Soler, and Leo Wanner. 2020 · 2020
Cited alongside, same era.
XHate-999: Analyzing and detecting abusive language across domains and languages
Goran Glavaš, Mladen Karan, and Ivan Vulić. 2020 · 2020
Cited alongside, same era.
Detoxify
Laura Hanu and Unitary team. 2020 · 2020
Cited alongside, same era.
Does gender matter? towards fairness in dialogue systems
Haochen Liu, Jamell Dacon, Wenqi Fan, Hui Liu, Zitao Liu, and Jiliang Tang. 2020 · 2020
Cited alongside, same era.
Bharathi Raja Chakravarthi, Ruba Priyadharshini, Rahul Ponnusamy, Prasanna Kumar Kumaresan, Kayalvizhi Sampath, Durairaj Thenmozhi, Sathiyaraj Thangasamy, Rajendran Nallathambi, and John Philip McCrae. 2021 · 2021
Later among the works it cites.
Does gender matter in the news? detecting and examining gender bias in news articles
Jamell Dacon and Haochen Liu. 2021 · 2021
Later among the works it cites.
Ammus: A survey of transformer-based pretrained models in natural language processing
Katikapalli Subramanyam Kalyan, Ajit Rajasekharan, and Sivanesan Sangeetha. 2021 · 2021
Later among the works it cites.
Rebuilding trust: Queer in ai approach to artificial intelligence risk management
Organizers of QueerInAI, Ashwin S, William Agnew, Hetvi Jethwani, and Arjun Subramonian. 2021 · 2021
Later among the works it cites.
“short is the road that leads from fear to hate”: Fear speech in indian whatsapp groups
Punyajoy Saha, Binny Mathew, Kiran Garimella, and Animesh Mukherjee. 2021 · 2021
Later among the works it cites.
Towards generalisable hate speech detection: a review on obstacles and solutions
Wenjie Yin and Arkaitz Zubiaga. 2021 · 2021
Later among the works it cites.
Towards a deep multi-layered dialectal language analysis: A case study of african-american english
Jamell Dacon. 2022 · 2022
Closest in time.
Controlled analyses of social biases in wikipedia bios
Anjalie Field, Chan Young Park, Kevin Z. Lin, and Yulia Tsvetkov. 2022 · 2022
Closest in time.