Fetching the paper…
Reading the bibliography…
We present a human-and-model-in-the-loop process for dynamically generating datasets and training better performing and more robust hate detection models.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Tackling online abuse: A survey of automated abuse detection methods
Pushkar Mishra, Helen Yannakoudakis, and Ekaterina Shutova. 2019 · 1908
Earlier work this paper cites.
Learning the difference that makes a difference with counterfactually-augmented data
Divyansh Kaushik, Eduard Hovy, and Zachary C Lipton. 2019 · 1909
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
Unsex me here: Revisiting sexism detection using psychological scales and adversarial samples
Mattia Samory, Indira Sen, Julian Kohne, Fabian Flock, and Claudia Wagner. 2020 · 2004
Earlier work this paper cites.
Detecting East Asian Prejudice on Social Media
Bertie Vidgen, Austin Botelho, David Broniatowski, Ella Guest, Matthew Hall, Helen Margetts, Rebekah Tromble, Zeerak Waseem, and Scott Hale. 2020 · 2005
Earlier work this paper cites.
Computing Inter-Rater Reliability for Observational Data: An Overview and Tutorial
Kevin A Hallgren. 2012 · 2012
Earlier work this paper cites.
Detecting hate speech on the world wide web
William Warner and Julia Hirschberg. 2012 · 2012
Earlier work this paper cites.
A method for taxonomy development and its application in information systems
Robert C Nickerson, Upkar Varshney, and Jan Muntermann. 2013 · 2013
Earlier work this paper cites.
Detecting threats of violence in online discussions using bigrams of important words
Hugo Lewi Hammer. 2014 · 2014
Earlier work this paper cites.
Recent research on dehumanization
N. Haslam and M. Stratemeyer. 2016 · 2016
Earlier work this paper cites.
Abusive Language Detection in Online User Content
Chikashi Nobata, Achint Thomas, Yashar Mehdad, Yi Chang, and Joel Tetreault. 2016 · 2016
Earlier work this paper cites.
Are you a racist or am I seeing things? annotator influence on hate speech detection on twitter
Zeerak Waseem. 2016 · 2016
Earlier work this paper cites.
Hateful Symbols or Hateful People? Predictive Features for Hate Speech Detection on Twitter
Zeerak Waseem and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Automated Hate Speech Detection and the Problem of Offensive Language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
A Large Labeled Corpus for Online Harassment Research
Jennifer Golbeck, Alicia A. Geller, Jayanee Thanki, Shalmali Naik, Kelly M. Hoffman, Derek Michael Wu, Alexandra Berlinger, Priyanka Vengataraman, Shivika Khare, Zahra Ashktorab, Marianna J. Martindale, Gaurav Shahane, Paul Cheakalos, Jenny Hottle, Siddharth Bhagwan, Raja Rajan Gunasekaran, Rajesh Kumar Gnanasekaran, Rashad O. Banjo, Piyush Ramachandran, Lisa Rogers, Kristine M. Rogers, Quint Gergory, Heather L. Nixon, Meghna Sardana Sarin, Zijian Wan, Cody Buntain, Ryan Lau, and Vichita Jienjitlert. 2017 · 2017
Earlier work this paper cites.
Deceiving Google’s Perspective API Built for Detecting Toxic Comments
Hossein Hosseini, Sreeram Kannan, Baosen Zhang, and Radha Poovendran. 2017 · 2017
Earlier work this paper cites.
A Survey on Hate Speech Detection using Natural Language Processing
Anna Schmidt and Michael Wiegand. 2017 · 2017
Earlier work this paper cites.
Understanding Abuse: A Typology of Abusive Language Detection Subtasks
Zeerak Waseem, Thomas Davidson, Dana Warmsley, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Vectors for counterspeech on Twitter
Lucas Wright, Derek Ruths, Kelly P Dillon, Haji Mohammad Saleem, and Susan Benesch. 2017 · 2017
Cited alongside, same era.
Automatic identification and classification of misogynistic language on Twitter
Maria Anzovino, Elisabetta Fersini, and Paolo Rosso. 2018 · 2018
Cited alongside, same era.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M. Bender and Batya Friedman. 2018 · 2018
Cited alongside, same era.
Measuring and Mitigating Unintended Bias in Text Classification
Lucas Dixon, John Li, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman. 2018 · 2018
Cited alongside, same era.
Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior
Antigoni-maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Cited alongside, same era.
Detection of Abusive Language: the Problem of Biased Datasets
Michael Wiegand, Josef Ruppenhofer, and Thomas Kleinbauer. 2019 · 2019
Later among the works it cites.
Predicting the type and target of offensive posts in social media
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019 · 2019
Later among the works it cites.
I Feel Offended, Don’t Be Abusive! Implicit/Explicit Messages in Offensive and Abusive Language
Tommaso Caselli, Valerio Basile, Jelena Mitrović, Inga Kartoziya, and Michael Granitzer. 2020 · 2020
Closest in time.
Evaluating NLP Models’ Local Decision Boundaries via Contrast Sets
Matt Gardner, Yoav Artzi, Victoria Basmova, Jonathan Berant, Ben Bogin, Sihao Chen, Pradeep Dasigi, Dheeru Dua, Yanai Elazar, Ananth Gottumukkala, Nitish Gupta, Hanna Hajishirzi, Gabriel Ilharco, Daniel Khashabi, Kevin Lin, Jiangming Liu, Nelson F. Liu, Phoebe Mulcaire, Qiang Ning, Sameer Singh, Noah A. Smith, Sanjay Subramanian, Reut Tsarfaty, Eric Wallace, Ally Zhang, and Ben Zhou. 2020 · 2020
Closest in time.
Contextualizing Hate Speech Classifiers with Post-hoc Explanation
Brendan Kennedy, Xisen Jin, Aida Mostafazadeh Davani, Morteza Dehghani, and Xiang Ren. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
All You Need is ”Love”: Evading Hate-speech Detection
Tommi Gröndahl, Luca Pajola, Mika Juuti, Mauro Conti, and N. Asokan. 2018 · 2018
Cited alongside, same era.
Cross-Domain Detection of Abusive Language Online
Mladen Karan and Jan Šnajder. 2018 · 2018
Cited alongside, same era.
Benchmarking Aggression Identification in Social Media
Ritesh Kumar, Atul Kr Ojha, Shervin Malmasi, and Marcos Zampieri. 2018 · 2018
Cited alongside, same era.
Detecting Offensive Language in Tweets Using Deep Learning
Georgios K. Pitsilis, Heri Ramampiaro, and Helge Langseth. 2018 · 2018
Cited alongside, same era.
Bridging the gaps: Multi task learning for domain transfer of hate speech detection
Zeerak Waseem, James Thorne, and Joachim Bingel. 2018 · 2018
Cited alongside, same era.
Sem Eval-2019 Task 5: Multilingual Detection of Hate Speech Against Immigrants and Women in Twitter
Valerio Basile, Cristina Bosco, Elisabetta Fersini, Debora Nozza, Viviana Patti, Francisco Manuel Rangel Pardo, Paolo Rosso, and Manuela Sanguinetti. 2019 · 2019
Cited alongside, same era.
CONAN - COunter NArratives through Nichesourcing: a Multilingual Dataset of Responses to Fight Online Hate Speech
Yi-Ling Chung, Elizaveta Kuzmenko, Serra Sinem Tekiroglu, and Marco Guerini. 2019 · 2019
Cited alongside, same era.
Closest in time.
Towards a Comprehensive Taxonomy and Large-Scale Annotated Corpus for Online Slur Usage
Jana Kurrek, Haji Mohammad Saleem, and Derek Ruths. 2020 · 2020
Closest in time.
Hatexplain: A benchmark dataset for explainable hate speech detection
Binny Mathew, Punyajoy Saha, Seid Muhie Yimam, Chris Biemann, Pawan Goyal, and Animesh Mukherjee. 2020 · 2020
Closest in time.
A Framework for the Computational Linguistic Analysis of Dehumanization
Julia Mendelsohn, Yulia Tsvetkov, and Dan Jurafsky. 2020 · 2020
Closest in time.
Adversarial NLI: A new benchmark for natural language understanding
Yixin Nie, Adina Williams, Emily Dinan, Mohit Bansal, Jason Weston, and Douwe Kiela. 2020 · 2020
Closest in time.
COLD: Annotation scheme and evaluation data set for complex offensive language in English
Alexis Palmer, Christine Carr, Melissa Robinson, and Jordan Sanders. 2020 · 2020
Closest in time.
Misogyny Detection in Twitter: a Multilingual and Cross-Domain Study
Endang Pamungkas, Valerio Basile, and Viviana Patti. 2020 · 2020
Closest in time.
Resources and benchmark corpora for hate speech detection: a systematic review
Fabio Poletto, Valerio Basile, Manuela Sanguinetti, Cristina Bosco, and Viviana Patti. 2020 · 2020
Closest in time.
Beyond accuracy: Behavioral testing of NLP models with CheckList
Marco Tulio Ribeiro, Tongshuang Wu, Carlos Guestrin, and Sameer Singh. 2020 · 2020
Closest in time.
Hatecheck: Functional tests for hate speech detection models
Paul Röttger, Bertram Vidgen, Dong Nguyen, Zeerak Waseem, Helen Margetts, and Janet Pierrehumbert. 2020 · 2020
Closest in time.
Developing an online hate classifier for multiple social media platforms
Joni Salminen, Maximilian Hopf, Shammur A. Chowdhury, Soon gyo Jung, Hind Almerekhi, and Bernard J. Jansen. 2020 · 2020
Closest in time.
Directions in Abusive Language Training Data: Garbage In, Garbage Out
Bertie Vidgen and Leon Derczynski. 2020 · 2020
Closest in time.
Marcos Zampieri, Preslav Nakov, Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Hamdy Mubarak, Leon Derczynski, Zeses Pitenis, and Çağrı Çöltekin. 2020 · 2020
Closest in time.
Introducing CAD: the contextual abuse dataset
Bertie Vidgen, Dong Nguyen, Helen Margetts, Patricia Rossini, and Rebekah Tromble. 2021 · 2021
Closest in time.