Fetching the paper…
Reading the bibliography…
Warning: this paper contains content that may be offensive or upsetting.
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Note on the sampling error of the difference between correlated proportions or percentages
Quinn McNemar. 1947 · 1947
Earlier work this paper cites.
Tests for departure from normality. Empirical results for the distributions of b2 and √b1
Ralph D’Agostino and E. S. Pearson. 1973 · 1973
Earlier work this paper cites.
Culture’s Consequences: International Differences in Work-Related Values
G. Hofstede. 1984 · 1984
Earlier work this paper cites.
The effect of national culture on the choice of entry mode
Bruce Kogut and Harbir Singh. 1988 · 1988
Earlier work this paper cites.
Education and colonial transition in singapore and hong kong: Comparisons and contrasts
Jason Tan. 1997 · 1997
Earlier work this paper cites.
Keep calm and carry on: Reflections on the anglosphere
Andrew Davies, Graeme Dobell, Peter Jennings, Sarah Norgrove, Andrew Smith, Nic Stuart, and Hugh White. 2013 · 2013
Earlier work this paper cites.
Language and culture
Claire Kramsch. 2014 · 2014
Earlier work this paper cites.
Lingua franca english of south africa
Irina Khokhlova. 2015 · 2015
Earlier work this paper cites.
Measuring the Reliability of Hate Speech Annotations: The Case of the European Refugee Crisis
Björn Ross, Michael Rist, Guillermo Carbonell, Benjamin Cabrera, Nils Kurowsky, and Michael Wojatzki. 2016 · 2016
Earlier work this paper cites.
Are you a racist or am I seeing things? Annotator influence on hate speech detection on Twitter
Zeerak Waseem. 2016 · 2016
Earlier work this paper cites.
Hateful symbols or hateful people? Predictive features for hate speech detection on Twitter
Zeerak Waseem and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Like trainer, like bot? Inheritance of bias in algorithmic content moderation
Reuben Binns, Michael Veale, Max Van Kleek, and Nigel Shadbolt. 2017 · 2017
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Hate speech dataset from a white supremacy forum
Ona de Gibert, Naiara Perez, Aitor García-Pablos, and Montse Cuadros. 2018 · 2018
Earlier work this paper cites.
Large scale crowdsourcing and characterization of Twitter abusive behavior
Antigoni Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Earlier work this paper cites.
Crowdsourcing subjective tasks: The case study of understanding toxicity in online discussions
Lora Aroyo, Lucas Dixon, Nithum Thain, Olivia Redfield, and Rachel Rosen. 2019 · 2019
Earlier work this paper cites.
Finding microaggressions in the wild: A case for locating elusive phenomena in social media posts
Luke Breitfeller, Emily Ahn, David Jurgens, and Yulia Tsvetkov. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
Multilingual and multi-aspect hate speech analysis
Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang, Yangqiu Song, and Dit-Yan Yeung. 2019 · 2019
Cited alongside, same era.
Predicting the type and target of offensive posts in social media
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019 · 2019
Cited alongside, same era.
TweetEval: Unified benchmark and comparative evaluation for tweet classification
Francesco Barbieri, Jose Camacho-Collados, Luis Espinosa Anke, and Leonardo Neves. 2020 · 2020
Cited alongside, same era.
That “Special Something”: The U.S.-Australia Alliance, Special Relationships, and Emotions
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Y. Zhao, Yanping Huang, Andrew M. Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2022 · 2022
Later among the works it cites.
Detox: A comprehensive dataset for German offensive language and conversation analysis
Christoph Demus, Jonas Pitz, Mina Schütz, Nadine Probol, Melanie Siegel, and Dirk Labudde. 2022 · 2022
Later among the works it cites.
COLD: A benchmark for Chinese offensive language detection
Jiawen Deng, Jingyan Zhou, Hao Sun, Chujie Zheng, Fei Mi, Helen Meng, and Minlie Huang. 2022 · 2022
Later among the works it cites.
Is your toxicity my toxicity? Exploring the impact of rater identity on toxicity annotation
Nitesh Goyal, Ian D. Kivlichan, Rachel Rosen, and Lucy Vasserman. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lloyd Cox and Brendon O’Connor. 2020 · 2020
Cited alongside, same era.
BERTweet: A pre-trained language model for English tweets
Dat Quoc Nguyen, Thanh Vu, and Anh Tuan Nguyen. 2020 · 2020
Cited alongside, same era.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A. Smith, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Cited alongside, same era.
HateBERT: Retraining BERT for abusive language detection in English
Tommaso Caselli, Valerio Basile, Jelena Mitrović, and Michael Granitzer. 2021 · 2021
Cited alongside, same era.
Latent hatred: A benchmark for understanding implicit hate speech
Mai ElSherief, Caleb Ziems, David Muchlinski, Vaishnavi Anupindi, Jordyn Seybolt, Munmun De Choudhury, and Diyi Yang. 2021 · 2021
Cited alongside, same era.
4: The Anglo–American World View1
Andrew Gamble. 2021 · 2021
Cited alongside, same era.
Srinivasan Iyer, Xi Victoria Lin, Ramakanth Pasunuru, Todor Mihaylov, Daniel Simig, Ping Yu, Kurt Shuster, Tianlu Wang, Qing Liu, Punit Singh Koura, Xian Li, Brian O’Horo, Gabriel Pereyra, Jeff Wang, Christopher Dewan, Asli Celikyilmaz, Luke Zettlemoyer, and Ves Stoyanov. 2022 · 2022
Later among the works it cites.
KOLD: Korean offensive language dataset
Younghoon Jeong, Juhyun Oh, Jongwon Lee, Jaimeen Ahn, Jihyung Moon, Sungjoon Park, and Alice Oh. 2022 · 2022
Later among the works it cites.
Dealing with disagreements: Looking beyond the majority vote in subjective annotations
Aida Mostafazadeh Davani, Mark Díaz, and Vinodkumar Prabhakaran. 2022 · 2022
Later among the works it cites.
Emojis as anchors to detect arabic offensive language and hate speech
Hamdy Mubarak, Sabit Hassan, and Shammur Absar Chowdhury. 2022 · 2022
Later among the works it cites.
Assessing annotator identity sensitivity via item response theory: A case study in a hate speech corpus
Pratik S. Sachdeva, Renata Barreto, Claudia von Vacano, and Chris J. Kennedy. 2022 · 2022
Later among the works it cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Later among the works it cites.
EPIC: Multi-perspective annotation of a corpus of irony
Simona Frenda, Alessandro Pedrani, Valerio Basile, Soda Marem Lo, Alessandra Teresa Cignarella, Raffaella Panizzon, Cristina Marco, Bianca Scarlini, Viviana Patti, Cristina Bosco, and Davide Bernardi. 2023 · 2023
Closest in time.
KoBBQ: Korean bias benchmark for question answering
Jiho Jin, Jiseon Kim, Nayeon Lee, Haneul Yoo, Alice Oh, and Hwaran Lee. 2023 · 2023
Closest in time.
Hate speech classifiers are culturally insensitive
Nayeon Lee, Chani Jung, and Alice Oh. 2023 · 2023
Closest in time.
Orca 2: Teaching small language models how to reason
Arindam Mitra, Luciano Del Corro, Shweti Mahajan, Andrés Codas, Clarisse Simões, Sahaj Agrawal, Xuxi Chen, Anastasia Razdaibiedina, Erik Jones, Kriti Aggarwal, Hamid Palangi, Guoqing Zheng, Corby Rosset, Hamed Khanpour, and Ahmed Awadallah. 2023 · 2023
Closest in time.
When do annotator demographics matter? measuring the influence of annotator demographics with the POPQUORN dataset
Jiaxin Pei and David Jurgens. 2023 · 2023
Closest in time.
Why don’t you do it right? Analysing annotators’ disagreement in subjective tasks
Marta Sandri, Elisa Leonardelli, Sara Tonelli, and Elisabetta Jezek. 2023 · 2023
Closest in time.
TwHIN-BERT: A socially-enriched pre-trained language model for multilingual tweet representations at Twitter
Xinyang Zhang, Yury Malkov, Omar Florez, Serim Park, Brian McWilliams, Jiawei Han, and Ahmed El-Kishky. 2023 · 2023
Closest in time.