Fetching the paper…
Reading the bibliography…
Hate speech has become pervasive in today's digital age.
Interrater reliability: the kappa statistic
Mary L McHugh. 2012 · 2012
Earlier work this paper cites.
Hate speech detection with comment embeddings. In Proceedings of the 24th international conference on world wide web . 29–30
Nemanja Djuric, Jing Zhou, Robin Morris, Mihajlo Grbovic, Vladan Radosavljevic, and Narayan Bhamidipati. 2015 · 2015
Earlier work this paper cites.
Governing hate speech by means of counterspeech on Facebook. In 66th ica annual conference, at fukuoka, japan . 1–23
Carla Schieb and Mike Preuss. 2016 · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language. In Proceedings of the international AAAI conference on web and social media , Vol. 11. 512–515
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
A survey on hate speech detection using natural language processing. In Proceedings of the fifth international workshop on natural language processing for social media . 1–10
Anna Schmidt and Michael Wiegand. 2017 · 2017
Earlier work this paper cites.
A survey on automatic detection of hate speech in text
Paula Fortuna and Sérgio Nunes. 2018 · 2018
Earlier work this paper cites.
Thou shalt not hate: Countering online hate speech. In Proceedings of the international AAAI conference on web and social media , Vol. 13. 369–380
Binny Mathew, Punyajoy Saha, Hardik Tharad, Subham Rajgaria, Prajwal Singhania, Suman Kalyan Maity, Pawan Goyal, and Animesh Mukherjee. 2019 · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . 3982–3992
Nils Reimers and Iryna Gurevych. 2019 · 2019
Earlier work this paper cites.
Countering hate on social media: Large scale classification of hate and counter speech. In Proceedings of the Fourth Workshop on Online Abuse and Harms . 102–112
Joshua Garland, Keyan Ghazi-Zahedi, Jean-Gabriel Young, Laurent Hébert-Dufresne, and Mirta Galesic. 2020 · 2020
Earlier work this paper cites.
Reformulating Unsupervised Style Transfer as Paraphrase Generation. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) . 737–762
Kalpesh Krishna, John Wieting, and Mohit Iyyer. 2020 · 2020
Earlier work this paper cites.
Hate speech: A systematized review
María Antonia Paz, Julio Montero-Díaz, and Alicia Moreno-Delgado. 2020 · 2020
Earlier work this paper cites.
Hate speech on social media: Content moderation in context
Richard Ashby Wilson and Molly K Land. 2020 · 2020
Earlier work this paper cites.
Countering online hate speech: An nlp perspective
Mudit Chaudhary, Chandni Saxena, and Helen Meng. 2021 · 2021
Earlier work this paper cites.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Earlier work this paper cites.
Prefix-tuning: Optimizing continuous prompts for generation
Xiang Lisa Li and Percy Liang. 2021 · 2021
Cited alongside, same era.
Generate, Prune, Select: A Pipeline for Counterspeech Generation against Online Hate Speech. In Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 . 134–149
Wanzheng Zhu and Suma Bhat. 2021 · 2021
Cited alongside, same era.
GraphNLI: A Graph-based Natural Language Inference Model for Polarity Prediction in Online Debates. In Proceedings of the ACM Web Conference 2022 . 2729–2737
Vibhor Agarwal, Sagar Joglekar, Anthony P Young, and Nishanth Sastry. 2022 · 2022
Cited alongside, same era.
Towards automatic generation of messages countering online hate speech and microaggressions. In Proceedings of the Sixth Workshop on Online Abuse and Harms (WOAH) . 11–23
Mana Ashida and Mamoru Komachi. 2022 · 2022
Cited alongside, same era.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung. 2023 · 2023
Closest in time.
Perspective API
Google Jigsaw. 2023 · 2023
Closest in time.
Automatic Translation of Hate Speech to Non-hate Speech in Social Media Texts
Yevhen Kostiuk, Atnafu Lambebo Tonja, Grigori Sidorov, and Olga Kolesnikova. 2023 · 2023
Closest in time.
Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models
Yiheng Liu, Tianle Han, Siyuan Ma, Jiayue Zhang, Yuanyuan Yang, Jiaming Tian, Hao He, Antong Li, Mengshen He, Zhengliang Liu, et al · 2023
Closest in time.
The psychological impacts of content moderation on content moderators: A qualitative study
Ruth Spence, Antonia Bifulco, Paula Bradbury, Elena Martellozzo, and Jeffrey DeMarco. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bosheng Ding, Chengwei Qin, Linlin Liu, Lidong Bing, Shafiq Joty, and Boyang Li. 2022 · 2022
Cited alongside, same era.
Impact and dynamics of hate and counter speech online
Joshua Garland, Keyan Ghazi-Zahedi, Jean-Gabriel Young, Laurent Hébert-Dufresne, and Mirta Galesic. 2022 · 2022
Cited alongside, same era.
ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 3309–3326
Thomas Hartvigsen, Saadia Gabriel, Hamid Palangi, Maarten Sap, Dipankar Ray, and Ece Kamar. 2022 · 2022
Cited alongside, same era.
Paradetox: Detoxification with parallel data. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 6804–6818
Varvara Logacheva, Daryna Dementieva, Sergey Ustyantsev, Daniil Moskovskiy, David Dale, Irina Krotova, Nikita Semenov, and Alexander Panchenko. 2022 · 2022
Cited alongside, same era.
Proactively Reducing the Hate Intensity of Online Posts via Hate Speech Normalization. In Proceedings of the 28th ACM SIGKDD Conference on Knowledge Discovery and Data Mining . 3524–3534
Sarah Masud, Manjot Bedi, Mohammad Aflah Khan, Md Shad Akhtar, and Tanmoy Chakraborty. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
Using Pre-Trained Language Models for Producing Counter Narratives Against Hate Speech: a Comparative Study. In Findings of the Association for Computational Linguistics: ACL 2022 . 3099–3114
Serra Sinem Tekiroğlu, Helena Bonaldi, Margherita Fanton, and Marco Guerini. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Cited alongside, same era.
Xiaofei Sun, Xiaoya Li, Jiwei Li, Fei Wu, Shangwei Guo, Tianwei Zhang, and Guoyin Wang. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Closest in time.
Evaluating GPT-3 Generated Explanations for Hateful Content Moderation
Han Wang, Ming Shan Hee, Md Rabiul Awal, Kenny Tsu Wei Choo, and Roy Ka-Wei Lee. 2023 · 2023
Closest in time.
Annobert: Effectively representing multiple annotators’ label choices to improve hate speech detection. In Proceedings of the International AAAI Conference on Web and Social Media , Vol. 17. 902–913
Wenjie Yin, Vibhor Agarwal, Aiqi Jiang, Arkaitz Zubiaga, and Nishanth Sastry. 2023 · 2023
Closest in time.
Siren’s Song in the AI Ocean: A Survey on Hallucination in Large Language Models
Yue Zhang, Yafu Li, Leyang Cui, Deng Cai, Lemao Liu, Tingchen Fu, Xinting Huang, Enbo Zhao, Yu Zhang, Yulong Chen, et al · 2023
Closest in time.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Closest in time.
Can chatgpt reproduce human-generated labels? a study of social computing tasks
Yiming Zhu, Peixian Zhang, Ehsan-Ul Haq, Pan Hui, and Gareth Tyson. 2023 · 2023
Closest in time.