Fetching the paper…
Reading the bibliography…
Due to the subtleness, implicity, and different possible interpretations perceived by different people, detecting undesirable content from text is a nuanced difficulty.
Controlling other people: The impact of power on stereotyping
Susan T Fiske. 1993 · 1993
Earlier work this paper cites.
Learning from bullying traces in social media
Jun-Ming Xu, Kwang-Sung Jun, Xiaojin Zhu, and Amy Bellmore. 2012 · 2012
Earlier work this paper cites.
"why should I trust you?": Explaining the predictions of any classifier
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Hateful symbols or hateful people? predictive features for hate speech detection on twitter
Zeerak Waseem and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
Detecting hate speech in social media
Shervin Malmasi and Marcos Zampieri. 2017 · 2017
Earlier work this paper cites.
Ex machina: Personal attacks seen at scale
Ellery Wulczyn, Nithum Thain, and Lucas Dixon. 2017 · 2017
Earlier work this paper cites.
Generative and discriminative text classification with recurrent neural networks
Dani Yogatama, Chris Dyer, Wang Ling, and Phil Blunsom. 2017 · 2017
Earlier work this paper cites.
Large scale crowdsourcing and characterization of twitter abusive behavior
Antigoni Maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Earlier work this paper cites.
Convolutional neural networks for toxic comment classification
Spiros V. Georgakopoulos, Sotiris K. Tasoulis, Aristidis G. Vrahatis, and Vassilis P. Plagianakos. 2018 · 2018
Earlier work this paper cites.
Benchmarking aggression identification in social media
Ritesh Kumar, Atul Kr. Ojha, Shervin Malmasi, and Marcos Zampieri. 2018 · 2018
Earlier work this paper cites.
Anatomy of online hate: Developing a taxonomy and machine learning models for identifying and classifying hate in online news media
Joni O. Salminen, Hind Almerekhi, Milica Milenkovic, Soon-Gyo Jung, Jisun An, Haewoon Kwak, and Bernard Jim Jansen. 2018 · 2018
Earlier work this paper cites.
Finding microaggressions in the wild: A case for locating elusive phenomena in social media posts
Luke Breitfeller, Emily Ahn, David Jurgens, and Yulia Tsvetkov. 2019 · 2019
Earlier work this paper cites.
Multilingual and multi-aspect hate speech analysis
Nedjma Ousidhoum, Zizheng Lin, Hongming Zhang, Yangqiu Song, and Dit-Yan Yeung. 2019 · 2019
Cited alongside, same era.
The risk of racial bias in hate speech detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A Smith. 2019 · 2019
Cited alongside, same era.
Predicting the type and target of offensive posts in social media
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019a · 2019
Cited alongside, same era.
SemEval-2019 task 6: Identifying and categorizing offensive language in social media (OffensEval)
Marcos Zampieri, Shervin Malmasi, Preslav Nakov, Sara Rosenthal, Noura Farra, and Ritesh Kumar. 2019b · 2019
Cited alongside, same era.
Hate speech detection using transformer ensembles on the hasoc dataset
Pedro Alonso, Rajkumar Saini, and György Kovács. 2020 · 2020
Cited alongside, same era.
Hate speech detection in twitter using transformer methods
Few-shot bot: Prompt-based learning for dialogue systems
Andrea Madotto, Zhaojiang Lin, Genta Indra Winata, and Pascale Fung. 2021 · 2021
Later among the works it cites.
Hatexplain: A benchmark dataset for explainable hate speech detection
Binny Mathew, Punyajoy Saha, Seid Muhie Yimam, Chris Biemann, Pawan Goyal, and Animesh Mukherjee. 2021 · 2021
Later among the works it cites.
Cins: Comprehensive instruction for few-shot learning in task-oriented dialog systems
Fei Mi, Yitong Li, Yasheng Wang, Xin Jiang, and Qun Liu. 2021 · 2021
Later among the works it cites.
Few-shot instruction prompts for pretrained language models to detect social biases
Shrimai Prabhumoye, Rafal Kocielnik, Mohammad Shoeybi, Anima Anandkumar, and Bryan Catanzaro. 2021 · 2021
Later among the works it cites.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in nlp
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Raymond T Mutanga, Nalindren Naicker, and Oludayo. O. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Cited alongside, same era.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A Smith, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Cross-lingual retrieval for iterative self-supervised training
Chau Tran, Yuqing Tang, Xian Li, and Jiatao Gu. 2020 · 2020
Cited alongside, same era.
SimCSE: Simple contrastive learning of sentence embeddings
Tianyu Gao, Xingcheng Yao, and Danqi Chen. 2021 · 2021
Cited alongside, same era.
Anna Glazkova, Michael Kadantsev, and Maksim Glazkov. 2021 · 2021
Cited alongside, same era.
Delphi: Towards machine ethics and norms
Liwei Jiang, Jena D Hwang, Chandra Bhagavatula, Ronan Le Bras, Maxwell Forbes, Jon Borchardt, Jenny Liang, Oren Etzioni, Maarten Sap, and Yejin Choi. 2021 · 2021
Cited alongside, same era.
Timo Schick, Sahana Udupa, and Hinrich Schütze. 2021 · 2021
Later among the works it cites.
Open aspect target sentiment classification with natural language prompts
Ronald Seoh, Ian Birle, Mrinal Tak, Haw-Shiuan Chang, Brian Pinette, and Alfred Hough. 2021 · 2021
Later among the works it cites.
Zero-shot image-to-text generation for visual-semantic arithmetic
Yoad Tewel, Yoav Shalev, Idan Schwartz, and Lior Wolf. 2021 · 2021
Later among the works it cites.
Finetuned language models are zero-shot learners
Jason Wei, Maarten Bosma, Vincent Y Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M Dai, and Quoc V Le. 2021 · 2021
Later among the works it cites.
An empirical study of gpt-3 for few-shot knowledge-based vqa
Zhengyuan Yang, Zhe Gan, Jianfeng Wang, Xiaowei Hu, Yumao Lu, Zicheng Liu, and Lijuan Wang. 2021 · 2021
Later among the works it cites.
Exploring prompt-based few-shot learning for grounded dialog generation
Chujie Zheng and Minlie Huang. 2021 · 2021
Later among the works it cites.
Zero-shot commonsense question answering with cloze translation and consistency optimization
Zi-Yi Dou and Nanyun Peng. 2022 · 2022
Closest in time.
Red teaming language models with language models
Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving. 2022 · 2022
Closest in time.