Fetching the paper…
Reading the bibliography…
The use of words to convey speaker's intent is traditionally distinguished from the `mention' of words for quoting what someone said, or pointing out properties of a word.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
The concept of truth in formalized languages
Alfred Tarski. 1931 · 1931
Earlier work this paper cites.
Use versus mention
Willard VO Quine. 1940 · 1940
Earlier work this paper cites.
Irony and the use-mention distinction
Dan Sperber and Deirdre Wilson. 1981 · 1981
Earlier work this paper cites.
Conversational adequacy: mistakes are the essence
Donald Perlis, Khemdut Purang, and Carl Andersen. 1998 · 1998
Earlier work this paper cites.
Quotation and the use-mention distinction
Paul Saka. 1998 · 1998
Earlier work this paper cites.
The use-mention distinction and its importance to HCI
Michael L Anderson, Yoshi Okamoto, Darsana Josyula, and Don Perlis. 2002 · 2002
Earlier work this paper cites.
Introduction: Counter-narratives and the power to oppose
Molly Andrews. 2002 · 2002
Earlier work this paper cites.
Fightin’words: Lexical feature selection and evaluation for identifying the content of political conflict
Burt L Monroe, Michael P Colaresi, and Kevin M Quinn. 2008 · 2008
Earlier work this paper cites.
Factbank: a corpus annotated with event factuality
Roser Saurí and James Pustejovsky. 2009 · 2009
Earlier work this paper cites.
Automatic committed belief tagging
Vinodkumar Prabhakaran, Owen Rambow, and Mona Diab. 2010 · 2010
Earlier work this paper cites.
Distinguishing use and mention in natural language
Shomir Wilson. 2010 · 2010
Earlier work this paper cites.
The N word: Its history and use in the African American community
Jacquelyn Rahman. 2012 · 2012
Earlier work this paper cites.
Are you sure that this happened? assessing the factuality degree of events in text
Roser Saurí and James Pustejovsky. 2012 · 2012
Earlier work this paper cites.
The creation of a corpus of english metalanguage
Shomir Wilson. 2012 · 2012
Earlier work this paper cites.
An In-depth Analysis of Implicit and Subtle Hate Speech Messages
Nicolas Ocampo, Ekaterina Sviridova, Elena Cabrio, and Serena Villata. 2023 · 2013
Earlier work this paper cites.
Toward automatic processing of english metalanguage
Shomir Wilson. 2013 · 2013
Earlier work this paper cites.
Countering dangerous speech: New ideas for genocide prevention
Susan Benesch. 2014 · 2014
Earlier work this paper cites.
A new dataset and evaluation for belief/factuality
Vinodkumar Prabhakaran, Julia Hirschberg, Owen Rambow, Samira Shaikh, Tomek Strzalkowski, Jennifer Tracey, Michael Arrigo, Rupayan Basu, Micah Clark, Adam Dalton, et al. 2015 · 2015
Earlier work this paper cites.
Governing hate speech by means of counterspeech on facebook
Carla Schieb and Mike Preuss. 2016 · 2016
Earlier work this paper cites.
Detection of online fake news using n-gram analysis and machine learning techniques
Hadeer Ahmed, Issa Traore, and Sherif Saad. 2017 · 2017
Earlier work this paper cites.
Vectors for Counterspeech on Twitter
Lucas Wright, Derek Ruths, Kelly P Dillon, Haji Mohammad Saleem, and Susan Benesch. 2017 · 2017
Earlier work this paper cites.
Measuring and mitigating unintended bias in text classification
Lucas Dixon, John Li, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman. 2018 · 2018
Earlier work this paper cites.
Reducing gender bias in abusive language detection
Ji Ho Park, Jamin Shin, and Pascale Fung. 2018 · 2018
Earlier work this paper cites.
Challenges for toxic comment classification: An in-depth error analysis
Betty Van Aken, Julian Risch, Ralf Krestel, and Alexander Löser. 2018 · 2018
Earlier work this paper cites.
Conan-counter narratives through nichesourcing: a multilingual dataset of responses to fight online hate speech
Yi-Ling Chung, Elizaveta Kuzmenko, Serra Sinem Tekiroğlu, and Marco Guerini. 2019 · 2019
Earlier work this paper cites.
The commitmentbank: Investigating projection in naturally occurring discourse
Marie-Catherine De Marneffe, Mandy Simons, and Judith Tonhauser. 2019 · 2019
Earlier work this paper cites.
Thou shalt not hate: Countering online hate speech
Binny Mathew, Punyajoy Saha, Hardik Tharad, Subham Rajgaria, Prajwal Singhania, Suman Kalyan Maity, Pawan Goyal, and Animesh Mukherjee. 2019 · 2019
Earlier work this paper cites.
A Benchmark Dataset for Learning to Intervene in Online Hate Speech
Jing Qian, Anna Bethke, Yinyin Liu, Elizabeth Belding, and William Yang Wang. 2019 · 2019
Earlier work this paper cites.
The risk of racial bias in hate speech detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A Smith. 2019 · 2019
Cited alongside, same era.
# MeToo, Time’s up, and Theories of Justice
Lesley Wexler, Jennifer K Robbennolt, and Colleen Murphy. 2019 · 2019
Cited alongside, same era.
Countering hate on social media: Large scale classification of hate and counter speech
Joshua Garland, Keyan Ghazi-Zahedi, Jean-Gabriel Young, Laurent Hébert-Dufresne, and Mirta Galesic. 2020 · 2020
Cited alongside, same era.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A Smith. 2020 · 2020
Cited alongside, same era.
Do You Really Want to Hurt Me? Predicting Abusive Swearing in Social Media
Endang Wahyu Pamungkas, Valerio Basile, and Viviana Patti. 2020 · 2020
Cited alongside, same era.
Social Bias Frames: Reasoning about Social and Power Implications of Language
Pile of law: Learning responsible data filtering from the law and a 256gb open-source legal dataset
Peter Henderson, Mark Krass, Lucia Zheng, Neel Guha, Christopher D Manning, Dan Jurafsky, and Daniel Ho. 2022 · 2022
Later among the works it cites.
Handling and Presenting Harmful Text in NLP Research
Hannah Kirk, Abeba Birhane, Bertie Vidgen, and Leon Derczynski. 2022 · 2022
Later among the works it cites.
Re-examining factbank: Predicting the author’s presentation of factuality
John Murzaku, Peter Zeng, Magdalena Markowska, and Owen Rambow. 2022 · 2022
Later among the works it cites.
Detecting Unintended Social Bias in Toxic Language Datasets
Nihar Sahoo, Himanshu Gupta, and Pushpak Bhattacharyya. 2022 · 2022
Later among the works it cites.
Defining and detecting toxicity on social media: context and knowledge are key
Amit Sheth, Valerie L Shalin, and Ugur Kursuncu. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A. Smith, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
#No2Sectarianism: Experimental Approaches to Reducing Sectarian Hate Speech Online
Alexandra Siegel and Vivienne Badaan. 2020 · 2020
Cited alongside, same era.
Towards knowledge-grounded counter narrative generation for hate speech
Yi-Ling Chung, Serra Sinem Tekiroğlu, and Marco Guerini. 2021 · 2021
Cited alongside, same era.
Fighting hate speech, silencing drag queens? artificial intelligence in content moderation and risks to LGBTQ voices online
Thiago Dias Oliva, Dennys Marcelo Antonialli, and Alessandra Gomes. 2021 · 2021
Cited alongside, same era.
Double standards in social media content moderation
Ángel Díaz and Laura Hecht-Felella. 2021 · 2021
Cited alongside, same era.
Human-in-the-Loop for Data Collection: a Multi-Target Counter Narrative Dataset to Fight Online Hate Speech
Margherita Fanton, Helena Bonaldi, Serra Sinem Tekiroğlu, and Marco Guerini. 2021 · 2021
Cited alongside, same era.
Detecting cross-geographic biases in toxicity modeling on social media
Sayan Ghosh, Dylan Baker, David Jurgens, and Vinodkumar Prabhakaran. 2021 · 2021
Cited alongside, same era.
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, and others. 2022 · 2022
Later among the works it cites.
Hate Speech and Counter Speech Detection: Conversational Context Does Matter
Xinchen Yu, Eduardo Blanco, and Lingzi Hong. 2022 · 2022
Later among the works it cites.
What makes a good counter-stereotype? evaluating strategies for automated responses to stereotypical text
Kathleen Fraser, Svetlana Kiritchenko, Isar Nejadgholi, and Anna Kerkhof. 2023 · 2023
Later among the works it cites.
Handling bias in toxic speech detection: A survey
Tanmay Garg, Sarah Masud, Tharun Suresh, and Tanmoy Chakraborty. 2023 · 2023
Later among the works it cites.
Beyond “Mention vs. Use”: The Linguistics of Slurs
Caitlin Green. 2023 · 2023
Later among the works it cites.
Reinforcement learning-based counter-misinformation response generation: a case study of covid-19 vaccine misinformation
Bing He, Mustaque Ahamad, and Srijan Kumar. 2023 · 2023
Later among the works it cites.
A fine-grained comparison of pragmatic language understanding in humans and language models
Jennifer Hu, Sammy Floyd, Olessia Jouravlev, Evelina Fedorenko, and Edward Gibson. 2023 · 2023
Later among the works it cites.
From Dogwhistles to Bullhorns: Unveiling Coded Rhetoric with Language Models
Julia Mendelsohn, Ronan Le Bras, Yejin Choi, and Maarten Sap. 2023 · 2023
Later among the works it cites.
Hate speech: Publisher and Creator Guidelines
Meta. 2023 · 2023
Later among the works it cites.
Beyond denouncing hate: Strategies for countering implied biases and stereotypes in language
Jimin Mun, Emily Allaway, Akhila Yerukola, Laura Vianna, Sarah-Jane Leslie, and Maarten Sap. 2023 · 2023
Later among the works it cites.
Toward Disambiguating the Definitions of Abusive, Offensive, Toxic, and Uncivil Comments
Pia Pachinger, Allan Hanbury, Julia Neidhardt, and Anna Planitzer. 2023 · 2023
Later among the works it cites.
Using machine learning to reduce toxicity online
Perspective. 2023 · 2023
Later among the works it cites.
Distinguishing address vs. reference mentions of personal names in text
Vinodkumar Prabhakaran, Aida Mostafazadeh Davani, Melissa Ferguson, and Stav Atir. 2023 · 2023
Later among the works it cites.
Evaluating neural word embeddings for Sanskrit
Jivnesh Sandhan, Om Adideva Paranjay, Komal Digumarthi, Laxmidhar Behra, and Pawan Goyal. 2023 · 2023
Later among the works it cites.
NLPositionality: Characterizing Design Biases of Datasets and Models
Sebastin Santy, Jenny T Liang, Ronan Le Bras, Katharina Reinecke, and Maarten Sap. 2023 · 2023
Later among the works it cites.
Grounding or Guesswork? Large Language Models are Presumptive Grounders
Omar Shaikh, Kristina Gligorić, Ashna Khetan, Matthias Gerstgrasser, Diyi Yang, and Dan Jurafsky. 2023 · 2023
Later among the works it cites.
Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them
Mirac Suzgun, Nathan Scales, Nathanael Schärli, Sebastian Gehrmann, Yi Tay, Hyung Won Chung, Aakanksha Chowdhery, Quoc Le, Ed Chi, Denny Zhou, and Jason Wei. 2023a · 2023
Later among the works it cites.
Safety and Civility
TikTok. 2023 · 2023
Later among the works it cites.
Does GPT-3 Grasp Metaphors? Identifying Metaphor Mappings with Generative Language Models
Lennart Wachowiak and Dagmar Gromann. 2023 · 2023
Later among the works it cites.
Current safeguards, risk mitigation, and transparency measures of large language models against the generation of health disinformation: repeated cross sectional analysis
Bradley D Menz, Nicole M Kuderer, Stephen Bacchi, Natansh D Modi, Benjamin Chin-Yee, Tiancheng Hu, Ceara Rickard, Mark Haseloff, Agnes Vitry, Ross A McKinnon, et al. 2024 · 2024
Closest in time.
Misinformation and harmful language are interconnected, rather than distinct, challenges
Mohsen Mosleh, Rocky Cole, and David G Rand. 2024 · 2024
Closest in time.
Elqa: A corpus of metalinguistic questions and answers about english
Shabnam Behzad, Keisuke Sakaguchi, Nathan Schneider, and Amir Zeldes. 2023 · 2047
Closest in time.
Don’t Let Me Be Misunderstood:Comparing Intentions and Perceptions in Online Discussions
Jonathan P. Chang, Justin Cheng, and Cristian Danescu-Niculescu-Mizil. 2020 · 2077
Closest in time.