Fetching the paper…
Reading the bibliography…
Counterspeech, i.e., responses to counteract potential harms of hateful speech, has become an increasingly popular solution to address online hate speech without censorship.
Advances in experimental social psychology
Leonard Berkowitz. 1984 · 1984
Earlier work this paper cites.
Stereotyping, prejudice, and discrimination
Susan Fiske. 1998 · 1998
Earlier work this paper cites.
When positive stereotypes threaten intellectual performance: the psychological hazards of “model minority” status
S Cheryan and G V Bodenhausen. 2000 · 2000
Earlier work this paper cites.
Social cognition: Thinking categorically about others
C. Neil Macrae and Galen V. Bodenhausen. 2000 · 2000
Earlier work this paper cites.
Generics: Cognition and acquisition
Sarah-Jane Leslie. 2008 · 2008
Earlier work this paper cites.
Generic statements require little evidence for acceptance but have powerful implications
Andrei Cimpian, Amanda C. Brandone, and Susan A. Gelman. 2010 · 2010
Earlier work this paper cites.
Extended Contact Effects: Is Exposure to Positive Outgroup Exemplars Sufficient or Is Interaction With Ingroup Members Necessary?
Vasile Cernat. 2011 · 2011
Earlier work this paper cites.
Misinformation and its correction: Continued influence and successful debiasing
Stephan Lewandowsky, Ullrich K. H. Ecker, Colleen M. Seifert, Norbert Schwarz, and John Cook. 2012 · 2012
Earlier work this paper cites.
Cultural transmission of social essentialism
Marjorie Rhodes, Sarah-Jane Leslie, and Christina M. Tworek. 2012 · 2012
Earlier work this paper cites.
Carving up the social world with generics
Sarah-Jane Leslie. 2014 · 2014
Earlier work this paper cites.
Counter-stereotypical pictures as a strategy for overcoming spontaneous gender stereotypes
Eimear Finnegan, Jane Oakhill, and Alan Garnham. 2015 · 2015
Earlier work this paper cites.
Considerations for Successful Counterspeech
Susan Benesch, Derek Ruths, Haji Mohammad Saleem, Kelly P. Dillon, and Lucas Wright. 2016 · 2016
Earlier work this paper cites.
Humanizing Outgroups Through Multiple Categorization: The Roles of Individuation and Threat
Francesca Prati, Richard J. Crisp, Rose Meleady, and Monica Rubini. 2016 · 2016
Earlier work this paper cites.
Winning Arguments: Interaction Dynamics and Persuasion Strategies in Good-faith Online Discussions
Chenhao Tan, Vlad Niculae, Cristian Danescu-Niculescu-Mizil, and Lillian Lee. 2016 · 2016
Earlier work this paper cites.
Discussion quality diffuses in the digital public square
George Berry and Sean J. Taylor. 2017 · 2017
Earlier work this paper cites.
The Original Sin of Cognition: Fear, Prejudice, and Generalization
Sarah-Jane Leslie. 2017 · 2017
Earlier work this paper cites.
Framing effects: Choice of slogans used to advertise online experiments can boost recruitment and lead to sample biases
Tal August, Nigini Oliveira, Chenhao Tan, Noah Smith, and Katharina Reinecke. 2018 · 2018
Earlier work this paper cites.
Is Civility Contagious? Examining the Impact of Modeling in Online Political Discussions
Soo-Hye Han, LeAnn M. Brazeal, and Natalie Pennington. 2018 · 2018
Earlier work this paper cites.
Censored, suspended, shadowbanned: User interpretations of content moderation on social media platforms
Sarah Myers West. 2018 · 2018
Earlier work this paper cites.
How Stereotypes Become Shared Knowledge: An Integrative Review on the Role of Biased Language Use in Communication about Categorized Individuals
Camiel J. Beukeboom and Christian Burgers. 2019 · 2019
Cited alongside, same era.
CONAN - COunter NArratives through nichesourcing: a multilingual dataset of responses to fight online hate speech
Yi-Ling Chung, Elizaveta Kuzmenko, Serra Sinem Tekiroglu, and Marco Guerini. 2019 · 2019
Cited alongside, same era.
Online toxicity is as old as the web itself but the return to communities may help
Kalev Leetaru. 2019 · 2019
Cited alongside, same era.
Can ’More Speech’ Counter Ignorant Speech? Tackling the Stickiness of Verbal Ignorance
Maxime Charles Lepoutre. 2019 · 2019
Cited alongside, same era.
Thou shalt not hate: Countering online hate speech
Binny Mathew, Punyajoy Saha, Hardik Tharad, Subham Rajgaria, Prajwal Singhania, Suman Kalyan Maity, Pawan Goyal, and Animesh Mukherjee. 2019 · 2019
The global effectiveness of fact-checking: Evidence from simultaneous experiments in Argentina, Nigeria, South Africa, and the United Kingdom
Ethan Porter and Thomas J. Wood. 2021 · 2021
Later among the works it cites.
Generate, prune, select: A pipeline for counterspeech generation against online hate speech
Wanzheng Zhu and Suma Bhat. 2021 · 2021
Later among the works it cites.
Human-machine collaboration approaches to build a dialogue dataset for hate speech countering
Helena Bonaldi, Sara Dellantonio, Serra Sinem Tekiroğlu, and Marco Guerini. 2022 · 2022
Later among the works it cites.
Why They Do It: Counterspeech Theories of Change
Catherine Buerger. 2022 · 2022
Later among the works it cites.
Speaking of kinds: How correcting generic statements can shape children’s concepts
Emily Foster-Hanson, Sarah-Jane Leslie, and Marjorie Rhodes. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A benchmark dataset for learning to intervene in online hate speech
Jing Qian, Anna Bethke, Yinyin Liu, Elizabeth Belding, and William Yang Wang. 2019 · 2019
Cited alongside, same era.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
The risk of racial bias in hate speech detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A Smith. 2019 · 2019
Cited alongside, same era.
Countering hate on social media: Large scale classification of hate and counter speech
Joshua Garland, Keyan Ghazi-Zahedi, Jean-Gabriel Young, Laurent Hébert-Dufresne, and Mirta Galesic. 2020 · 2020
Cited alongside, same era.
Generics and typicality: A bounded rationality approach
Robert Rooij and Katrin Schulz. 2020 · 2020
Cited alongside, same era.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A Smith, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Structural thinking about social categories: Evidence from formal explanations, generics, and generalization
Nadya Vasilyeva and Tania Lombrozo. 2020 · 2020
Cited alongside, same era.
Paul Röttger, Bertie Vidgen, Dirk Hovy, and Janet B Pierrehumbert. 2022 · 2022
Later among the works it cites.
Countergedi: A controllable approach to generate polite, detoxified and emotional counterspeech
Punyajoy Saha, Kanishk Singh, Adarsh Kumar, Binny Mathew, and Animesh Mukherjee. 2022 · 2022
Later among the works it cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Later among the works it cites.
Belief traps: Tackling the inertia of harmful beliefs
Marten Scheffer, Denny Borsboom, Sander Nieuwenhuis, and Frances Westley. 2022 · 2022
Later among the works it cites.
What makes online communities ‘better’? measuring values, consensus, and conflict across thousands of subreddits
Galen Weld, Amy X. Zhang, and Tim Althoff. 2022 · 2022
Later among the works it cites.
Hate Speech and Counter Speech Detection: Conversational Context Does Matter
Xinchen Yu, Eduardo Blanco, and Lingzi Hong. 2022 · 2022
Later among the works it cites.
Towards countering essentialism through social bias reasoning
Emily Allaway, Nina Taneja, Sarah-Jane Leslie, and Maarten Sap. 2023 · 2023
Closest in time.
Co-writing with opinionated language models affects users’ views
Maurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson, and Mor Naaman. 2023 · 2023
Closest in time.
Collective moderation of hate, toxicity, and extremity in online discussions
Jana Lasser, Alina Herderich, Joshua Garland, Segun Taofeek Aroyehun, David Garcia, and Mirta Galesic. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Is hate speech detection the solution the world wants?
Sara Parker and Derek Ruths. 2023 · 2023
Closest in time.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto. 2023 · 2023
Closest in time.
Can large language models transform computational social science?
Caleb Ziems, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang. 2023 · 2023
Closest in time.