Fetching the paper…
Reading the bibliography…
Toxicity annotators and content moderators often default to mental shortcuts when making decisions.
Understanding oppression: Strategies in addressing power and privilege
Leticia Nieto and Margot Boyer. 2006 · 2006
Earlier work this paper cites.
Statistical power analyses using G*Power 3.1: Tests for correlation and regression analyses
Franz Faul, Edgar Erdfelder, Axel Buchner, and Albert-Georg Lang. 2009 · 2009
Earlier work this paper cites.
Thinking, fast and slow
Daniel Kahneman. 2011 · 2011
Earlier work this paper cites.
Measuring the Reliability of Hate Speech Annotations: The Case of the European Refugee Crisis
Björn Ross, Michael Rist, Guillermo Carbonell, Benjamin Cabrera, Nils Kurowsky, and Michael Wojatzki. 2016 · 2016
Earlier work this paper cites.
Discrimination in america: experiences and views
RWJF. 2017 · 2017
Earlier work this paper cites.
e-snli: Natural language inference with natural language explanations
Oana-Maria Camburu, Tim Rocktäschel, Thomas Lukasiewicz, and Phil Blunsom. 2018 · 2018
Earlier work this paper cites.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller. 2019 · 2019
Earlier work this paper cites.
Behind the screen
Sarah T Roberts. 2019 · 2019
Earlier work this paper cites.
The risk of racial bias in hate speech detection
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A Smith. 2019 · 2019
Earlier work this paper cites.
Feature-based explanations don’t help people detect misclassifications of online toxicity
Samuel Carton, Qiaozhu Mei, and Paul Resnick. 2020 · 2020
Earlier work this paper cites.
Expanding the debate about content moderation: Scholarly research agendas for the coming policy debates
Tarleton Gillespie, Patricia Aufderheide, Elinor Carmi, Ysabel Gerrard, Robert Gorwa, Ariadna Matamoros-Fernandez, Sarah T Roberts, Aram Sinnreich, and Sarah Myers West. 2020 · 2020
Earlier work this paper cites.
Fortifying toxic speech detectors against veiled toxicity
Xiaochuang Han and Yulia Tsvetkov. 2020 · 2020
Cited alongside, same era.
"Why is ’Chicago’ Deceptive?" towards building model-driven tutorials for humans
Vivian Lai, Han Liu, and Chenhao Tan. 2020 · 2020
Cited alongside, same era.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A Smith, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Microaggressions: Clarification, evidence, and impact
Monnica T. Williams. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-Art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Cited alongside, same era.
Cascading biases: Investigating the effect of heuristic annotation strategies on data and models
Chaitanya Malaviya, Sudeep Bhatia, and Mark Yatskar. 2022 · 2022
Later among the works it cites.
Listening to Affected Communities to Define Extreme Speech: Dataset and Experiments
Antonis Maronikolakis, Axel Wisiorek, Leah Nann, Haris Jabbar, Sahana Udupa, and Hinrich Schuetze. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
Two contrasting data annotation paradigms for subjective NLP tasks
Paul Rottger, Bertie Vidgen, Dirk Hovy, and Janet Pierrehumbert. 2022 · 2022
Later among the works it cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Does the whole exceed its parts? the effect of ai explanations on complementary team performance
Gagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok, Besmira Nushi, Ece Kamar, Marco Tulio Ribeiro, and Daniel Weld. 2021 · 2021
Cited alongside, same era.
Double Standards in Social Media Content Moderation
Ángel Díaz and Laura Hecht-Felella. 2021 · 2021
Cited alongside, same era.
DEBERTA: DECODING-ENHANCED BERT WITH DISENTANGLED ATTENTION
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
Misinfo reaction frames: Reasoning about readers’ reactions to news headlines
Saadia Gabriel, Skyler Hallinan, Maarten Sap, Pemi Nguyen, Franziska Roesner, Eunsol Choi, and Yejin Choi. 2022 · 2022
Cited alongside, same era.
Human-AI collaboration via conditional delegation: A case study of content moderation
Vivian Lai, Samuel Carton, Rajat Bhatnagar, Q. Vera Liao, Yunfeng Zhang, and Chenhao Tan. 2022 · 2022
Cited alongside, same era.
The principles of psychology , volume 1
William James, Frederick Burkhardt, Fredson Bowers, and Ignas K Skrupskelis. 1890
Cited in the paper.
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Later among the works it cites.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Cobra frames: Contextual reasoning about effects and harms of offensive statements
Xuhui Zhou, Hao Zhu, Akhila Yerukola, Thomas Davidson, Jena D. Hwang, Swabha Swayamdipta, and Maarten Sap. 2023 · 2022
Later among the works it cites.
Selective Explanations: Leveraging Human Input to Align Explainable AI
Vivian Lai, Yiming Zhang, Chacha Chen, Q. Vera Liao, and Chenhao Tan. 2023 · 2023
Closest in time.
Explanations Can Reduce Overreliance on AI Systems During Decision-Making
Helena Vasconcelos, Matthew Jörke, Madeleine Grunde-McLaughlin, Tobias Gerstenberg, Michael Bernstein, and Ranjay Krishna. 2023 · 2023
Closest in time.
Self-Consistency Improves Chain of Thought Reasoning in Language Models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. 2023 · 2023
Closest in time.