Fetching the paper…
Reading the bibliography…
Text data can pose a risk of harm.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
The relationship among cognitive schemas, job-related traumatic exposure, and posttraumatic stress disorder in journalists
Caroline M Pyevich, Elana Newman, and Eric Daleiden. 2003 · 2003
Earlier work this paper cites.
Language (Technology) is Power: A Critical Survey of "Bias" in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé, and Hanna Wallach. 2020 · 2005
Earlier work this paper cites.
A phenomenology of whiteness
Sara Ahmed. 2007 · 2007
Earlier work this paper cites.
Against scale: Provocations and resistances to scale thinking
Alex Hanna and Tina M Park. 2020 · 2010
Earlier work this paper cites.
A Guide to Effective Fact Checking On-air and Online
The Annenberg Public Policy Center. 2012 · 2012
Earlier work this paper cites.
The harm in hate speech
Jeremy Waldron. 2012 · 2012
Earlier work this paper cites.
Google search: Hyper-visibility as a means of rendering black women and girls invisible
Safiya Umoja Noble. 2013 · 2013
Earlier work this paper cites.
Fact checking: A studio workshop
Politifact. 2014 · 2014
Earlier work this paper cites.
Pheme: Computing veracity—the fourth challenge of big social data
Leon Derczynski, Kalina Bontcheva, Michal Lukasik, Thierry Declerck, Arno Scharl, Georgi Georgiev, Petya Osenova, Toms Pariente Lobo, Anna Kolliakou, Robert Stewart, et al. 2015 · 2015
Earlier work this paper cites.
Making secondary trauma a primary issue: A study of eyewitness media and vicarious trauma on the digital frontline
Sam Dubberley, Elizabeth Griffin, and Haluk Mert Bal. 2015 · 2015
Earlier work this paper cites.
Evidencing the Harms of Hate Speech
Katharine Gelber and Luke McNamara. 2016 · 2016
Earlier work this paper cites.
Best Practices for Conducting Risky Research and Protecting Yourself from Online Harassment (Data & Society Guide)
A. Marwick, L. Blackwell, and K. Lo. 2016 · 2016
Earlier work this paper cites.
Arkaitz Zubiaga, Maria Liakata, Rob Procter, Geraldine Wong Sak Hoi, and Peter Tolmie. 2016 · 2016
Earlier work this paper cites.
Limited Vision: The Undersampled Majority
Joy Buolamwini. 2017 · 2017
Earlier work this paper cites.
Ethical by Design: Ethics Best Practices for Natural Language Processing
Jochen L. Leidner and Vassilis Plachouras. 2017 · 2017
Earlier work this paper cites.
Datafication & Discrimination
Koen Leurs and Tamara Shepherd. 2017 · 2017
Earlier work this paper cites.
Spreadable spectacle in digital culture: Civic expression, fake news, and the role of media literacies in “post-fact” society
Paul Mihailidis and Samantha Viotty. 2017 · 2017
Earlier work this paper cites.
Say the Right Thing Right: Ethics Issues in Natural Language Generation Systems
Charese Smiley, Frank Schilder, Vassilis Plachouras, and Jochen L. Leidner. 2017 · 2017
Earlier work this paper cites.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M Bender and Batya Friedman. 2018 · 2018
Earlier work this paper cites.
Online targeting of researchers/academics: Ethical obligations and best practices
Devon Greyson, Nicole Cooke, Amelia Gibson, and Heidi Julien. 2018 · 2018
Earlier work this paper cites.
Anger, fear, and games: The long event of# GamerGate
Torill Elvira Mortensen. 2018 · 2018
Earlier work this paper cites.
Prior exposure increases perceived accuracy of fake news
Gordon Pennycook, Tyrone D Cannon, and David G Rand. 2018 · 2018
Cited alongside, same era.
Race After Technology: Abolitionist Tools for the New Jim Code
Ruha Benjamin. 2019 · 2019
Cited alongside, same era.
Social group identity and perceptions of online hate
Matthew Costello, James Hawdon, Colin Bernatzky, and Kelly Mendes. 2019 · 2019
Cited alongside, same era.
Testing stylistic interventions to reduce emotional impact of content moderation workers
Sowmya Karunakaran and Rashmi Ramakrishan. 2019 · 2019
Cited alongside, same era.
Model cards for model reporting
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru. 2019 · 2019
Cited alongside, same era.
Language Models Are Unsupervised Multitask Learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Abeba Birhane, Vinay Uday Prabhu, and Emmanuel Kahembwe. 2021 · 2021
Later among the works it cites.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al. 2021 · 2021
Later among the works it cites.
Danger! Negative memories ahead: the effect of warnings on reactions to and recall of negative memories
Victoria ME Bridgland and Melanie KT Takarangi. 2021 · 2021
Later among the works it cites.
Documenting the English Colossal Clean Crawled Corpus
Jesse Dodge, Maarten Sap, Ana Marasovic, William Agnew, Gabriel Ilharco, Dirk Groeneveld, and Matt Gardner. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Challenges and frontiers in abusive content detection
Bertie Vidgen, Alex Harris, Dong Nguyen, Rebekah Tromble, Scott Hale, and Helen Margetts. 2019 · 2019
Cited alongside, same era.
A Unified Taxonomy of Harmful Content
Michele Banko, Brendon MacKeen, and Laurie Ray. 2020 · 2020
Cited alongside, same era.
SemEval-2020 task 11: Detection of propaganda techniques in news articles
Giovanni Da San Martino, Alberto Barrón-Cedeno, Henning Wachsmuth, Rostislav Petrov, and Preslav Nakov. 2020 · 2020
Cited alongside, same era.
Fast, accurate, and healthier: Interactive blurring helps moderators reduce exposure to harmful content
Anubrata Das, Brandon Dang, and Matthew Lease. 2020 · 2020
Cited alongside, same era.
Data feminism
Catherine D’ignazio and Lauren F Klein. 2020 · 2020
Cited alongside, same era.
RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A. Smith. 2020 · 2020
Cited alongside, same era.
Report on the State of Resources Provided to Support Scholars Against Harassment, Trolling, and Doxxing While Doing Public Media Work
Alex Ketchum. 2021 · 2021
Later among the works it cites.
Bias Out-of-the-Box: An Empirical Analysis of Intersectional Occupational Biases in Popular Generative Language Models
Hannah Rose Kirk, Filippo Volpin, Haider Iqbal, Elias Benussi, Frederic Dreyer, Aleksandar Shtedritski, Yuki Asano, et al. 2021 · 2021
Later among the works it cites.
Reduced, Reused and Recycled: The Life of a Dataset in Machine Learning Research
Bernard Koch, Emily Denton, Alex Hanna, and Jacob G Foster. 2021 · 2021
Later among the works it cites.
What’s in the Box? A Preliminary Analysis of Undesirable Content in the Common Crawl Corpus
Alexandra Sasha Luccioni and Joseph D Viviano. 2021 · 2021
Later among the works it cites.
Data and its (dis) contents: A survey of dataset development and use in machine learning research
Amandalynne Paullada, Inioluwa Deborah Raji, Emily M Bender, Emily Denton, and Alex Hanna. 2021 · 2021
Later among the works it cites.
‘Just What do You Think You’re Doing, Dave?’A Checklist for Responsible Data Use in NLP
Anna Rogers, Timothy Baldwin, and Kobi Leins. 2021 · 2021
Later among the works it cites.
HateCheck: Functional Tests for Hate Speech Detection Models
Paul Röttger, Bertie Vidgen, Dong Nguyen, Zeerak Waseem, Helen Margetts, and Janet Pierrehumbert. 2021 · 2021
Later among the works it cites.
“Everyone wants to do the model work, not the data work”: Data Cascades in High-Stakes AI
Nithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo. 2021 · 2021
Later among the works it cites.
The psychological well-being of content moderators: the emotional labor of commercial moderation and avenues for improving support
Miriah Steiger, Timir J Bharucha, Sukrit Venkatagiri, Martin J Riedl, and Matthew Lease. 2021 · 2021
Later among the works it cites.
The state of online harassment
Emily A Vogels. 2021 · 2021
Later among the works it cites.
Ethical and social risks of harm from Language Models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al. 2021 · 2021
Later among the works it cites.
DataCLUE: A Benchmark Suite for Data-centric NLP
Liang Xu, Jiacheng Liu, Xiang Pan, Xiaojing Lu, and Xiaofeng Hou. 2021 · 2021
Later among the works it cites.
Annotating online misogyny
Philine Zeinert, Nanna Inie, and Leon Derczynski. 2021 · 2021
Later among the works it cites.
Hannah Rose Kirk, Bertram Vidgen, Paul Röttger, Tristan Thrush, and Scott A. Hale. 2022 · 2022
Closest in time.
Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets
Julia Kreutzer, Isaac Caswell, Lisa Wang, Ahsan Wahab, Daan van Esch, Nasanbayar Ulzii-Orshikh, Allahsera Tapo, Nishant Subramani, Artem Sokolov, Claytone Sikasote, et al. 2022 · 2022
Closest in time.
Characteristics of Harmful Text: Towards Rigorous Benchmarking of Language Models
Maribeth Rauh, John Mellor, Jonathan Uesato, Po-Sen Huang, Johannes Welbl, Laura Weidinger, Sumanth Dathathri, Amelia Glaese, Geoffrey Irving, Iason Gabriel, William Isaac, and Lisa Anne Hendricks. 2022 · 2022
Closest in time.
Annotators with Attitudes: How Annotator Beliefs And Identities Bias Toxic Language Detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Closest in time.
How war videos on social media can trigger secondary trauma
Jochen Spangenberg. 2022 · 2022
Closest in time.