Fetching the paper…
Reading the bibliography…
Bias research in NLP seeks to analyse models for social biases, thus helping NLP practitioners uncover, measure, and mitigate social harms.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Abigail Z. Jacobs and Hanna Wallach. 2021 · 1912
Earlier work this paper cites.
Yi Zhou, Masahiro Kaneko, and Danushka Bollegala. 2022 · 1935
Earlier work this paper cites.
Intrinsic bias metrics do not correlate with application bias
Seraphina Goldfarb-Tarrant, Rebecca Marchant, Ricardo Muñoz Sánchez, Mugdha Pandya, and Adam Lopez. 2021 · 1940
Earlier work this paper cites.
Factors relevant to the validity of experiments in social settings
Donald T. Campbell. 1957 · 1957
Earlier work this paper cites.
Crows-pairs: A challenge dataset for measuring social biases in masked language models
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel R. Bowman. 2020 · 1967
Earlier work this paper cites.
Harms of gender exclusivity and challenges in non-binary representation in language technologies
Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian, Jeff Phillips, and Kai-Wei Chang. 2021 · 1994
Earlier work this paper cites.
Confirmation bias: A ubiquitous phenomenon in many guises
Raymond S Nickerson. 1998 · 1998
Earlier work this paper cites.
Measurement validity: A shared standard for qualitative and quantitative research
Robert Adcock and David Collier. 2001 · 2001
Earlier work this paper cites.
Experimental research
Susan Gass. 2010 · 2010
Earlier work this paper cites.
How to analyze political attention with minimal assumptions and costs
Kevin M Quinn, Burt L Monroe, Michael Colaresi, Michael H Crespin, and Dragomir R Radev. 2010 · 2010
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y. Zou, Venkatesh Saligrama, and Adam Tauman Kalai. 2016 · 2016
Earlier work this paper cites.
Detecting independent pronoun bias with partially-synthetic data generation
Robert Munro and Alex (Carmen) Morrison. 2020 · 2017
Earlier work this paper cites.
Examining gender and race bias in two hundred sentiment analysis systems
Svetlana Kiritchenko and Saif Mohammad. 2018 · 2018
Earlier work this paper cites.
Identifying and reducing gender bias in word-level language models
Shikha Bordia and Samuel R. Bowman. 2019 · 2019
Earlier work this paper cites.
Measuring bias in contextualized word representations
Keita Kurita, Nidhi Vyas, Ayush Pareek, Alan W Black, and Yulia Tsvetkov. 2019 · 2019
Earlier work this paper cites.
The woman worked as a babysitter: On biases in language generation
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2019 · 2019
Cited alongside, same era.
A defense and definition of construct validity in psychology
Caroline Stone. 2019 · 2019
Cited alongside, same era.
Multi-dimensional gender bias classification
Emily Dinan, Angela Fan, Ledell Wu, Jason Weston, Douwe Kiela, and Adina Williams. 2020 · 2020
Cited alongside, same era.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A. Smith. 2020 · 2020
Cited alongside, same era.
Investigating African-American Vernacular English in transformer-based text generation
Sophie Groenwold, Lily Ou, Aesha Parekh, Samhita Honnavalli, Sharon Levy, Diba Mirza, and William Yang Wang. 2020 · 2020
Cited alongside, same era.
Towards a critical race methodology in algorithmic fairness
Unpacking the interdependent systems of discrimination: Ableist bias in NLP systems through an intersectional lens
Saad Hassan, Matt Huenerfauth, and Cecilia Ovesdotter Alm. 2021 · 2021
Later among the works it cites.
Ecological validity and “ecological validity”
John F. Kihlstrom. 2021 · 2021
Later among the works it cites.
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2021 · 2021
Later among the works it cites.
Ensemble of MRR and NDCG models for visual dialog
Idan Schwartz. 2021 · 2021
Later among the works it cites.
Hi, my name is martha: Using names to measure and mitigate bias in generative dialogue models
Eric Michael Smith and Adina Williams. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alex Hanna, Emily Denton, Andrew Smart, and Jamila Smith-Loud. 2020 · 2020
Cited alongside, same era.
Reducing sentiment bias in language models via counterfactual evaluation
Po-Sen Huang, Huan Zhang, Ray Jiang, Robert Stanforth, Johannes Welbl, Jack Rae, Vishal Maini, Dani Yogatama, and Pushmeet Kohli. 2020 · 2020
Cited alongside, same era.
UNQOVERing stereotyping biases via underspecified questions
Tao Li, Daniel Khashabi, Tushar Khot, Ashish Sabharwal, and Vivek Srikumar. 2020a · 2020
Cited alongside, same era.
Unqovering stereotyping biases via underspecified questions
Tao Li, Daniel Khashabi, Tushar Khot, Ashish Sabharwal, and Vivek Srikumar. 2020b · 2020
Cited alongside, same era.
Investigating societal biases in a poetry composition system
Emily Sheng and David Uthus. 2020 · 2020
Cited alongside, same era.
Mitigating language-dependent ethnic bias in BERT
Jaimeen Ahn and Alice Oh. 2021 · 2021
Cited alongside, same era.
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Abeba Birhane, Vinay Uday Prabhu, and Emmanuel Kahembwe. 2021 · 2021
Cited alongside, same era.
2022 · 2022
Later among the works it cites.
Challenges in measuring bias via open-ended language generation
Afra Feyza Akyürek, Muhammed Yusuf Kocyigit, Sejin Paik, and Derry Tanti Wijaya. 2022 · 2022
Later among the works it cites.
Inferring gender: A scalable methodology for gender detection with online lexical databases
Marion Bartl and Susan Leavy. 2022 · 2022
Later among the works it cites.
Re-contextualizing fairness in NLP: The case of India
Shaily Bhatt, Sunipa Dev, Partha Talukdar, Shachi Dave, and Vinodkumar Prabhakaran. 2022 · 2022
Later among the works it cites.
On the intrinsic and extrinsic fairness evaluation metrics for contextualized language representations
Yang Cao, Yada Pruksachatkun, Kai-Wei Chang, Rahul Gupta, Varun Kumar, Jwala Dhamala, and Aram Galstyan. 2022 · 2022
Later among the works it cites.
Measuring fairness with biased rulers: A comparative study on bias metrics for pre-trained language models
Pieter Delobelle, Ewoenam Tokpo, Toon Calders, and Bettina Berendt. 2022 · 2022
Later among the works it cites.
SafetyKit: First aid for measuring safety in open-domain conversational systems
Emily Dinan, Gavin Abercrombie, A. Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser. 2022 · 2022
Later among the works it cites.
Pangubot: Efficient generative dialogue pre-training from pre-trained language model
Fei Mi, Yitong Li, Yulong Zeng, Jingyan Zhou, Yasheng Wang, Chuanfei Xu, Lifeng Shang, Xin Jiang, Shiqi Zhao, and Qun Liu. 2022 · 2022
Later among the works it cites.
Quantifying social biases using templates is unreliable
Preethi Seshadri, Pouya Pezeshkpour, and Sameer Singh. 2022 · 2022
Later among the works it cites.
Upstream Mitigation Is Not
Ryan Steed, Swetasudha Panda, Ari Kobren, and Michael Wick. 2022 · 2022
Later among the works it cites.