Fetching the paper…
Reading the bibliography…
Recent studies show that Natural Language Processing (NLP) technologies propagate societal biases about demographic groups associated with attributes such as gender, race, and nationality.
Towards understanding gender bias in relation extraction
Andrew Gaut, Tony Sun, Shirlyn Tang, Yuxin Huang, Jing Qian, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth M. Belding, Kai-Wei Chang, and William Yang Wang. 2019 · 1911
Earlier work this paper cites.
The nature of prejudice
Gordon Willard Allport, Kenneth Clark, and Thomas Pettigrew. 1954 · 1954
Earlier work this paper cites.
CrowS-pairs: A challenge dataset for measuring social biases in masked language models
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel R. Bowman. 2020 · 1967
Earlier work this paper cites.
Judgment under uncertainty: Heuristics and biases: Biases in judgments reveal some heuristics of thinking under uncertainty
Amos Tversky and Daniel Kahneman. 1974 · 1974
Earlier work this paper cites.
Categorization of natural objects
Carolyn B Mervis, Eleanor Rosch, et al. 1981 · 1981
Earlier work this paper cites.
Stereotypes and prejudice: Their automatic and controlled components
Patricia G Devine. 1989 · 1989
Earlier work this paper cites.
Conceptualizing stigma
Bruce G. Link and Jo C. Phelan. 2001 · 2001
Earlier work this paper cites.
Social Dominance: An Intergroup Theory of Social Hierarchy and Oppression
J. Sidanius and F. Pratto. 2001 · 2001
Earlier work this paper cites.
Fair hate speech detection through evaluation of social group counterfactuals
Aida Mostafazadeh Davani, Ali Omrani, Brendan Kennedy, Mohammad Atari, Xiang Ren, and Morteza Dehghani. 2020 · 2010
Earlier work this paper cites.
Lexical and semantic relations , Cambridge Textbooks in Linguistics, page 108–132. Cambridge University Press
M. Lynne Murphy. 2010 · 2010
Earlier work this paper cites.
Intrinsic bias metrics do not correlate with application bias
Seraphina Goldfarb-Tarrant, Rebecca Marchant, Ricardo Muñoz Sanchez, Mugdha Pandya, and Adam Lopez. 2020 · 2012
Earlier work this paper cites.
Demographic dialectal variation in social media: A case study of African-American English
Su Lin Blodgett, Lisa Green, and Brendan O’Connor. 2016 · 2016
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
T Bolukbasi, K W Chang, J Zou, V Saligrama, and A Kalai. 2016 · 2016
Earlier work this paper cites.
Recent research on dehumanization
Nick Haslam and Michelle Stratemeyer. 2016 · 2016
Earlier work this paper cites.
The social impact of natural language processing
Dirk Hovy and Shannon L. Spruit. 2016 · 2016
Earlier work this paper cites.
Are you a racist or am I seeing things? annotator influence on hate speech detection on Twitter
Zeerak Waseem. 2016 · 2016
Earlier work this paper cites.
Hateful symbols or hateful people? predictive features for hate speech detection on Twitter
Zeerak Waseem and Dirk Hovy. 2016 · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Aylin Caliskan, Joanna J. Bryson, and Arvind Narayanan. 2017 · 2017
Earlier work this paper cites.
The trouble with bias
Kate Crawford. 2017 · 2017
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber. 2017 · 2017
Earlier work this paper cites.
A large labeled corpus for online harassment research
Jennifer Golbeck, Zahra Ashktorab, Rashad O. Banjo, Alexandra Berlinger, Siddharth Bhagwan, Cody Buntain, Paul Cheakalos, Alicia A. Geller, Quint Gergory, Rajesh Kumar Gnanasekaran, Raja Rajan Gunasekaran, Kelly M. Hoffman, Jenny Hottle, Vichita Jienjitlert, Shivika Khare, Ryan Lau, Marianna J. Martindale, Shalmali Naik, Heather L. Nixon, Piyush Ramachandran, Kristine M. Rogers, Lisa Rogers, Meghna Sardana Sarin, Gaurav Shahane, Jayanee Thanki, Priyanka Vengataraman, Zijian Wan, and Derek Michael Wu. 2017 · 2017
Earlier work this paper cites.
Social bias in elicited natural language inferences
Rachel Rudinger, Chandler May, and Benjamin Van Durme. 2017 · 2017
Earlier work this paper cites.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M. Bender and Batya Friedman. 2018 · 2018
Earlier work this paper cites.
Addressing age-related bias in sentiment analysis
Mark Díaz, Isaac Johnson, Amanda Lazar, Anne Marie Piper, and Darren Gergle. 2018 · 2018
Earlier work this paper cites.
Measuring and mitigating unintended bias in text classification
Lucas Dixon, John Li, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman. 2018 · 2018
Earlier work this paper cites.
Large scale crowdsourcing and characterization of twitter abusive behavior
Antigoni Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos, and Nicolas Kourtellis. 2018 · 2018
Earlier work this paper cites.
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé III, and Kate Crawford. 2018 · 2018
Earlier work this paper cites.
Examining gender and race bias in two hundred sentiment analysis systems
Svetlana Kiritchenko and Saif Mohammad. 2018 · 2018
Earlier work this paper cites.
Comparing prescriptive and descriptive gender stereotypes about children, adults, and the elderly
Anne M Koenig. 2018 · 2018
Earlier work this paper cites.
Obtaining reliable human ratings of valence, arousal, and dominance for 20,000 English words
Saif Mohammad. 2018 · 2018
Earlier work this paper cites.
Reducing gender bias in abusive language detection
Ji Ho Park, Jamin Shin, and Pascale Fung. 2018 · 2018
Earlier work this paper cites.
User-level race and ethnicity predictors from Twitter text
Daniel Preoţiuc-Pietro and Lyle Ungar. 2018 · 2018
Earlier work this paper cites.
Event2Mind: Commonsense inference on events, intents, and reactions
Hannah Rashkin, Maarten Sap, Emily Allaway, Noah A. Smith, and Yejin Choi. 2018 · 2018
Cited alongside, same era.
Gender bias in coreference resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard, and Benjamin Van Durme. 2018 · 2018
Cited alongside, same era.
Mind the GAP: A balanced corpus of gendered ambiguous pronouns
Kellie Webster, Marta Recasens, Vera Axelrod, and Jason Baldridge. 2018 · 2018
Cited alongside, same era.
Gender bias in coreference resolution: Evaluation and debiasing methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2018 · 2018
Cited alongside, same era.
SemEval-2019 task 5: Multilingual detection of hate speech against immigrants and women in Twitter
Valerio Basile, Cristina Bosco, Elisabetta Fersini, Debora Nozza, Viviana Patti, Francisco Manuel Rangel Pardo, Paolo Rosso, and Manuela Sanguinetti. 2019 · 2019
Cited alongside, same era.
AraWEAT: Multidimensional analysis of biases in Arabic word embeddings
Anne Lauscher, Rafik Takieddin, Simone Paolo Ponzetto, and Goran Glavaš. 2020 · 2020
Later among the works it cites.
UNQOVERing stereotyping biases via underspecified questions
Tao Li, Daniel Khashabi, Tushar Khot, Ashish Sabharwal, and Vivek Srikumar. 2020 · 2020
Later among the works it cites.
Mitigating gender bias for neural dialogue generation with adversarial learning
Haochen Liu, Wentao Wang, Yiqi Wang, Hui Liu, Zitao Liu, and Jiliang Tang. 2020b · 2020
Later among the works it cites.
Gender bias in neural natural language processing
Kaiji Lu, Piotr Mardziel, Fangjing Wu, Preetam Amancharla, and Anupam Datta. 2020 · 2020
Later among the works it cites.
Social, psychological, and demographic characteristics of dehumanization toward immigrants
David M. Markowitz and Paul Slovic. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nuanced metrics for measuring unintended bias with real data for text classification
Daniel Borkan, Lucas Dixon, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman. 2019 · 2019
Cited alongside, same era.
Racial bias in hate speech and abusive language detection datasets
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 2019
Cited alongside, same era.
Bias in bios: A case study of semantic representation bias in a high-stakes setting
Maria De-Arteaga, Alexey Romanov, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, and Adam Tauman Kalai. 2019 · 2019
Cited alongside, same era.
On measuring and mitigating biased inferences of word embeddings
Sunipa Dev, Tao Li, Jeff Phillips, and Vivek Srikumar. 2019 · 2019
Cited alongside, same era.
Attenuating bias in word vectors
Sunipa Dev and Jeff Phillips. 2019 · 2019
Cited alongside, same era.
Build it break it fix it for dialogue safety: Robustness from adversarial human attack
Emily Dinan, Samuel Humeau, Bharath Chintagunta, and Jason Weston. 2019 · 2019
Cited alongside, same era.
Women’s syntactic resilience and men’s grammatical luck: Gender-bias in part-of-speech tagging and dependency parsing
Aparna Garimella, Carmen Banea, E. Hovy, and Rada Mihalcea. 2019 · 2019
Cited alongside, same era.
Ninareh Mehrabi, Thamme Gowda, Fred Morstatter, Nanyun Peng, and A. Galstyan. 2020 · 2020
Later among the works it cites.
Detecting independent pronoun bias with partially-synthetic data generation
R. Munro and Alex Morrison. 2020 · 2020
Later among the works it cites.
Null it out: Guarding protected attributes by iterative nullspace projection
Shauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg. 2020 · 2020
Later among the works it cites.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A. Smith, and Yejin Choi. 2020 · 2020
Later among the works it cites.
“you are grounded!”: Latent name artifacts in pre-trained language models
Vered Shwartz, Rachel Rudinger, and Oyvind Tafjord. 2020 · 2020
Later among the works it cites.
Demoting racial bias in hate speech detection
Mengzhou Xia, Anjalie Field, and Yulia Tsvetkov. 2020 · 2020
Later among the works it cites.
Gender bias in multilingual embeddings and cross-lingual transfer
Jieyu Zhao, Subhabrata Mukherjee, Saghar Hosseini, Kai-Wei Chang, and Ahmed Hassan Awadallah. 2020 · 2020
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Closest in time.
Sociolinguistically driven approaches for just natural language processing
Su Lin Blodgett. 2021 · 2021
Closest in time.
Stereotyping norwegian salmon: an inventory of pitfalls in fairness benchmark datasets
Su Lin Blodgett, Gilsinia Lopez, Alexandra Olteanu, Robert Sim, and Hanna Wallach. 2021 · 2021
Closest in time.
OSCaR: Orthogonal subspace correction and rectification of biases in word embeddings
Sunipa Dev, Tao Li, Jeff M Phillips, and Vivek Srikumar. 2021a · 2021
Closest in time.
Bold: Dataset and metrics for measuring biases in open-ended language generation
Jwala Dhamala, Tony Sun, Varun Kumar, Satyapriya Krishna, Yada Pruksachatkun, Kai-Wei Chang, and Rahul Gupta. 2021 · 2021
Closest in time.
Measurement and fairness
Abigail Z Jacobs and Hanna Wallach. 2021 · 2021
Closest in time.
Lawyers are dishonest? quantifying representational harms in commonsense knowledge resources
Ninareh Mehrabi, Pei Zhou, Fred Morstatter, J. Pujara, Xiang Ren, and A. Galstyan. 2021 · 2021
Closest in time.
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2021 · 2021
Closest in time.
Societal biases in language generation: Progress and challenges
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2021 · 2021
Closest in time.
Ethical-advice taker: Do language models understand natural language interventions?
Jieyu Zhao, Daniel Khashabi, Tushar Khot, Ashish Sabharwal, and Kai-Wei Chang. 2021 · 2021
Closest in time.
Using natural sentences for understanding biases in language models
Sarah Alnegheimish, Alicia Guo, and Yi Sun. 2022 · 2022
Closest in time.
Responsible language technologies: Foreseeing and mitigating harms
Su Lin Blodgett, Q. Vera Liao, Alexandra Olteanu, Rada Mihalcea, Michael Muller, Morgan Klaus Scheuerman, Chenhao Tan, and Qian Yang. 2022 · 2022
Closest in time.
Fairlex: A multilingual benchmark for evaluating fairness in legal text processing
Ilias Chalkidis, Tommaso Pasini, Sheng Zhang, Letizia Tomada, Sebastian Felix Schwemer, and Anders Søgaard. 2022 · 2022
Closest in time.
Measuring fairness with biased rulers: A comparative study on bias metrics for pre-trained language models
Pieter Delobelle, Ewoenam Tokpo, Toon Calders, and Bettina Berendt. 2022 · 2022
Closest in time.
Socially aware bias measurements for Hindi language representations
Vijit Malik, Sunipa Dev, Akihiro Nishi, Nanyun Peng, and Kai-Wei Chang. 2022 · 2022
Closest in time.
Ethics sheets for AI tasks
Saif Mohammad. 2022 · 2022
Closest in time.
French crows-pairs: Extending a challenge dataset for measuring social bias in masked language models to a language other than english
Aurelie Neveol, Yoann Dupont, Julien Bezançon, and Karën Fort. 2022 · 2022
Closest in time.
You reap what you sow: On the challenges of bias evaluation under multilingual settings
Zeerak Talat, Aurélie Névéol, Stella Biderman, Miruna Clinciu, Manan Dey, Shayne Longpre, Sasha Luccioni, Maraim Masoud, Margaret Mitchell, Dragomir Radev, Shanya Sharma, Arjun Subramonian, Jaesung Tae, Samson Tan, Deepak Tunuguntla, and Oskar Van Der Wal. 2022 · 2022
Closest in time.
Measuring and mitigating name biases in neural machine translation
Jun Wang, Benjamin Rubinstein, and Trevor Cohn. 2022 · 2022
Closest in time.