Fetching the paper…
Reading the bibliography…
Stereotype benchmark datasets are crucial to detect and mitigate social stereotypes about groups of people in NLP models.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Unifiedqa: Crossing format boundaries with a single qa system
Daniel Khashabi, Sewon Min, Tushar Khot, Ashish Sabharwal, Oyvind Tafjord, Peter Clark, and Hannaneh Hajishirzi. 2020 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
The scientific study of national stereotypes
Otto Klineberg. 1951 · 1951
Earlier work this paper cites.
The measurement of meaning
Charles Egerton Osgood, George J Suci, and Percy H Tannenbaum. 1957 · 1957
Earlier work this paper cites.
Linguistic stereotypes and social distance
Ramdas Borude. 1966 · 1966
Earlier work this paper cites.
Crows-pairs: A challenge dataset for measuring social biases in masked language models
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel Bowman. 2020 · 1967
Earlier work this paper cites.
Electra: Pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V Le, and Christopher D Manning. 2020 · 2003
Earlier work this paper cites.
Deberta: Decoding-enhanced bert with disentangled attention
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2020 · 2006
Earlier work this paper cites.
Stereotyping and impression formation: How categorical thinking shapes person perception
Kimberly A Quinn, C Neil Macrae, and Galen V Bodenhausen. 2007 · 2007
Earlier work this paper cites.
Warmth and competence as universal dimensions of social perception: The stereotype content model and the bias map
Amy JC Cuddy, Susan T Fiske, and Peter Glick. 2008 · 2008
Earlier work this paper cites.
Accuracy of united states regional personality stereotypes
Katherine H. Rogers and Dustin Wood. 2010 · 2010
Earlier work this paper cites.
Henry the nurse is a doctor too: Implicitly examining children’s gender stereotypes for male and female occupational roles
Makeba Parramore Wilbourn and Daniel W Kee. 2010 · 2010
Earlier work this paper cites.
Communal and agentic content in social cognition: A dual perspective model
Andrea E Abele and Bogdan Wojciszke. 2014 · 2014
Earlier work this paper cites.
A dictionary of psychology
Andrew M Colman. 2015 · 2015
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Earlier work this paper cites.
The abc of stereotypes about groups: Agency/socioeconomic success, conservative–progressive beliefs, and communion
Alex Koch, Roland Imhoff, Ron Dotsch, Christian Unkelbach, and Hans Alves. 2016 · 2016
Earlier work this paper cites.
A model of (often mixed) stereotype content: Competence and warmth respectively follow from perceived status and competition
Susan T Fiske, Amy JC Cuddy, Peter Glick, and Jun Xu. 2018 · 2018
Cited alongside, same era.
Studying the cognitive map of the u.s. states: Ideology and prosperity stereotypes predict interstate prejudice
Alex Koch, Nicolas Kervyn, Matthieu Kervyn, and Roland Imhoff. 2018 · 2018
Cited alongside, same era.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Gender bias in coreference resolution: Evaluation and debiasing methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2018 · 2018
Cited alongside, same era.
Bias and fairness in natural language processing
Kai-Wei Chang, Vinodkumar Prabhakaran, and Vicente Ordonez. 2019 · 2019
Cited alongside, same era.
OSCaR: Orthogonal subspace correction and rectification of biases in word embeddings
Sunipa Dev, Tao Li, Jeff M Phillips, and Vivek Srikumar. 2021 · 2021
Later among the works it cites.
The importance of modeling social factors of language: Theory and practice
Dirk Hovy and Diyi Yang. 2021 · 2021
Later among the works it cites.
Stereoset: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2021 · 2021
Later among the works it cites.
On releasing annotator-level labels and information in datasets
Vinodkumar Prabhakaran, Aida Mostafazadeh Davani, and Mark Diaz. 2021 · 2021
Later among the works it cites.
Multitask prompted training enables zero-shot task generalization
Victor Sanh, Albert Webson, Colin Raffel, Stephen Bach, Lintang Sutawika, Zaid Alyafeai, Antoine Chaffin, Arnaud Stiegler, Arun Raja, Manan Dey, et al. 2021 · 2021
Later among the works it cites.
Re-contextualizing fairness in nlp: The case of india
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Quantifying social biases in contextual word representations
Keita Kurita, Nidhi Vyas, Ayush Pareek, Alan W Black, and Yulia Tsvetkov. 2019 · 2019
Cited alongside, same era.
The woman worked as a babysitter: On biases in language generation
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2019 · 2019
Cited alongside, same era.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Cited alongside, same era.
Language (technology) is power: A critical survey of “bias” in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé III, and Hanna Wallach. 2020 · 2020
Cited alongside, same era.
On measuring and mitigating biased inferences of word embeddings
Sunipa Dev, Tao Li, Jeff M Phillips, and Vivek Srikumar. 2020 · 2020
Cited alongside, same era.
UNQOVERing stereotyping biases via underspecified questions
Tao Li, Daniel Khashabi, Tushar Khot, Ashish Sabharwal, and Vivek Srikumar. 2020 · 2020
Cited alongside, same era.
Between subjectivity and imposition: Power dynamics in data annotation for computer vision
Milagros Miceli, Martin Schuessler, and Tianling Yang. 2020 · 2020
Cited alongside, same era.
Shaily Bhatt, Sunipa Dev, Partha Talukdar, Shachi Dave, and Vinodkumar Prabhakaran. 2022 · 2022
Later among the works it cites.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al. 2022 · 2022
Later among the works it cites.
Crowdworksheets: Accounting for individual and collective identities underlying crowdsourced dataset annotation
Mark Díaz, Ian Kivlichan, Rachel Rosen, Dylan Baker, Razvan Amironesei, Vinodkumar Prabhakaran, and Emily Denton. 2022 · 2022
Later among the works it cites.
Challenges and strategies in cross-cultural NLP
Daniel Hershcovich, Stella Frank, Heather Lent, Miryam de Lhoneux, Mostafa Abdou, Stephanie Brandl, Emanuele Bugliarello, Laura Cabello Piqueras, Ilias Chalkidis, Ruixiang Cui, Constanza Fierro, Katerina Margatina, Phillip Rust, and Anders Søgaard. 2022 · 2022
Later among the works it cites.
French CrowS-pairs: Extending a challenge dataset for measuring social bias in masked language models to a language other than English
Aurélie Névéol, Yoann Dupont, Julien Bezançon, and Karën Fort. 2022 · 2022
Later among the works it cites.
A spontaneous stereotype content model: Taxonomy, properties, and prediction
Gandalf Nicolas, Xuechunzi Bai, and Susan T Fiske. 2022 · 2022
Later among the works it cites.
Cultural incongruencies in artificial intelligence
Vinodkumar Prabhakaran, Rida Qadri, and Ben Hutchinson. 2022 · 2022
Later among the works it cites.
Data cards: Purposeful and transparent dataset documentation for responsible ai
Mahima Pushkarna, Andrew Zaldivar, and Oddur Kjartansson. 2022 · 2022
Later among the works it cites.
Taxonomy of risks posed by language models
Laura Weidinger, Jonathan Uesato, Maribeth Rauh, Conor Griffin, Po-Sen Huang, John Mellor, Amelia Glaese, Myra Cheng, Borja Balle, Atoosa Kasirzadeh, et al. 2022 · 2022
Later among the works it cites.
Building stereotype repositories with complementary approaches for scale and depth
Sunipa Dev, Akshita Jha, Jaya Goyal, Dinesh Tewari, Shachi Dave, and Vinodkumar Prabhakaran. 2023 · 2023
Closest in time.
Bbq: A hand-built bias benchmark for question answering
Alicia Parrish, Angelica Chen, Nikita Nangia, Vishakh Padmakumar, Jason Phang, Jana Thompson, Phu Mon Htut, and Samuel Bowman. 2022 · 2086
Closest in time.