Fetching the paper…
Reading the bibliography…
To recognize and mitigate harms from large language models (LLMs), we need to understand the prevalence and nuances of stereotypes in LLM outputs.
The mark of oppression; a psychosocial study of the american negro
Abram Kardiner and Lionel Ovesey. 1951 · 1951
Earlier work this paper cites.
RedditBias: A real-world resource for bias evaluation and debiasing of conversational language models
Soumya Barikeri, Anne Lauscher, Ivan Vulić, and Goran Glavaš. 2021 · 1955
Earlier work this paper cites.
Structural anthropology
Claude Lévi-Strauss. 1963 · 1963
Earlier work this paper cites.
Crows-pairs: A challenge dataset for measuring social biases in masked language models
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel Bowman. 2020 · 1967
Earlier work this paper cites.
The second sex, trans
Simone De Beauvoir. 1952 · 1974
Earlier work this paper cites.
Logic and conversation
Herbert P Grice. 1975 · 1975
Earlier work this paper cites.
Orientalism: Western concepts of the orient
Edward Said. 1978 · 1978
Earlier work this paper cites.
Marked and unmarked: A choice between unequals in semiotic structure
Linda R Waugh. 1982 · 1982
Earlier work this paper cites.
The combahee river collective statement
Combahee River Collective. 1983 · 1983
Earlier work this paper cites.
Asian-american women: Psychological responses to sexual exploitation and cultural stereotypes
Connie S Chan. 1988 · 1988
Earlier work this paper cites.
Black feminist thought in the matrix of domination
Patricia Hill Collins. 1990 · 1990
Earlier work this paper cites.
Gender stereotypes
Kay Deaux and Mary Kite. 1993 · 1993
Earlier work this paper cites.
White women, race matters: The social construction of whiteness
Ruth Frankenburg. 1993 · 1993
Earlier work this paper cites.
“Holy terror”: The implications of terrorism motivated by a religious imperative
Bruce Hoffman. 1995 · 1995
Earlier work this paper cites.
Race and the education of desire: Foucault’s history of sexuality and the colonial order of things
Ann Laura Stoler et al. 1995 · 1995
Earlier work this paper cites.
The Meaning of Difference: American Constructions of Race, Sex , volume 52
Karen E Rosenblum and Toni-Michelle C Travis. 1996 · 1996
Earlier work this paper cites.
Identity and difference , volume 3
Kathryn Woodward. 1997 · 1997
Earlier work this paper cites.
A sociology of the unmarked: Redirecting our focus
Wayne Brekhus. 1998 · 1998
Earlier work this paper cites.
The orientalization of asian women in america
Aki Uchida. 1998 · 1998
Earlier work this paper cites.
The inmates are running the asylum
Alan Cooper. 1999 · 1999
Earlier work this paper cites.
Feminist theory: From margin to center
Bell Hooks. 2000 · 2000
Earlier work this paper cites.
The deathly embrace: Orientalism and Asian American identity
Sheng-mei Ma. 2000 · 2000
Earlier work this paper cites.
Description and prescription: How gender stereotypes prevent women’s ascent up the organizational ladder
Madeline E Heilman. 2001 · 2001
Earlier work this paper cites.
Ethnic and national stereotypes: The princeton trilogy revisited and revised
Stephanie Madon, Max Guyll, Kathy Aboufadel, Eulices Montiel, Alison Smith, Polly Palumbo, and Lee Jussim. 2001 · 2001
Earlier work this paper cites.
The user as a personality-using personas as a tool for design
Stefan Blomkvist. 2002 · 2002
Earlier work this paper cites.
Adaptive testing: effects on user performance
Eva Jettmar and Clifford Nass. 2002 · 2002
Earlier work this paper cites.
Design as a minority discipline in a software company: toward requirements for a community of practice
Michael J Muller and Kenneth Carey. 2002 · 2002
Earlier work this paper cites.
Arab/muslim’otherness’: The role of racial constructions in the gulf war and the continuing crisis with iraq
Sina Ali Muscati. 2002 · 2002
Earlier work this paper cites.
Embracing the East: White women and American orientalism
Mari Yoshihara. 2002 · 2002
Earlier work this paper cites.
Common racial stereotypes
Szu-Hsien Chang and Brian H Kleiner. 2003 · 2003
Earlier work this paper cites.
Reel bad arabs: How hollywood vilifies a people
Jack G Shaheen. 2003 · 2003
Earlier work this paper cites.
What group?” studying whites and whiteness in the era of “color-blindness
Amanda E Lewis. 2004 · 2004
Earlier work this paper cites.
Black immigrants in the united states and the" cultural narratives" of ethnicity
Jemima Pierre. 2004 · 2004
Earlier work this paper cites.
Racism without racists: Color-blind racism and the persistence of racial inequality in the United States
Eduardo Bonilla-Silva. 2006 · 2006
Earlier work this paper cites.
The spirit of neoliberalism: From racial liberalism to neoliberal multiculturalism
Jodi Melamed. 2006 · 2006
Earlier work this paper cites.
Leader as Martial Artist: Techniques and Strategies for Resolving Conflict and Creating Community
Arnold Mindell. 2006 · 2006
Earlier work this paper cites.
Fightin’words: Lexical feature selection and evaluation for identifying the content of political conflict
Burt L Monroe, Michael P Colaresi, and Kevin M Quinn. 2008 · 2008
Cited alongside, same era.
Dangerous curves: Latina bodies in the media , volume 5
Isabel Molina-Guzmán. 2010 · 2010
Cited alongside, same era.
Superwoman schema: African american women’s views on stress, strength, and health
Cheryl L Woods-Giscombé. 2010 · 2010
Cited alongside, same era.
Othering, identity formation and agency
Sune Qvotrup Jensen. 2011 · 2011
Cited alongside, same era.
Arabs and muslims in the media
Evelyn Alsultany. 2012 · 2012
Cited alongside, same era.
Cultural identity, representation and othering
Fred Dervin. 2012 · 2012
Cited alongside, same era.
Social bias frames: Reasoning about social and power implications of language
Maarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky, Noah A Smith, and Yejin Choi. 2020 · 2020
Later among the works it cites.
Race and distant reading
Richard Jean So and Edwin Roland. 2020 · 2020
Later among the works it cites.
“You’re so exotic looking”: An intersectional analysis of asian american and pacific islander stereotypes
Sameena Azhar, Antonia RG Alvarez, Anne SJ Farina, and Susan Klumpner. 2021 · 2021
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Later among the works it cites.
Stereotyping norwegian salmon: an inventory of pitfalls in fairness benchmark datasets
Su Lin Blodgett, Gilsinia Lopez, Alexandra Olteanu, Robert Sim, and Hanna Wallach. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning latent personas of film characters
David Bamman, Brendan O’Connor, and Noah A Smith. 2013 · 2013
Cited alongside, same era.
Cisgenderism in family therapy: How everyday clinical practices can delegitimize people’s gender self-designations
Markie LC Blumer, Y Gavriel Ansara, and Courtney M Watson. 2013 · 2013
Cited alongside, same era.
An intersectional analysis of gender and ethnic stereotypes: Testing three hypotheses
Negin Ghavami and Letitia Anne Peplau. 2013 · 2013
Cited alongside, same era.
A different kind of whiteness: Marking and unmarking of social boundaries in the construction of hegemonic ethnicity
Orna Sasson-Levy. 2013 · 2013
Cited alongside, same era.
Vader: A parsimonious rule-based model for sentiment analysis of social media text
Clayton Hutto and Eric Gilbert. 2014 · 2014
Cited alongside, same era.
The rise of neoliberal feminism
Catherine Rottenberg. 2014 · 2014
Cited alongside, same era.
Generalized word shift graphs: a method for visualizing and explaining pairwise comparisons between texts
Ryan J Gallagher, Morgan R Frank, Lewis Mitchell, Aaron J Schwartz, Andrew J Reagan, Christopher M Danforth, and Peter Sheridan Dodds. 2021 · 2021
Later among the works it cites.
Detecting emergent intersectional biases: Contextualized word embeddings contain a distribution of human-like biases
Wei Guo and Aylin Caliskan. 2021 · 2021
Later among the works it cites.
Bias out-of-the-box: An empirical analysis of intersectional occupational biases in popular generative language models
Hannah Rose Kirk, Yennie Jun, Filippo Volpin, Haider Iqbal, Elias Benussi, Frederic Dreyer, Aleksandar Shtedritski, and Yuki Asano. 2021 · 2021
Later among the works it cites.
Pollution is colonialism
Max Liboiron. 2021 · 2021
Later among the works it cites.
Gender and representation bias in gpt-3 generated stories
Li Lucy and David Bamman. 2021 · 2021
Later among the works it cites.
Stereoset: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2021 · 2021
Later among the works it cites.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in nlp
Timo Schick, Sahana Udupa, and Hinrich Schütze. 2021 · 2021
Later among the works it cites.
When the echo chamber shatters: Examining the use of community-specific language post-subreddit ban
Milo Trujillo, Sam Rosenblatt, Guillermo De Anda-Jáuregui, Emily Moog, Briane Paul V Samson, Laurent Hébert-Dufresne, and Allison M Roth. 2021 · 2021
Later among the works it cites.
Ethical and social risks of harm from language models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al. 2021 · 2021
Later among the works it cites.
Re-claiming resilience and re-imagining welfare: A response to angela mcrobbie
Kim Allen. 2022 · 2022
Later among the works it cites.
Based on billions of words on the internet, people= men
April H Bailey, Adina Williams, and Andrei Cimpian. 2022 · 2022
Later among the works it cites.
Theory-grounded measurement of us social stereotypes in english language models
Yang Cao, Anna Sotnikova, Hal Daumé III, Rachel Rudinger, and Linda Zou. 2022 · 2022
Later among the works it cites.
“I’m a strong independent black woman”: The strong black woman schema and mental health in college-aged black women
Stephanie Castelin and Grace White. 2022 · 2022
Later among the works it cites.
Misogynistic terrorism: it has always been here
Caron E Gentry. 2022 · 2022
Later among the works it cites.
Surfacing racial stereotypes through identity portrayal
Gauri Kambhatla, Ian Stewart, and Rada Mihalcea. 2022 · 2022
Later among the works it cites.
Coauthor: Designing a human-ai collaborative writing dataset for exploring language model capabilities
Mina Lee, Percy Liang, and Qian Yang. 2022 · 2022
Later among the works it cites.
Holistic evaluation of language models
Percy Liang, Rishi Bommasani, Tony Lee, Dimitris Tsipras, Dilara Soylu, Michihiro Yasunaga, Yian Zhang, Deepak Narayanan, Yuhuai Wu, Ananya Kumar, et al. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
Bloom: A 176b-parameter open-access multilingual language model
Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, Matthias Gallé, et al. 2022 · 2022
Later among the works it cites.
“I’m sorry to hear that”: Finding new biases in language models with a holistic descriptor dataset
Eric Michael Smith, Melissa Hall, Melanie Kambadur, Eleonora Presani, and Adina Williams. 2022 · 2022
Later among the works it cites.
Evidence for hypodescent in visual semantic ai
Robert Wolfe, Mahzarin R. Banaji, and Aylin Caliskan. 2022 · 2022
Later among the works it cites.
American== white in multimodal language-and-image ai
Robert Wolfe and Aylin Caliskan. 2022a · 2022
Later among the works it cites.
Markedness in visual semantic ai
Robert Wolfe and Aylin Caliskan. 2022b · 2022
Later among the works it cites.
Long time no see! open-domain conversation with long-term persona memory
Xinchao Xu, Zhibin Gou, Wenquan Wu, Zheng-Yu Niu, Hua Wu, Haifeng Wang, and Shihang Wang. 2022 · 2022
Later among the works it cites.
Opt: Open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2022
Later among the works it cites.
SODAPOP: Open-ended discovery of social biases in social commonsense reasoning models
Haozhe An, Zongxia Li, Jieyu Zhao, and Rachel Rudinger. 2023 · 2023
Closest in time.
Openai: Introducing chatgpt
OpenAI. 2022 · 2023
Closest in time.
Gpt-4 technical report
OpenAI. 2023 · 2023
Closest in time.
Whose opinions do language models reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee, Percy Liang, and Tatsunori Hashimoto. 2023 · 2023
Closest in time.
Bbq: A hand-built bias benchmark for question answering
Alicia Parrish, Angelica Chen, Nikita Nangia, Vishakh Padmakumar, Jason Phang, Jana Thompson, Phu Mon Htut, and Samuel Bowman. 2022 · 2086
Closest in time.