Fetching the paper…
Reading the bibliography…
Large language models (LLMs) reflect societal norms and biases, especially about gender.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 1910
Earlier work this paper cites.
Intrinsic bias metrics do not correlate with application bias
Seraphina Goldfarb-Tarrant, Rebecca Marchant, Ricardo Muñoz Sánchez, Mugdha Pandya, and Adam Lopez. 2021 · 1940
Earlier work this paper cites.
CrowS-pairs: A challenge dataset for measuring social biases in masked language models
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel R. Bowman. 2020 · 1967
Earlier work this paper cites.
An argument for basic emotions
Paul Ekman. 1992 · 1992
Earlier work this paper cites.
Evidence for universality and cultural variation of differential emotion response patterning
Klaus R Scherer and Harald G Wallbott. 1994 · 1994
Earlier work this paper cites.
The group as a basis for emergent stereotype consensus
S Alexander Haslam, John C Turner, Penelope J Oakes, Craig McGarty, and Katherine J Reynolds. 1997 · 1997
Earlier work this paper cites.
The gender stereotyping of emotions
E Ashby Plant, Janet Shibley Hyde, Dacher Keltner, and Patricia G Devine. 2000 · 2000
Earlier work this paper cites.
Epistemic injustice: Power and the ethics of knowing
Miranda Fricker. 2007 · 2007
Earlier work this paper cites.
Aristotle’s account of the subjection of women
Dana Jalbert Stauffer. 2008 · 2008
Earlier work this paper cites.
Reflections on the silencing the self scale and its origins
Dana Crowley Jack. 2011 · 2011
Earlier work this paper cites.
Gender and emotion: What we think we know, what we need to know, and why it matters
Stephanie A Shields. 2013 · 2013
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? Debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Aylin Caliskan, Joanna J. Bryson, and Arvind Narayanan. 2017 · 2017
Earlier work this paper cites.
The moral psychology of anger
Myisha Cherry and Owen Flanagan. 2017 · 2017
Earlier work this paper cites.
The trouble with bias
Kate Crawford. 2017 · 2017
Earlier work this paper cites.
The moral psychology of sadness
Anna Gotlib. 2017 · 2017
Earlier work this paper cites.
Gender stereotypes
Naomi Ellemers. 2018 · 2018
Cited alongside, same era.
IEST: WASSA-2018 implicit emotions shared task
Roman Klinger, Orphée De Clercq, Saif Mohammad, and Alexandra Balahur. 2018 · 2018
Cited alongside, same era.
SemEval-2018 task 1: Affect in tweets
Saif Mohammad, Felipe Bravo-Marquez, Mohammad Salameh, and Svetlana Kiritchenko. 2018 · 2018
Cited alongside, same era.
Gender bias in coreference resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard, and Benjamin Van Durme. 2018 · 2018
Cited alongside, same era.
On measuring gender bias in translation of gender-neutral pronouns
Won Ik Cho, Ji Won Kim, Seok Min Kim, and Nam Soo Kim. 2019 · 2019
Cited alongside, same era.
The woman worked as a babysitter: On biases in language generation
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2019 · 2019
Cited alongside, same era.
HONEST: Measuring hurtful sentence completion in language models
Debora Nozza, Federico Bianchi, and Dirk Hovy. 2021 · 2021
Later among the works it cites.
Gender bias in machine translation
Beatrice Savoldi, Marco Gaido, Luisa Bentivogli, Matteo Negri, and Marco Turchi. 2021 · 2021
Later among the works it cites.
Revealing persona biases in dialogue systems
Emily Sheng, Josh Arnold, Zhou Yu, Kai-Wei Chang, and Nanyun Peng. 2021 · 2021
Later among the works it cites.
Theory-grounded measurement of U.S. social stereotypes in English language models
Yang Trista Cao, Anna Sotnikova, Hal Daumé III, Rachel Rudinger, and Linda Zou. 2022 · 2022
Later among the works it cites.
Marked personas: Using natural language prompts to measure stereotypes in language models
Myra Cheng, Esin Durmus, and Dan Jurafsky. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Generating responses with a specific emotion in dialog
Zhenqiao Song, Xiaoqing Zheng, Lu Liu, Mu Xu, and Xuanjing Huang. 2019 · 2019
Cited alongside, same era.
Evaluating gender bias in machine translation
Gabriel Stanovsky, Noah A. Smith, and Luke Zettlemoyer. 2019 · 2019
Cited alongside, same era.
Mitigating gender bias in natural language processing: Literature review
Tony Sun, Andrew Gaut, Shirlyn Tang, Yuxin Huang, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth Belding, Kai-Wei Chang, and William Yang Wang. 2019 · 2019
Cited alongside, same era.
Emotion-aware chat machine: Automatic emotional response generation for human-like emotional interaction
Wei Wei, Jiayi Liu, Xianling Mao, Guibing Guo, Feida Zhu, Pan Zhou, and Yuchong Hu. 2019 · 2019
Cited alongside, same era.
Multi-dimensional gender bias classification
Emily Dinan, Angela Fan, Ledell Wu, Jason Weston, Douwe Kiela, and Adina Williams. 2020 · 2020
Cited alongside, same era.
“you sound just like your father” commercial machine translation systems include stylistic biases
Dirk Hovy, Federico Bianchi, and Tommaso Fornaciari. 2020 · 2020
Cited alongside, same era.
Ameet Deshpande, Vishvak Murahari, Tanmay Rajpurohit, Ashwin Kalyan, and Karthik Narasimhan. 2023 · 2023
Later among the works it cites.
Regulation of the European Parliament and of the Council on laying down harmonised rules on artificial intelligence (Artificial Intelligence Act)
European Commission. 2023 · 2023
Later among the works it cites.
Bias runs deep: Implicit reasoning biases in persona-assigned llms
Shashank Gupta, Vaishnavi Shrivastava, Ameet Deshpande, Ashwin Kalyan, Peter Clark, Ashish Sabharwal, and Tushar Khot. 2023 · 2023
Later among the works it cites.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, et al. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
“kelly is a warm person, joseph is a role model”: Gender biases in LLM-generated reference letters
Yixin Wan, George Pu, Jiao Sun, Aparna Garimella, Kai-Wei Chang, and Nanyun Peng. 2023a · 2023
Later among the works it cites.
Are personalized stochastic parrots more dangerous? evaluating persona biases in dialogue systems
Yixin Wan, Jieyu Zhao, Aman Chadha, Nanyun Peng, and Kai-Wei Chang. 2023b · 2023
Later among the works it cites.
Decodingtrust: A comprehensive assessment of trustworthiness in GPT models
Boxin Wang, Weixin Chen, Hengzhi Pei, Chulin Xie, Mintong Kang, Chenhui Zhang, Chejian Xu, Zidi Xiong, Ritik Dutta, Rylan Schaeffer, et al. 2023 · 2023
Later among the works it cites.
MentalRiskES: A new corpus for early detection of mental disorders in Spanish
Alba M. Mármol Romero, Adrián Moreno Muñoz, Flor Miriam Plaza-del-Arco, M. Dolores Molina González, María Teresa Martín Valdivia, L. Alfonso Ureña-López, and Arturo Montejo Ráez. 2024 · 2024
Closest in time.
Emotion analysis in NLP: Trends, gaps and roadmap for future directions
Flor Miriam Plaza-del-Arco, Alba Curry, Amanda Cercas Curry, and Dirk Hovy. 2024 · 2024
Closest in time.