Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have emerged as an integral part of modern societies, powering user-facing applications such as personal assistants and enterprise applications like recruitment tools.
Mitigating gender bias in natural language processing: Literature review
Tony Sun, Andrew Gaut, Shirlyn Tang, Yuxin Huang, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth Belding, Kai-Wei Chang, and William Yang Wang. 2019 · 1906
Earlier work this paper cites.
Race, caste, and other invidious distinctions in social stratification
Gerald D Berreman. 1972 · 1972
Earlier work this paper cites.
The measurement of observer agreement for categorical data
J. Richard Landis and Gary G. Koch. 1977 · 1977
Earlier work this paper cites.
A comparison of symbolic racism theory and social dominance theory as explanations for racial policy attitude
Jim Sidanius, Erik Devereux, and Felicia Pratto. 1992 · 1992
Earlier work this paper cites.
An integrated threat theory of prejudice.” in stuart oskamp (ed.)
Walter Stephan and W.S. Cookie. 2000 · 2000
Earlier work this paper cites.
Self and social identity*
Naomi Ellemers, Russell Spears, and Bertjan Doosje. 2002 · 2002
Earlier work this paper cites.
A model of (often mixed) stereotype content: Competence and warmth respectively follow from perceived status and competition
Susan Fiske, Amy Cuddy, Peter Glick, and Jun Xu. 2002 · 2002
Earlier work this paper cites.
The social identity theory of intergroup behavior
Henri Tajfel and John C Turner. 2004 · 2004
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2005
Earlier work this paper cites.
Convokit: A toolkit for the analysis of conversations
Jonathan P Chang, Caleb Chiam, Liye Fu, Andrew Z Wang, Justine Zhang, and Cristian Danescu-Niculescu-Mizil. 2020 · 2005
Earlier work this paper cites.
Encyclopedia of race, ethnicity, and society , volume 1
Richard T Schaefer. 2008 · 2008
Earlier work this paper cites.
Social identity and self-categorization
Dominic Abrams and Michael A Hogg. 2010 · 2010
Earlier work this paper cites.
Hatebert: Retraining bert for abusive language detection in english
Tommaso Caselli, Valerio Basile, Jelena Mitrović, and Michael Granitzer. 2020 · 2010
Earlier work this paper cites.
Key female characters in film have more to talk about besides men: Automating the bechdel test
Apoorv Agarwal, Jiehan Zheng, Shruti Kamath, Sriramkumar Balasubramanian, and Shirin Ann Dey. 2015 · 2015
Earlier work this paper cites.
Caste and care: is Indian healthcare delivery system favourable for Dalits? , volume 350
Sobin George. 2015 · 2015
Earlier work this paper cites.
Rethinking employment discrimination harms
Jessica L Roberts. 2015 · 2015
Earlier work this paper cites.
History of medicine between tradition and modernity
Cristian Barsu. 2017 · 2017
Earlier work this paper cites.
Measuring the reliability of hate speech annotations: The case of the european refugee crisis
Björn Ross, Michael Rist, Guillermo Carbonell, Benjamin Cabrera, Nils Kurowsky, and Michael Wojatzki. 2017 · 2017
Earlier work this paper cites.
Ex machina: Personal attacks seen at scale
Ellery Wulczyn, Nithum Thain, and Lucas Dixon. 2017 · 2017
Earlier work this paper cites.
Content analysis: An introduction to its methodology
Klaus Krippendorff. 2018 · 2018
Earlier work this paper cites.
Reconciliations of caste and medical power in rural public health services
Sobin George. 2019 · 2019
Earlier work this paper cites.
Quantification of gender representation bias in commercial films based on image analysis
Ji Yoon Jang, Sangyoon Lee, and Byungjoo Lee. 2019 · 2019
Earlier work this paper cites.
Ethical considerations in ai-based recruitment
Dena F Mujtaba and Nihar R Mahapatra. 2019 · 2019
Earlier work this paper cites.
California accuses cisco of job discrimination based on indian employee’s caste
Paresh Dave. 2020 · 2020
Earlier work this paper cites.
RealToxicityPrompts: Evaluating neural toxic degeneration in language models
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A. Smith. 2020 · 2020
Earlier work this paper cites.
Detoxify
Laura Hanu and Unitary team. 2020 · 2020
Earlier work this paper cites.
Impact of films: Changes in young people’s attitudes after watching a movie
Tina Kubrak. 2020 · 2020
Earlier work this paper cites.
Hatexplain: A benchmark dataset for explainable hate speech detection
Binny Mathew, Punyajoy Saha, Seid Muhie Yimam, Chris Biemann, Pawan Goyal, and Animesh Mukherjee. 2020 · 2020
Earlier work this paper cites.
Mitigating bias in algorithmic hiring: Evaluating claims and practices
Manish Raghavan, Solon Barocas, Jon Kleinberg, and Karen Levy. 2020 · 2020
Earlier work this paper cites.
Just say no: Analyzing the stance of neural dialogue generation in offensive contexts
Ashutosh Baheti, Maarten Sap, Alan Ritter, and Mark Riedl. 2021 · 2021
Earlier work this paper cites.
Workplace bullying in healthcare facilities: Role of caste and reservation
Mrinal Prakash Barua and Anita Verma. 2021 · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
Stereotyping norwegian salmon: An inventory of pitfalls in fairness benchmark datasets
Su Lin Blodgett, Gilsinia Lopez, Alexandra Olteanu, Robert Sim, and Hanna Wallach. 2021 · 2021
Cited alongside, same era.
A Brief History of Education – From Ancient Greece to the Enlightenment , pages 39–55
Chris Brown and Ruth Luzmore. 2021 · 2021
Cited alongside, same era.
Understanding and countering stereotypes: A computational approach to the stereotype content model
Kathleen C. Fraser, Isar Nejadgholi, and Svetlana Kiritchenko. 2021 · 2021
Cited alongside, same era.
Ai recruitment algorithms and the dehumanization problem
Megan Fritts and Frank Cabrera. 2021 · 2021
Cited alongside, same era.
Knowledge distillation: A survey
Jianping Gou, Baosheng Yu, Stephen J Maybank, and Dacheng Tao. 2021 · 2021
Cited alongside, same era.
Chatgpt outperforms crowd workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli. 2023 · 2023
Later among the works it cites.
Chatgpt is reshaping crowd work
Caitlin Harrington. 2023 · 2023
Later among the works it cites.
Is ai recruiting (un) ethical? a human rights perspective on the use of ai for hiring
Anna Lena Hunkenschroer and Alexander Kriebitz. 2023 · 2023
Later among the works it cites.
Khyati Khandelwal, Manuel Tonneau, Andrew M Bean, Hannah Rose Kirk, and Scott A Hale. 2023 · 2023
Later among the works it cites.
Trustworthy llms: a survey and guideline for evaluating large language models’ alignment
Yang Liu, Yuanshun Yao, Jean-Francois Ton, Xiaoying Zhang, Ruocheng Guo Hao Cheng, Yegor Klochkov, Muhammad Faaiz Taufiq, and Hang Li. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
The perils of using Mechanical Turk to evaluate open-ended text generation
Marzena Karpinska, Nader Akoury, and Mohit Iyyer. 2021 · 2021
Cited alongside, same era.
The medical profession must urgently act on caste-based discrimination and harassment in their midst
Kiran Kumbhar. 2021 · 2021
Cited alongside, same era.
An empirical survey of the effectiveness of debiasing techniques for pre-trained language models
Nicholas Meade, Elinor Poole-Dayan, and Siva Reddy. 2021 · 2021
Cited alongside, same era.
A survey on bias and fairness in machine learning
Ninareh Mehrabi, Fred Morstatter, Nripsuta Saxena, Kristina Lerman, and Aram Galstyan. 2021 · 2021
Cited alongside, same era.
Re-imagining algorithmic fairness in india and beyond
Nithya Sambasivan, Erin Arnesen, Ben Hutchinson, Tulsee Doshi, and Vinodkumar Prabhakaran. 2021 · 2021
Cited alongside, same era.
Hi, my name is martha: Using names to measure and mitigate bias in generative dialogue models
Eric Michael Smith and Adina Williams. 2021 · 2021
Cited alongside, same era.
When my group is under attack: The development of a social identity threat scale
Rong Ma, Edward L. Fink, and Anita Atwell Seate. 2023 · 2023
Later among the works it cites.
Co-writing screenplays and theatre scripts with language models: Evaluation by industry professionals
Piotr Mirowski, Kory W Mathewson, Jaylen Pittman, and Richard Evans. 2023 · 2023
Later among the works it cites.
Exploring chatgpt for toxicity detection in github
Shyamal Mishra and Preetha Chatterjee. 2023 · 2023
Later among the works it cites.
A human-centered evaluation of a toxicity detection api: Testing transferability and unpacking latent attributes
Meena Devii Muralikumar, Yun Shan Yang, and David W. McDonald. 2023 · 2023
Later among the works it cites.
Having beer after prayer? measuring cultural bias in large language models
Tarek Naous, Michael J Ryan, and Wei Xu. 2023 · 2023
Later among the works it cites.
Caste identities and structures of threats: Stigma, prejudice, and social representation in indian universities
Gaurav J. Pathania, Sushrut Jadhav, Amit Thorat, David Mosse, and Sumeet Jain. 2023 · 2023
Later among the works it cites.
Fairness in language models beyond english: Gaps and challenges
Krithika Ramesh, Sunayana Sitaram, and Monojit Choudhury. 2023 · 2023
Later among the works it cites.
Petter Törnberg. 2023 · 2023
Later among the works it cites.
Performance and risk trade-offs for multi-word text prediction at scale
Aniket Vashishtha, S Sai Prasad, Payal Bajaj, Vishrav Chaudhary, Kate Cook, Sandipan Dandapat, Sunayana Sitaram, and Monojit Choudhury. 2023 · 2023
Later among the works it cites.
Investigating hiring bias in large language models
Akshaj Kumar Veldanda, Fabian Grob, Shailja Thakur, Hammond Pearce, Benjamin Tan, Ramesh Karri, and Siddharth Garg. 2023 · 2023
Later among the works it cites.
Veniamin Veselovsky, Manoel Horta Ribeiro, and Robert West. 2023 · 2023
Later among the works it cites.
On the robustness of chatgpt: An adversarial and out-of-distribution perspective
Jindong Wang, Xixu Hu, Wenxin Hou, Hao Chen, Runkai Zheng, Yidong Wang, Linyi Yang, Haojun Huang, Wei Ye, Xiubo Geng, et al. 2023 · 2023
Later among the works it cites.
Rongwu Xu, Brian S Lin, Shujian Yang, Tianqi Zhang, Weiyan Shi, Tianwei Zhang, Zhixuan Fang, Wei Xu, and Han Qiu. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P. Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica. 2023 · 2023
Later among the works it cites.
Lima: Less is more for alignment
Chunting Zhou, Pengfei Liu, Puxin Xu, Srini Iyer, Jiao Sun, Yuning Mao, Xuezhe Ma, Avia Efrat, Ping Yu, Lili Yu, et al. 2023 · 2023
Later among the works it cites.
Terry Yue Zhuo, Zhuang Li, Yujin Huang, Fatemeh Shiri, Weiqing Wang, Gholamreza Haffari, and Yuan-Fang Li. 2023 · 2023
Later among the works it cites.
Large legal fictions: Profiling legal hallucinations in large language models
Matthew Dahl, Varun Magesh, Mirac Suzgun, and Daniel E. Ho. 2024 · 2024
Closest in time.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer. 2024 · 2024
Closest in time.
MiniLLM: Knowledge distillation of large language models
Yuxian Gu, Li Dong, Furu Wei, and Minlie Huang. 2024 · 2024
Closest in time.
Dialect prejudice predicts ai decisions about people’s character, employability, and criminality
Valentin Hofmann, Pratyusha Ria Kalluri, Dan Jurafsky, and Sharese King. 2024 · 2024
Closest in time.
50 years of software
Michael Martinez. 2019 · 2024
Closest in time.
Best practices for prompt engineering with the openai api
OpenAI. 2024a · 2024
Closest in time.
Prompt engineering
OpenAI. 2024c · 2024
Closest in time.
Text generation models
OpenAI. 2024 · 2024
Closest in time.
Efficient toxic content detection by bootstrapping and distilling large language models
Jiang Zhang, Qiong Wu, Yiming Xu, Cheng Cao, Zheng Du, and Konstantinos Psounis. 2024 · 2024
Closest in time.
Llamafactory: Unified efficient fine-tuning of 100+ language models
Yaowei Zheng, Richong Zhang, Junhao Zhang, Yanhan Ye, Zheyan Luo, and Yongqiang Ma. 2024 · 2024
Closest in time.