Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have revolutionized natural language processing, but their susceptibility to biases poses significant challenges.
Language models are few-shot learners
Brown, T.B.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al · 1901
Earlier work this paper cites.
On measuring social biases in sentence encoders
May, C.; Wang, A.; Bordia, S.; Bowman, S.R.; Rudinger, R · 1903
Earlier work this paper cites.
Gender bias in contextualized word embeddings
Zhao, J.; Wang, T.; Yatskar, M.; Cotterell, R.; Ordonez, V.; Chang, K.W · 1904
Earlier work this paper cites.
Evaluating the underlying gender bias in contextualized word embeddings
Basta, C.; Costa-Jussà, M.R.; Casas, N · 1904
Earlier work this paper cites.
Improving neural conversational models with entropy-based data filtering
Csáky, R.; Purgai, P.; Recski, G · 1905
Earlier work this paper cites.
Mitigating gender bias in natural language processing: Literature review
Sun, T.; Gaut, A.; Tang, S.; Huang, Y.; ElSherief, M.; Zhao, J.; Mirza, D.; Belding, E.; Chang, K.W.; Wang, W.Y · 1906
Earlier work this paper cites.
Lost in translation: Loss and decay of linguistic richness in machine translation
Vanmassenhove, E.; Shterionov, D.; Way, A · 1906
Earlier work this paper cites.
Measuring bias in contextualized word representations
Kurita, K.; Vyas, N.; Pareek, A.; Black, A.W.; Tsvetkov, Y · 1906
Earlier work this paper cites.
Dialogpt: Large-scale generative pre-training for conversational response generation
Zhang, Y.; Sun, S.; Galley, M.; Chen, Y.C.; Brockett, C.; Gao, X.; Gao, J.; Liu, J.; Dolan, B · 1911
Earlier work this paper cites.
Dropout: A Simple Way to Prevent Neural Networks from Overfitting
Srivastava, N.; Hinton, G.; Krizhevsky, A.; Sutskever, I.; Salakhutdinov, R · 1958
Earlier work this paper cites.
A New Look at the Statistical Model Identification
Akaike, H · 1974
Earlier work this paper cites.
Estimating Causal Effects of Treatments in Randomized and Nonrandomized Studies
Rubin, D.B · 1974
Earlier work this paper cites.
Sample Selection Bias as a Specification Error
Heckman, J.J · 1979
Earlier work this paper cites.
Nonparametric Estimation in the Presence of Length Bias
Vardi, Y · 1982
Earlier work this paper cites.
The dictionary of affect in language. In The Measurement of Emotions
Whissell, C.M · 1989
Earlier work this paper cites.
Regression Shrinkage and Selection via the Lasso
Tibshirani, R · 1996
Earlier work this paper cites.
FinBERT: A Pretrained Language Model for Financial Communications
Yang, Y.; Yuan, Y.; Liu, L · 2006
Earlier work this paper cites.
From search engines to question answering systems—The problems of world knowledge, relevance, deduction and precisiation. In Capturing Intelligence
Zadeh, L.A · 2006
Earlier work this paper cites.
Opinion mining and sentiment analysis
Pang, B.; Lee, L · 2008
Earlier work this paper cites.
The Elements of Statistical Learning: Data Mining, Inference, and Prediction
Hastie, T.; Tibshirani, R.; Friedman, J · 2009
Earlier work this paper cites.
Bias in odds ratios by logistic regression modelling and sample size
Nemes, S.; Jonasson, J.M.; Genell, A.; Steineck, G · 2009
Earlier work this paper cites.
Natural Language Inference
MacCartney, B · 2009
Earlier work this paper cites.
Text classification using label names only: A language model self-training approach
Meng, Y.; Zhang, Y.; Huang, J.; Xiong, C.; Ji, H.; Zhang, C.; Han, J · 2010
Earlier work this paper cites.
Matching and sorting in online dating
Hitsch, G.J.; Hortaçsu, A.; Ariely, D · 2010
Earlier work this paper cites.
Measuring and reducing gendered correlations in pre-trained models
Webster, K.; Wang, X.; Tenney, I.; Beutel, A.; Pitler, E.; Pavlick, E.; Chen, J.; Chi, E.; Petrov, S · 2010
Earlier work this paper cites.
The Filter Bubble: How the New Personalized Web is Changing What We Read and How We Think
Pariser, E · 2011
Earlier work this paper cites.
Building and annotating the linguistically diverse NTU-MC (NTU-multilingual corpus)
Tan, L.; Bond, F · 2011
Earlier work this paper cites.
Cross-lingual semantic similarity of words as the similarity of their semantic word responses
Vulic, I.; Moens, M.F · 2013
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Bowman, S.R.; Angeli, G.; Potts, C.; Manning, C.D · 2015
Earlier work this paper cites.
Cyber hate speech on twitter: An application of machine classification and statistical modeling for policy and decision making
Burnap, P.; Williams, M.L · 2015
Earlier work this paper cites.
Tagging performance correlates with author age
Hovy, D.; Søgaard, A · 2015
Earlier work this paper cites.
Abstractive text summarization using sequence-to-sequence rnns and beyond
Nallapati, R.; Zhou, B.; Gulcehre, C.; Xiang, B · 2016
Earlier work this paper cites.
Machine Bias
Angwin, J.; Larson, J.; Mattu, S.; Kirchner, L · 2016
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? Debiasing word embeddings
Bolukbasi, T.; Chang, K.W.; Zou, J.Y.; Saligrama, V.; Kalai, A.T · 2016
Earlier work this paper cites.
Query expansion with locally-trained word embeddings
Diaz, F.; Mitra, B.; Craswell, N · 2016
Earlier work this paper cites.
Enhanced LSTM for natural language inference
Chen, Q.; Zhu, X.; Ling, Z.; Wei, S.; Jiang, H.; Inkpen, D · 2016
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Rajpurkar, P.; Zhang, J.; Lopyrev, K.; Liang, P · 2016
Earlier work this paper cites.
Women through the glass ceiling: Gender asymmetries in Wikipedia
Wagner, C.; Graells-Garrido, E.; Garcia, D.; Menczer, F · 2016
Earlier work this paper cites.
The social impact of natural language processing
Hovy, D.; Spruit, S.L · 2016
Earlier work this paper cites.
Neural Machine Translation ;
Sennrich, R · 2016
Earlier work this paper cites.
Analyzing the targets of hate in online social media
Silva, L.; Mondal, M.; Correa, D.; Benevenuto, F.; Weber, I · 2016
Earlier work this paper cites.
Causal Inference in Statistics: A Primer
Pearl, J.; Glymour, M.; Jewell, N.P · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Caliskan, A.; Bryson, J.J.; Narayanan, A · 2017
Earlier work this paper cites.
End-to-end neural coreference resolution
Lee, K.; He, L.; Lewis, M.; Zettlemoyer, L · 2017
Earlier work this paper cites.
Semeval-2017 task 1: Semantic textual similarity-multilingual and cross-lingual focused evaluation
Cer, D.; Diab, M.; Agirre, E.; Lopez-Gazpio, I.; Specia, L · 2017
Earlier work this paper cites.
Men also like shopping: Reducing gender bias amplification using corpus-level constraints
Zhao, J.; Wang, T.; Yatskar, M.; Ordonez, V.; Chang, K.W · 2017
Earlier work this paper cites.
Blodgett, S.L.; O’Connor, B · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A · 2017
Earlier work this paper cites.
Competent men and warm women: Gender stereotypes and backlash in image search results
Otterbacher, J.; Bates, J.; Clough, P · 2017
Earlier work this paper cites.
Style transfer from non-parallel text by cross-alignment
Shen, T.; Lei, T.; Barzilay, R.; Jaakkola, T · 2017
Earlier work this paper cites.
Inter-annotator agreement
Artstein, R · 2017
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.W.; Lee, K.; Toutanova, K · 2018
Earlier work this paper cites.
Deep Learning for Sentiment Analysis: A Survey
Zhang, L.; Wang, S.; Liu, B · 2018
Earlier work this paper cites.
Ensuring Fairness in Machine Learning to Advance Health Equity
Rajkomar, A.; Hardt, M.; Howell, M.D.; Corrado, G.; Chin, M.H · 2018
Earlier work this paper cites.
My Fair LADY: Detecting and Mitigating Bias in Job Advertisements
Chen, M.; Ma, Z.; Hannak, A.; Wilson, C · 2018
Earlier work this paper cites.
Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods
Zhao, J.; Wang, T.; Yatskar, M.; Ordonez, V.; Chang, K.W · 2018
Earlier work this paper cites.
Examining Gender and Race Bias in Two Hundred Sentiment Analysis Systems
Kiritchenko, S.; Mohammad, S.M · 2018
Earlier work this paper cites.
Learning Gender-Neutral Word Embeddings
Zhao, J.; Wang, T.; Yatskar, M.; Ordonez, V.; Chang, K.W · 2018
Earlier work this paper cites.
Word embeddings quantify 100 years of gender and ethnic stereotypes
Garg, N.; Schiebinger, L.; Jurafsky, D.; Zou, J · 2018
Earlier work this paper cites.
Women also snowboard: Overcoming bias in captioning models
Hendricks, L.A.; Burns, K.; Saenko, K.; Darrell, T.; Rohrbach, A · 2018
Earlier work this paper cites.
Gender bias in coreference resolution
Rudinger, R.; Naradowsky, J.; Leonard, B.; Van Durme, B · 2018
Earlier work this paper cites.
Addressing age-related bias in sentiment analysis
Díaz, M.; Johnson, I.; Lazar, A.; Piper, A.M.; Gergle, D · 2018
Earlier work this paper cites.
Know what you don’t know: Unanswerable questions for SQuAD
Rajpurkar, P.; Jia, R.; Liang, P · 2018
Earlier work this paper cites.
Mind the GAP: A balanced corpus of gendered ambiguous pronouns
Webster, K.; Recasens, M.; Axelrod, V.; Baldridge, J · 2018
Earlier work this paper cites.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Bender, E.M.; Friedman, B · 2018
Earlier work this paper cites.
Allennlp: A deep semantic natural language processing platform
Gardner, M.; Grus, J.; Neumann, M.; Tafjord, O.; Dasigi, P.; Liu, N.; Peters, M.; Schmitz, M.; Zettlemoyer, L · 2018
Earlier work this paper cites.
All the cool kids, how do they fit in?: Popularity and demographic biases in recommender evaluation and effectiveness
Ekstrand, M.D.; Tian, M.; Azpiazu, I.M.; Ekstrand, J.D.; Anuyah, O.; McNeill, D.; Pera, M.S · 2018
Earlier work this paper cites.
Equity of attention: Amortizing individual fairness in rankings
Biega, A.J.; Gummadi, K.P.; Weikum, G · 2018
Earlier work this paper cites.
Balanced neighborhoods for multi-sided fairness in recommendation
Burke, R.; Sonboli, N.; Ordonez-Gauger, A · 2018
Earlier work this paper cites.
News recommender systems–Survey and roads ahead
Karimi, M.; Jannach, D.; Jugovac, M · 2018
Earlier work this paper cites.
Multilingual neural machine translation for low-resource languages
Lakew, S.M.; Federico, M.; Negri, M.; Turchi, M · 2018
Cited alongside, same era.
Reducing Gender Bias in Abusive Language Detection
Park, J.H.; Shin, J.; Fung, P · 2018
Cited alongside, same era.
CTRL: A Conditional Transformer Language Model for Controllable Generation
Keskar, N.S.; McCann, B.; Varshney, L.R.; Xiong, C.; Socher, R · 2019
Cited alongside, same era.
The Risk of Racial Bias in Hate Speech Detection
Sap, M.; Card, D.; Gabriel, S.; Choi, Y.; Smith, N.A · 2019
Cited alongside, same era.
Dissecting Racial Bias in an Algorithm Used to Manage the Health of Populations
Obermeyer, Z.; Powers, B.; Vogeli, C.; Mullainathan, S · 2019
Cited alongside, same era.
Social Bias Frames: Reasoning about Social and Power Implications of Language
Sap, M.; Gabriel, S.; Qin, L.; Jurafsky, D.; Smith, N.A.; Choi, Y · 2019
“Im not Racist but…”: Discovering Bias in the Internal Knowledge of Large Language Models
Salinas, A.; Penafiel, L.; McCormack, R.; Morstatter, F · 2023
Later among the works it cites.
Who Authors the Internet? Analyzing Gender Diversity in ChatGPT-3 Training Data ;
Kuntz, J.B.; Silva, E.C · 2023
Later among the works it cites.
Gender bias and stereotypes in large language models
Kotek, H.; Dockum, R.; Sun, D · 2023
Later among the works it cites.
Breaking the bias: Gender fairness in LLMs using prompt engineering and in-context learning
Dwivedi, S.; Ghosh, S.; Dwivedi, S · 2023
Later among the works it cites.
Language model tokenizers introduce unfairness between languages
Petrov, A.; La Malfa, E.; Torr, P.; Bibi, A · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A Structural Probe for Finding Syntax in Word Representations
Hewitt, J.; Manning, C.D · 2019
Cited alongside, same era.
Bias in bios: A case study of semantic representation bias in a high-stakes setting
De-Arteaga, M.; Romanov, A.; Wallach, H.; Chayes, J.; Borgs, C.; Chouldechova, A.; Geyik, S.; Kenthapadi, K.; Kalai, A.T · 2019
Cited alongside, same era.
Density-dependent speed-up of particle transport in channels
Misiunas, K.; Keyser, U.F · 2019
Cited alongside, same era.
Natural questions: A benchmark for question answering research
Kwiatkowski, T.; Palomaki, J.; Redfield, O.; Collins, M.; Parikh, A.; Alberti, C.; Epstein, D.; Polosukhin, I.; Devlin, J.; Lee, K.; et al · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Radford, A.; Wu, J.; Child, R.; Luan, D.; Amodei, D.; Sutskever, I · 2019
Cited alongside, same era.
Deep learning based recommender system: A survey and new perspectives
Zhang, S.; Yao, L.; Sun, A.; Tay, Y · 2019
Cited alongside, same era.
Ahia, O.; Kumar, S.; Gonen, H.; Kasai, J.; Mortensen, D.R.; Smith, N.A.; Tsvetkov, Y · 2023
Later among the works it cites.
A survey on fairness in large language models
Li, Y.; Du, M.; Song, R.; Wang, X.; Wang, Y · 2023
Later among the works it cites.
Large language models as subpopulation representative models: A review
Simmons, G.; Hare, C · 2023
Later among the works it cites.
Slimpajama-dc: Understanding data combinations for llm training
Shen, Z.; Tao, T.; Ma, L.; Neiswanger, W.; Hestness, J.; Vassilieva, N.; Soboleva, D.; Xing, E · 2023
Later among the works it cites.
BiCapsHate: Attention to the linguistic context of hate via bidirectional capsules and hatebase
Kamal, A.; Anwar, T.; Sejwal, V.K.; Fazil, M · 2023
Later among the works it cites.
ROBBIE: Robust bias evaluation of large generative language models
Esiobu, D.; Tan, X.; Hosseini, S.; Ung, M.; Zhang, Y.; Fernandes, J.; Dwivedi-Yu, J.; Presani, E.; Williams, A.; Smith, E · 2023
Later among the works it cites.
Llama guard: Llm-based input-output safeguard for human-ai conversations
Inan, H.; Upasani, K.; Chi, J.; Rungta, R.; Iyer, K.; Mao, Y.; Tontchev, M.; Hu, Q.; Fuller, B.; Testuggine, D.; et al · 2023
Later among the works it cites.
Should chatgpt be biased? Challenges and risks of bias in large language models
Ferrara, E · 2023
Later among the works it cites.
OpenAGI: When llm meets domain experts
Ge, Y.; Hua, W.; Mei, K.; Tan, J.; Xu, S.; Li, Z.; Zhang, Y · 2023
Later among the works it cites.
Comparing Biases and the Impact of Multilingual Training across Multiple Languages
Levy, S.; John, N.; Liu, L.; Vyas, Y.; Ma, J.; Fujinuma, Y.; Ballesteros, M.; Castelli, V.; Roth, D · 2023
Later among the works it cites.
Artificial Intelligence Act. 2023
European Commission · 2023
Later among the works it cites.
Issues in cox proportional hazards model with unequal randomization
Li, H.; Qian H.; Li, C.T.; Hou, K · 2024
Closest in time.
Bias and fairness in large language models: A survey
Gallegos, I.O.; Rossi, R.A.; Barrow, J.; Tanjim, M.M.; Kim, S.; Dernoncourt, F.; Yu, T.; Zhang, R.; Ahmed, N.K · 2024
Closest in time.
Fairness in Large Language Models in three hours
Doan, T.V.; Wang, Z.; Nguyen, M.N.; Zhang, W · 2024
Closest in time.
Fairness in Transfer Learning for Natural Language Processing
Goldfarb-Tarrant, S · 2024
Closest in time.
Statistical Challenges with Dataset Construction: Why You Will Never Have Enough Images
Goldman, J.; Tsotsos, J.K · 2024
Closest in time.
Under the surface: Tracking the artifactuality of llm-generated data
Das, D.; De Langis, K.; Martin, A.; Kim, J.; Lee, M.; Kim, Z.M.; Hayati, S.; Owan, R.; Hu, B.; Parkar, R.; et al · 2024
Closest in time.
Large Language Models, Social Demography, and Hegemony: Comparing Authorship in Human and Synthetic Text
Alvero, A.; Lee, J.; Regla-Vargas, A.; Kizilec, R.; Joachims, T.; Antonio, A.L · 2024
Closest in time.
Challenging Systematic Prejudices: An Investigation into Bias Against Women and Girls in Large Language Models
UNESCO; IRCAI · 2024
Closest in time.
Bias in Language Models: A Survey
Ahmad, A.; Bhattacharyya, P · 2024
Closest in time.
Global Gallery: The Fine Art of Painting Culture Portraits through Multilingual Instruction Tuning
Mukherjee, A.; Caliskan, A.; Zhu, Z.; Anastasopoulos, A · 2024
Closest in time.
Neural embedding of beliefs reveals the role of relative dissonance in human decision-making
Lee, B.; Aiyappa, R.; Ahn, Y.Y.; Kwak, H.; An, J · 2024
Closest in time.
Shi, W.; Li, R.; Zhang, Y.; Ziems, C.; Horesh, R.; de Paula, R.A.; Yang, D · 2024
Closest in time.
Having beer after prayer? Measuring cultural bias in large language models
Naous, T.; Ryan, M.J.; Ritter, A.; Xu, W · 2024
Closest in time.
Not all countries celebrate thanksgiving: On the cultural dominance in large language models
Wang, W.; Jiao, W.; Huang, J.; Dai, R.; Huang, J.T.; Tu, Z.; Lyu, M · 2024
Closest in time.
Evaluating Racial Bias in Large Language Models: The Necessity for “SMOKY”
Bussaja, J · 2024
Closest in time.
Performance and biases of Large Language Models in public opinion simulation
Qu, Y.; Wang, J · 2024
Closest in time.
Semantic Change Characterization with LLMs using Rhetorics
de Sá, J.M.C.; Da Silveira, M.; Pruski, C · 2024
Closest in time.
Tokenization matters: Navigating data-scarce tokenization for gender inclusive language technologies
Ovalle, A.; Mehrabi, N.; Goyal, P.; Dhamala, J.; Chang, K.W.; Zemel, R.; Galstyan, A.; Pinter, Y.; Gupta, R · 2024
Closest in time.
A survey on evaluation of large language models
Chang, Y.; Wang, X.; Wang, J.; Wu, Y.; Yang, L.; Zhu, K.; Chen, H.; Yi, X.; Wang, C.; Wang, Y.; et al · 2024
Closest in time.
Large language models cannot replace human participants because they cannot portray identity groups
Wang, A.; Morgenstern, J.; Dickerson, J.P · 2024
Closest in time.
Unboxing Occupational Bias: Grounded Debiasing LLMs with US Labor Data
Gorti, A.; Gaur, M.; Chadha, A · 2024
Closest in time.
The political preferences of LLMs
Rozado, D · 2024
Closest in time.
Evaluating Vocabulary Usage in LLMs
Durward, M.; Thomson, C · 2024
Closest in time.
Data quality in NLP: Metrics and a comprehensive taxonomy
Dang, V.M.H.; Verma, R.M · 2024
Closest in time.
Human-LLM collaborative annotation through effective verification of LLM labels
Wang, X.; Kim, H.; Rahman, S.; Mitra, K.; Miao, Z · 2024
Closest in time.
Wu, Z.; Bulathwela, S.; Perez-Ortiz, M.; Koshiyama, A.S · 2024
Closest in time.
Towards trustworthy LLMs: A review on debiasing and dehallucinating in large language models
Lin, Z.; Guan, S.; Zhang, W.; Zhang, H.; Li, Y.; Zhang, H · 2024
Closest in time.
CausalBench: A Comprehensive Benchmark for Evaluating Causal Reasoning Capabilities of Large Language Models
Wang, Z · 2024
Closest in time.
All Should Be Equal in the Eyes of LMs: Counterfactually Aware Fair Text Generation
Banerjee, P.; Java, A.; Jandial, S.; Shahid, S.; Furniturewala, S.; Krishnamurthy, B.; Bhatia, S · 2024
Closest in time.
Interactive Analysis of LLMs using Meaningful Counterfactuals
Cheng, F.; Zouhar, V.; Chan, R.S.M.; Fürst, D.; Strobelt, H.; El-Assady, M · 2024
Closest in time.
FairMonitor: A Dual-framework for Detecting Stereotypes and Biases in Large Language Models
Bai, Y.; Zhao, J.; Shi, J.; Xie, Z.; Wu, X.; He, L · 2024
Closest in time.
The Bias that Lies Beneath: Qualitative Uncovering of Stereotypes in Large Language Models
Babonnaud, W.; Delouche, E.; Lahlouh, M · 2024
Closest in time.
Can LLMs Recognize Toxicity? Structured Toxicity Investigation Framework and Semantic-Based Metric
Koh, H.; Kim, D.; Lee, M.; Jung, K · 2024
Closest in time.
Hu, Z.; Piet, J.; Zhao, G.; Jiao, J.; Wagner, D · 2024
Closest in time.
An, H.; Acquaye, C.; Wang, C.; Li, Z.; Rudinger, R · 2024
Closest in time.
Evaluating the moral beliefs encoded in llms
Scherrer, N.; Shi, C.; Feder, A.; Blei, D · 2024
Closest in time.
Cognitive bias in high-stakes decision-making with llms
Echterhoff, J.; Liu, Y.; Alessa, A.; McAuley, J.; He, Z · 2024
Closest in time.
Do Multilingual Large Language Models Mitigate Stereotype Bias?
Nie, S.; Fromm, M.; Welch, C.; Görge, R.; Karimi, A.; Plepi, J.; Mowmita, N.; Flores-Herr, N.; Ali, M.; Flek, L · 2024
Closest in time.
Mitigating Large Language Model Bias: Automated Dataset Augmentation and Prejudice Quantification
Mondal, D.; Lipizzi, C · 2024
Closest in time.
Exploring large language models for bias mitigation and fairness
Serouis, I.M.; Sèdes, F · 2024
Closest in time.
Unified Transfer Learning in High-Dimensional Linear Regression
Liu, S.S · 2024
Closest in time.
Reducing Bias in Sentiment Analysis Models Through Causal Mediation Analysis and Targeted Counterfactual Training
Da, Y.; Bossa, M.N.; Berenguer, A.D.; Sahli, H · 2024
Closest in time.
Locating and mitigating gender bias in large language models
Cai, Y.; Cao, D.; Guo, R.; Wen, Y.; Liu, G.; Chen, E · 2024
Closest in time.
Steering LLMs Towards Unbiased Responses: A Causality-Guided Debiasing Framework
Li, J.; Tang, Z.; Liu, X.; Spirtes, P.; Zhang, K.; Leqi, L.; Liu, Y · 2024
Closest in time.
Causal Prompting: Debiasing Large Language Model Prompting based on Front-Door Adjustment
Zhang, C.; Zhang, L.; Zhou, D.; Xu, G · 2024
Closest in time.
EEOC Guidance on AI in Employment. 2023
EEOC · 2024
Closest in time.
Local Law 144: Automated Employment Decision Tools. 2021
New York City Council · 2024
Closest in time.
Section 1557 Final Rule. 2024
U.S. Department of Health and Human Services · 2024
Closest in time.
Assessing the potential of GPT-4 to perpetuate racial and gender biases in health care: A model evaluation study
Zack, T.; Lehman, E.; Suzgun, M.; Rodriguez, J.A · 2024
Closest in time.
Citation-Enhanced Generation for LLM-based Chatbot
Li, W.; Li, J.; Ma, W.; Liu, Y · 2024
Closest in time.
Up5: Unbiased foundation model for fairness-aware recommendation
Hua, W.; Ge, Y.; Xu, S.; Ji, J.; Zhang, Y · 2024
Closest in time.
From prejudice to parity: A new approach to debiasing large language model word embeddings
Rakshit, A.; Singh, S.; Keshari, S.; Chowdhury, A.G.; Jain, V.; Chadha, A · 2025
Closest in time.
Reducing Gender Bias in Word-Level Language Models with a Gender-Equalizing Loss Function
Qian, Y.; Muaz, U.; Zhang, B.; Hyun, J.W · 2031
Closest in time.