Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated remarkable success across various domains.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Improvements in beam search.. In ICSLP , Vol. 94. 2143–2146
Volker Steinbiss, Bach-Hiep Tran, and Hermann Ney. 1994 · 1994
Earlier work this paper cites.
Statistical methods for speech recognition
Frederick Jelinek. 1998 · 1998
Earlier work this paper cites.
Two decades of statistical language modeling: Where do we go from here?
Ronald Rosenfeld. 2000 · 2000
Earlier work this paper cites.
Sentiment analysis: Capturing favorability using natural language processing. In Proceedings of the 2nd international conference on Knowledge capture . 70–77
Tetsuya Nasukawa and Jeonghee Yi. 2003 · 2003
Earlier work this paper cites.
UCI machine learning repository
Arthur Asuncion and David Newman. 2007 · 2007
Earlier work this paper cites.
Speech and language processing
Constituency Parsing. 2009 · 2009
Earlier work this paper cites.
Recurrent neural network based language model.. In Interspeech , Vol. 2. Makuhari, 1045–1048
Tomas Mikolov, Martin Karafiát, Lukas Burget, Jan Cernockỳ, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Towards Enhanced Opinion Classification using NLP Techniques.. In Proceedings of the workshop on Sentiment Analysis where AI meets Psychology (SAAIP 2011) . 101–107
Akshat Bakliwal, Piyush Arora, Ankit Patil, and Vasudeva Varma. 2011 · 2011
Earlier work this paper cites.
Fairness through awareness. In Proceedings of the 3rd innovations in theoretical computer science conference . 214–226
Cynthia Dwork, Moritz Hardt, Toniann Pitassi, Omer Reingold, and Richard Zemel. 2012 · 2012
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Glove: Global vectors for word representation. In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) . 1532–1543
Jeffrey Pennington, Richard Socher, and Christopher D Manning. 2014 · 2014
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Earlier work this paper cites.
Equality of opportunity in supervised learning
Moritz Hardt, Eric Price, and Nati Srebro. 2016 · 2016
Earlier work this paper cites.
Using the machine learning approach to predict patient survival from high-dimensional survival data. In IEEE International Conference on Bioinformatics and Biomedicine (BIBM)
Wenbin Zhang, Jian Tang, and Nuo Wang. 2016 · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Aylin Caliskan, Joanna J Bryson, and Arvind Narayanan. 2017 · 2017
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Earlier work this paper cites.
Conscientious classification: A data scientist’s guide to discrimination-aware classification
Brian d’Alessandro, Cathy O’Neil, and Tom LaGatta. 2017 · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Scott M Lundberg and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Phd forum: Recognizing human posture from time-changing wearable sensor data streams. In IEEE International Conference on Smart Computing (SMARTCOMP)
Wenbin Zhang. 2017 · 2017
Earlier work this paper cites.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M Bender and Batya Friedman. 2018 · 2018
Earlier work this paper cites.
Gender shades: Intersectional accuracy disparities in commercial gender classification. In Conference on fairness, accountability and transparency . PMLR, 77–91
Joy Buolamwini and Timnit Gebru. 2018 · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Word embeddings quantify 100 years of gender and ethnic stereotypes
Nikhil Garg, Londa Schiebinger, Dan Jurafsky, and James Zou. 2018 · 2018
Earlier work this paper cites.
Annotation artifacts in natural language inference data
Suchin Gururangan, Swabha Swayamdipta, Omer Levy, Roy Schwartz, Samuel R Bowman, and Noah A Smith. 2018 · 2018
Earlier work this paper cites.
Examining gender and race bias in two hundred sentiment analysis systems
Svetlana Kiritchenko and Saif M Mohammad. 2018 · 2018
Earlier work this paper cites.
Hypothesis only baselines in natural language inference
Adam Poliak, Jason Naradowsky, Aparajita Haldar, Rachel Rudinger, and Benjamin Van Durme. 2018 · 2018
Earlier work this paper cites.
Gender bias in coreference resolution
Rachel Rudinger, Jason Naradowsky, Brian Leonard, and Benjamin Van Durme. 2018 · 2018
Earlier work this paper cites.
Aequitas: A bias and fairness audit toolkit
Pedro Saleiro, Benedict Kuester, Loren Hinkson, Jesse London, Abby Stevens, Ari Anisfeld, Kit T Rodolfa, and Rayid Ghani. 2018 · 2018
Earlier work this paper cites.
Fairness definitions explained. In Proceedings of the international workshop on software fairness . 1–7
Sahil Verma and Julia Rubin. 2018 · 2018
Earlier work this paper cites.
Mind the GAP: A balanced corpus of gendered ambiguous pronouns
Kellie Webster, Marta Recasens, Vera Axelrod, and Jason Baldridge. 2018 · 2018
Earlier work this paper cites.
Content-bootstrapped collaborative filtering for medical article recommendations. In IEEE International Conference on Bioinformatics and Biomedicine (BIBM)
Wenbin Zhang and Jianwu Wang. 2018 · 2018
Earlier work this paper cites.
A deterministic self-organizing map approach and its application on satellite data based cloud type classification. In IEEE International Conference on Big Data (Big Data)
Wenbin Zhang, Jianwu Wang, Daeho Jin, Lazaros Oreopoulos, and Zhibo Zhang. 2018 · 2018
Earlier work this paper cites.
Gender bias in coreference resolution: Evaluation and debiasing methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2018 · 2018
Earlier work this paper cites.
AI Fairness 360: An extensible toolkit for detecting and mitigating algorithmic bias
Rachel KE Bellamy, Kuntal Dey, Michael Hind, Samuel C Hoffman, Stephanie Houde, Kalapriya Kannan, Pranay Lohia, Jacquelyn Martino, Sameep Mehta, Aleksandra Mojsilović, et al · 2019
Earlier work this paper cites.
Identifying and reducing gender bias in word-level language models
Shikha Bordia and Samuel R Bowman. 2019 · 2019
Earlier work this paper cites.
Understanding the origins of bias in word embeddings. In International conference on machine learning . PMLR, 803–811
Marc-Etienne Brunet, Colleen Alkalay-Houlihan, Ashton Anderson, and Richard Zemel. 2019 · 2019
Earlier work this paper cites.
Bias and fairness in natural language processing. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP): Tutorial Abstracts
Kai-Wei Chang, Vinodkumar Prabhakaran, and Vicente Ordonez. 2019 · 2019
Earlier work this paper cites.
Bias in bios: A case study of semantic representation bias in a high-stakes setting. In proceedings of the Conference on Fairness, Accountability, and Transparency . 120–128
Maria De-Arteaga, Alexey Romanov, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, and Adam Tauman Kalai. 2019 · 2019
Earlier work this paper cites.
Reducing sentiment bias in language models via counterfactual evaluation
Po-Sen Huang, Huan Zhang, Ray Jiang, Robert Stanforth, Johannes Welbl, Jack Rae, Vishal Maini, Dani Yogatama, and Pushmeet Kohli. 2019 · 2019
Earlier work this paper cites.
Classification of Sindhi headline news documents based on TF-IDF text analysis scheme
Irfan Ali Kandhro, Sahar Zafar Jumani, Ajab Ali Lashari, Saima Sipy Nangraj, Qurban Ali Lakhan, Mirza Taimoor Baig, and Subhash Guriro. 2019 · 2019
Earlier work this paper cites.
Measuring bias in contextualized word representations
Keita Kurita, Nidhi Vyas, Ayush Pareek, Alan W Black, and Yulia Tsvetkov. 2019 · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Earlier work this paper cites.
Deep learning in computer vision: Methods, interpretation, causation, and fairness
Nikhil Malik and Param Vir Singh. 2019 · 2019
Earlier work this paper cites.
On measuring social biases in sentence encoders
Chandler May, Alex Wang, Shikha Bordia, Samuel R Bowman, and Rachel Rudinger. 2019 · 2019
Earlier work this paper cites.
Challenges for mitigating bias in algorithmic hiring
Manish Raghavan and Solon Barocas. 2019 · 2019
Earlier work this paper cites.
Julian Salazar, Davis Liang, Toan Q Nguyen, and Katrin Kirchhoff. 2019 · 2019
Earlier work this paper cites.
The risk of racial bias in hate speech detection. In Proceedings of the 57th annual meeting of the association for computational linguistics . 1668–1678
Maarten Sap, Dallas Card, Saadia Gabriel, Yejin Choi, and Noah A Smith. 2019 · 2019
Earlier work this paper cites.
Predictive biases in natural language processing models: A conceptual framework and overview
Deven Shah, H Andrew Schwartz, and Dirk Hovy. 2019 · 2019
Earlier work this paper cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2019 · 2019
Earlier work this paper cites.
FAHT: an adaptive fairness-aware decision tree classifier. In International Joint Conference on Artificial Intelligence (IJCAI) . 1480–1486
Wenbin Zhang and Eirini Ntoutsi. 2019 · 2019
Earlier work this paper cites.
On fairness-aware learning for non-discriminative decision-making. In International Conference on Data Mining Workshops (ICDMW) . 1072–1079
Wenbin Zhang, Xuejiao Tang, and Jianwu Wang. 2019 · 2019
Earlier work this paper cites.
Gender bias in contextualized word embeddings
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Ryan Cotterell, Vicente Ordonez, and Kai-Wei Chang. 2019 · 2019
Earlier work this paper cites.
Counterfactual data augmentation for mitigating gender stereotypes in languages with rich morphology
Ran Zmigrod, Sabrina J Mielke, Hanna Wallach, and Ryan Cotterell. 2019 · 2019
Earlier work this paper cites.
Unmasking contextual stereotypes: Measuring and mitigating BERT’s gender bias
Marion Bartl, Malvina Nissim, and Albert Gatt. 2020 · 2020
Earlier work this paper cites.
Sex and gender differences and biases in artificial intelligence for biomedicine and healthcare
Davide Cirillo, Silvina Catuara-Solarz, Czuee Morey, Emre Guney, Laia Subirats, Simona Mellino, Annalisa Gigante, Alfonso Valencia, María José Rementeria, Antonella Santuccione Chadha, et al · 2020
Earlier work this paper cites.
Mitigating demographic Bias in AI-based resume filtering. In Adjunct publication of the 28th ACM conference on user modeling, adaptation and personalization . 268–275
Ketki V Deshpande, Shimei Pan, and James R Foulds. 2020 · 2020
Earlier work this paper cites.
On measuring and mitigating biased inferences of word embeddings. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 34. 7659–7666
Sunipa Dev, Tao Li, Jeff M Phillips, and Vivek Srikumar. 2020 · 2020
Earlier work this paper cites.
Social chemistry 101: Learning to reason about social and moral norms
Maxwell Forbes, Jena D Hwang, Vered Shwartz, Maarten Sap, and Yejin Choi. 2020 · 2020
Earlier work this paper cites.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A Smith. 2020 · 2020
Earlier work this paper cites.
Intrinsic bias metrics do not correlate with application bias
Seraphina Goldfarb-Tarrant, Rebecca Marchant, Ricardo Muñoz Sánchez, Mugdha Pandya, and Adam Lopez. 2020 · 2020
Earlier work this paper cites.
UNQOVERing stereotyping biases via underspecified questions
Tao Li, Tushar Khot, Daniel Khashabi, Ashish Sabharwal, and Vivek Srikumar. 2020 · 2020
Earlier work this paper cites.
Gender bias in neural natural language processing
Kaiji Lu, Piotr Mardziel, Fangjing Wu, Preetam Amancharla, and Anupam Datta. 2020 · 2020
Earlier work this paper cites.
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2020 · 2020
Earlier work this paper cites.
CrowS-pairs: A challenge dataset for measuring social biases in masked language models
Nikita Nangia, Clara Vania, Rasika Bhalerao, and Samuel R Bowman. 2020 · 2020
Earlier work this paper cites.
Null it out: Guarding protected attributes by iterative nullspace projection
Shauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg. 2020 · 2020
Earlier work this paper cites.
" Nice Try, Kiddo": Investigating Ad Hominems in Dialogue Responses
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, and Nanyun Peng. 2020 · 2020
Earlier work this paper cites.
Using machine learning to automate mammogram images analysis. In IEEE International Conference on Bioinformatics and Biomedicine (BIBM) . 757–764
Xuejiao Tang, Liuhua Zhang, et al · 2020
Earlier work this paper cites.
Measuring and reducing gendered correlations in pre-trained models
Kellie Webster, Xuezhi Wang, Ian Tenney, Alex Beutel, Emily Pitler, Ellie Pavlick, Jilin Chen, Ed Chi, and Slav Petrov. 2020 · 2020
Earlier work this paper cites.
Learning fairness and graph deep generation in dynamic environments
Wenbin Zhang. 2020 · 2020
Earlier work this paper cites.
Online decision trees with fairness
Wenbin Zhang and Liang Zhao. 2020 · 2020
Earlier work this paper cites.
Gender bias in multilingual embeddings and cross-lingual transfer
Jieyu Zhao, Subhabrata Mukherjee, Saghar Hosseini, Kai-Wei Chang, and Ahmed Hassan Awadallah. 2020 · 2020
Earlier work this paper cites.
Persistent anti-muslim bias in large language models. In Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society . 298–306
Abubakar Abid, Maheen Farooqi, and James Zou. 2021 · 2021
Earlier work this paper cites.
Soumya Barikeri, Anne Lauscher, Ivan Vulić, and Goran Glavaš. 2021 · 2021
Earlier work this paper cites.
Harms of gender exclusivity and challenges in non-binary representation in language technologies
Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian, Jeff M Phillips, and Kai-Wei Chang. 2021a · 2021
Earlier work this paper cites.
On measures of biases and harms in NLP
Sunipa Dev, Emily Sheng, Jieyu Zhao, Aubrie Amstutz, Jiao Sun, Yu Hou, Mattie Sanseverino, Jiin Kim, Akihiro Nishi, Nanyun Peng, et al · 2021
Earlier work this paper cites.
Bold: Dataset and metrics for measuring biases in open-ended language generation. In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency . 862–872
Jwala Dhamala, Tony Sun, Varun Kumar, Satyapriya Krishna, Yada Pruksachatkun, Kai-Wei Chang, and Rahul Gupta. 2021 · 2021
Earlier work this paper cites.
Improving gender fairness of pre-trained language models without catastrophic forgetting
Zahra Fatemi, Chen Xing, Wenhao Liu, and Caiming Xiong. 2021 · 2021
Earlier work this paper cites.
Five sources of bias in natural language processing
Dirk Hovy and Shrimai Prabhumoye. 2021 · 2021
Earlier work this paper cites.
Bias out-of-the-box: An empirical analysis of intersectional occupational biases in popular generative language models
Hannah Rose Kirk, Yennie Jun, Filippo Volpin, Haider Iqbal, Elias Benussi, Frederic Dreyer, Aleksandar Shtedritski, and Yuki Asano. 2021 · 2021
Cited alongside, same era.
Considering the possibilities and pitfalls of Generative Pre-trained Transformer 3 (GPT-3) in healthcare delivery
Diane M Korngiebel and Sean D Mooney. 2021 · 2021
Cited alongside, same era.
Sustainable modular debiasing of language models
Anne Lauscher, Tobias Lueken, and Goran Glavaš. 2021 · 2021
Cited alongside, same era.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Cited alongside, same era.
Collecting a large-scale gender bias dataset for coreference resolution and machine translation
ROBBIE: Robust Bias Evaluation of Large Generative Language Models. In The 2023 Conference on Empirical Methods in Natural Language Processing
David Esiobu, Xiaoqing Tan, Saghar Hosseini, Megan Ung, Yuchen Zhang, Jude Fernandes, Jane Dwivedi-Yu, Eleonora Presani, Adina Williams, and Eric Michael Smith. 2023 · 2023
Later among the works it cites.
Winoqueer: A community-in-the-loop benchmark for anti-lgbtq+ bias in large language models
Virginia K Felkner, Ho-Chun Herbert Chang, Eugene Jang, and Jonathan May. 2023 · 2023
Later among the works it cites.
Shangbin Feng, Chan Young Park, Yuhan Liu, and Yulia Tsvetkov. 2023 · 2023
Later among the works it cites.
Should chatgpt be biased? challenges and risks of bias in large language models
Emilio Ferrara. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shahar Levy, Koren Lazar, and Gabriel Stanovsky. 2021 · 2021
Cited alongside, same era.
Research on unsupervised feature learning for Android malware detection based on Restricted Boltzmann Machines
Zhen Liu, Ruoyu Wang, Nathalie Japkowicz, Deyu Tang, Wenbin Zhang, and Jie Zhao. 2021 · 2021
Cited alongside, same era.
An empirical survey of the effectiveness of debiasing techniques for pre-trained language models
Nicholas Meade, Elinor Poole-Dayan, and Siva Reddy. 2021 · 2021
Cited alongside, same era.
A survey on bias and fairness in machine learning
Ninareh Mehrabi, Fred Morstatter, Nripsuta Saxena, Kristina Lerman, and Aram Galstyan. 2021 · 2021
Cited alongside, same era.
Eric Mitchell, Charles Lin, Antoine Bosselut, Chelsea Finn, and Christopher D Manning. 2021 · 2021
Cited alongside, same era.
Mitigating harm in language models with conditional-likelihood filtration
Helen Ngo, Cooper Raterink, João GM Araújo, Ivan Zhang, Carol Chen, Adrien Morisot, and Nicholas Frosst. 2021 · 2021
Cited alongside, same era.
HONEST: Measuring hurtful sentence completion in language models. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics
Debora Nozza, Federico Bianchi, Dirk Hovy, et al · 2021
Cited alongside, same era.
Probing toxic content in large pre-trained language models. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . 4262–4274
Nedjma Ousidhoum, Xinran Zhao, Tianqing Fang, Yangqiu Song, and Dit-Yan Yeung. 2021 · 2021
Cited alongside, same era.
When the majority is wrong: Modeling annotator disagreement for subjective tasks. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 6715–6726
Eve Fleisig, Rediet Abebe, and Dan Klein. 2023a · 2023
Later among the works it cites.
Bias and fairness in large language models: A survey
Isabel O Gallegos, Ryan A Rossi, Joe Barrow, Md Mehrab Tanjim, Sungchul Kim, Franck Dernoncourt, Tong Yu, Ruiyi Zhang, and Nesreen K Ahmed. 2023 · 2023
Later among the works it cites.
Not what you’ve signed up for: Compromising real-world llm-integrated applications with indirect prompt injection. In Proceedings of the 16th ACM Workshop on Artificial Intelligence and Security . 79–90
Kai Greshake, Sahar Abdelnabi, Shailesh Mishra, Christoph Endres, Thorsten Holz, and Mario Fritz. 2023 · 2023
Later among the works it cites.
Evaluating large language models: A comprehensive survey
Zishan Guo, Renren Jin, Chuang Liu, Yufei Huang, Dan Shi, Linhao Yu, Yan Liu, Jiaxuan Li, Bojian Xiong, Deyi Xiong, et al · 2023
Later among the works it cites.
Vision-language models performing zero-shot tasks exhibit gender-based disparities
Melissa Hall, Laura Gustafson, Aaron Adcock, Ishan Misra, and Candace Ross. 2023 · 2023
Later among the works it cites.
Blind judgement: Agent-based supreme court modelling with gpt
Sil Hamilton. 2023 · 2023
Later among the works it cites.
Bias assessment and mitigation in llm-based code generation
Dong Huang, Qingwen Bu, Jie Zhang, Xiaofei Xie, Junjie Chen, and Heming Cui. 2023a · 2023
Later among the works it cites.
Trustgpt: A benchmark for trustworthy and responsible large language models
Yue Huang, Qihui Zhang, Lichao Sun, et al · 2023
Later among the works it cites.
ChatGPT by OpenAI: The End of Litigation Lawyers?
Kwan Yuen Iu and Vanessa Man-Yi Wong. 2023 · 2023
Later among the works it cites.
Co-writing with opinionated language models affects users’ views. In Proceedings of the 2023 CHI conference on human factors in computing systems . 1–15
Maurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson, and Mor Naaman. 2023 · 2023
Later among the works it cites.
Better to Ask in English: Cross-Lingual Evaluation of Large Language Models for Healthcare Queries
Yiqiao Jin, Mohit Chandra, Gaurav Verma, Yibo Hu, Munmun De Choudhury, and Srijan Kumar. 2023 · 2023
Later among the works it cites.
Gender bias and stereotypes in large language models. In Proceedings of The ACM Collective Intelligence Conference . 12–24
Hadas Kotek, Rikker Dockum, and David Sun. 2023 · 2023
Later among the works it cites.
A survey on fairness in large language models. arXiv. doi: 10.48550
Y Li, M Du, R Song, X Wang, and Y Wang. 2023 · 2023
Later among the works it cites.
Debiasing algorithm through model adaptation
Tomasz Limisiewicz, David Mareček, and Tomáš Musil. 2023 · 2023
Later among the works it cites.
Gili Lior and Gabriel Stanovsky. 2023 · 2023
Later among the works it cites.
SeGDroid: An Android malware detection method based on sensitive function call graph learning
Zhen Liu, Ruoyu Wang, Nathalie Japkowicz, Heitor Murilo Gomes, Bitao Peng, and Wenbin Zhang. 2023 · 2023
Later among the works it cites.
Gender-inclusive grammatical error correction through augmentation
Gunnar Lund, Kostiantyn Omelianchuk, and Igor Samokhin. 2023 · 2023
Later among the works it cites.
Logic against bias: Textual entailment mitigates stereotypical sentence reasoning
Hongyin Luo and James Glass. 2023 · 2023
Later among the works it cites.
Queenie Luo, Michael J Puett, and Michael D Smith. 2023 · 2023
Later among the works it cites.
Automating Bias Testing of LLMs. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 1705–1707
Sergio Morales, Robert Clarisó, and Jordi Cabot. 2023 · 2023
Later among the works it cites.
Never too late to learn: Regularizing gender bias in coreference resolution. In Proceedings of the Sixteenth ACM International Conference on Web Search and Data Mining . 15–23
SunYoung Park, Kyuri Choi, Haeun Yu, and Youngjoong Ko. 2023 · 2023
Later among the works it cites.
AI Psychometrics: Using psychometric inventories to obtain psychological profiles of large language models
Max Pellert, Clemens M Lechner, Claudia Wagner, Beatrice Rammstedt, and Markus Strohmaier. 2023 · 2023
Later among the works it cites.
Layered bias: Interpreting bias in pretrained large language models. In Proceedings of the 6th BlackboxNLP Workshop: Analyzing and Interpreting Neural Networks for NLP . 284–295
Nirmalendu Prakash and Roy Ka-Wei Lee. 2023 · 2023
Later among the works it cites.
Computational approaches streamlining drug discovery
Anastasiia V Sadybekov and Vsevolod Katritch. 2023 · 2023
Later among the works it cites.
Drug discovery companies are customizing ChatGPT: here’s how
Neil Savage. 2023 · 2023
Later among the works it cites.
Test Suites Task: Evaluation of Gender Fairness in MT with MuST-SHE and INES
Beatrice Savoldi, Marco Gaido, Matteo Negri, and Luisa Bentivogli. 2023 · 2023
Later among the works it cites.
Missed Opportunities in Fair AI. In Proceedings of the 2023 SIAM International Conference on Data Mining (SDM) . SIAM, 961–964
Nripsuta Ani Saxena, Wenbin Zhang, and Cyrus Shahabi. 2023 · 2023
Later among the works it cites.
Large language models encode clinical knowledge
Karan Singhal, Shekoofeh Azizi, Tao Tu, S Sara Mahdavi, Jason Wei, Hyung Won Chung, Nathan Scales, Ajay Tanwani, Heather Cole-Lewis, Stephen Pfohl, et al · 2023
Later among the works it cites.
Measuring gender bias in natural language processing: Incorporating gender-neutral linguistic forms for non-binary gender identities in abusive speech detection. In Proceedings of the 14th International Conference on Recent Advances in Natural Language Processing . 1121–1131
Nasim Sobhani, Kinshuk Sengupta, and Sarah Jane Delany. 2023 · 2023
Later among the works it cites.
Comparing traditional and llm-based search for consumer choice: A randomized experiment
Sofia Eleni Spatharioti, David M Rothschild, Daniel G Goldstein, and Jake M Hofman. 2023 · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Later among the works it cites.
Nationality bias in text generation
Pranav Narayanan Venkit, Sanjana Gautam, Ruchi Panchanadikar, Ting-Hao’Kenneth’ Huang, and Shomir Wilson. 2023 · 2023
Later among the works it cites.
" kelly is a warm person, joseph is a role model": Gender biases in llm-generated reference letters
Yixin Wan, George Pu, Jiao Sun, Aparna Garimella, Kai-Wei Chang, and Nanyun Peng. 2023a · 2023
Later among the works it cites.
Pre-trained language models in biomedical domain: A systematic survey
Benyou Wang, Qianqian Xie, Jiahuan Pei, Zhihong Chen, Prayag Tiwari, Zhao Li, and Jie Fu. 2023g · 2023
Later among the works it cites.
Are Large Language Models Really Robust to Word-Level Perturbations?
Haoyu Wang, Guozheng Ma, Cong Yu, Ning Gui, Linrui Zhang, Zhiqi Huang, Suwei Ma, Yongzhe Chang, Sen Zhang, Li Shen, et al · 2023
Later among the works it cites.
Chatcad: Interactive computer-aided diagnosis on medical image using large language models
Sheng Wang, Zihao Zhao, Xi Ouyang, Qian Wang, and Dinggang Shen. 2023h · 2023
Later among the works it cites.
All languages matter: On the multilingual safety of large language models
Wenxuan Wang, Zhaopeng Tu, Chang Chen, Youliang Yuan, Jen-tse Huang, Wenxiang Jiao, and Michael R Lyu. 2023d · 2023
Later among the works it cites.
Mitigating multisource biases in graph neural networks via real counterfactual samples. In 2023 IEEE International Conference on Data Mining (ICDM) . IEEE, 638–647
Zichong Wang, Giri Narasimhan, Xin Yao, and Wenbin Zhang. 2023b · 2023
Later among the works it cites.
Preventing discriminatory decision-making in evolving data streams. In Proceedings of the 2023 ACM Conference on Fairness, Accountability, and Transparency . 149–159
Zichong Wang, Nripsuta Saxena, Tongjia Yu, Sneha Karki, Tyler Zetty, Israat Haque, Shan Zhou, Dukka Kc, Ian Stockwell, Xuyu Wang, et al · 2023
Later among the works it cites.
Zichong Wang, Yang Zhou, Meikang Qiu, Israat Haque, Laura Brown, Yi He, Jianwu Wang, David Lo, and Wenbin Zhang. 2023i · 2023
Later among the works it cites.
Adept: A debiasing prompt framework. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 37. 10780–10788
Ke Yang, Charles Yu, Yi R Fung, Manling Li, and Heng Ji. 2023 · 2023
Later among the works it cites.
Empowering LLM-based machine translation with cultural awareness
Binwei Yao, Ming Jiang, Diyi Yang, and Junjie Hu. 2023a · 2023
Later among the works it cites.
Editing large language models: Problems, methods, and opportunities
Yunzhi Yao, Peng Wang, Bozhong Tian, Siyuan Cheng, Zhoubo Li, Shumin Deng, Huajun Chen, and Ningyu Zhang. 2023b · 2023
Later among the works it cites.
Tackling Bias in Pre-trained Language Models: Current Trends and Under-represented Societies
Vithya Yogarajan, Gillian Dobbie, Te Taka Keegan, and Rostam J Neuwirth. 2023 · 2023
Later among the works it cites.
Unlearning bias in language models by partitioning gradients. In Findings of the Association for Computational Linguistics: ACL 2023 . 6032–6048
Charles Yu, Sullam Jeoung, Anish Kasi, Pengfei Yu, and Heng Ji. 2023 · 2023
Later among the works it cites.
Should we attend more or less? modulating attention for fairness
Abdelrahman Zayed, Goncalo Mordido, Samira Shabanian, and Sarath Chandar. 2023a · 2023
Later among the works it cites.
Individual fairness under uncertainty
Wenbin Zhang, Zichong Wang, Juyong Kim, Cheng Cheng, Thomas Oommen, Pradeep Ravikumar, and Jeremy Weiss. 2023c · 2023
Later among the works it cites.
Fairness with censorship and group constraints
Wenbin Zhang and Jeremy C Weiss. 2023 · 2023
Later among the works it cites.
Zhiyuan Zhang, Deli Chen, Hao Zhou, Fandong Meng, Jie Zhou, and Xu Sun. 2023a · 2023
Later among the works it cites.
Chbias: Bias evaluation and mitigation of chinese conversational language models
Jiaxu Zhao, Meng Fang, Zijing Shi, Yitong Li, Ling Chen, and Mykola Pechenizkiy. 2023a · 2023
Later among the works it cites.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Later among the works it cites.
Causal-debias: Unifying debiasing in pretrained language models and fine-tuning via causal invariant learning. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 4227–4241
Fan Zhou, Yuzhou Mao, Liu Yu, Yi Yang, and Ting Zhong. 2023 · 2023
Later among the works it cites.
FairAIED: Navigating fairness, bias, and ethics in educational AI applications
Sribala Vidyadhari Chinta, Zichong Wang, Zhipeng Yin, Nhat Hoang, Matthew Gonzalez, Tai Le Quy, and Wenbin Zhang. 2024a · 2024
Closest in time.
AI-Driven Healthcare: A Survey on Ensuring Fairness and Mitigating Bias
Sribala Vidyadhari Chinta, Zichong Wang, Xingyu Zhang, Thang Doan Viet, Ayesha Kashif, Monique Antoinette Smith, and Wenbin Zhang. 2024b · 2024
Closest in time.
History, Development, and Principles of Large Language Models-An Introductory Survey
Zhibo Chu, Shiwen Ni, Zichong Wang, Xi Feng, Chengming Li, Xiping Hu, Ruifeng Xu, Min Yang, and Wenbin Zhang. 2024a · 2024
Closest in time.
Fairness in Large Language Models: A Taxonomic Survey
Zhibo Chu, Zichong Wang, and Wenbin Zhang. 2024b · 2024
Closest in time.
Risk taxonomy, mitigation, and assessment benchmarks of large language model systems
Tianyu Cui, Yanling Wang, Chuanpu Fu, Yong Xiao, Sijia Li, Xinhao Deng, Yunpeng Liu, Qinglin Zhang, Ziyi Qiu, Peiyang Li, et al · 2024
Closest in time.
Fairness definitions in language models explained
Thang Viet Doan, Zhibo Chu, Zichong Wang, and Wenbin Zhang. 2024a · 2024
Closest in time.
Uncertain Boundaries: Multidisciplinary Approaches to Copyright Issues in Generative AI
Jocelyn Dzuong, Zichong Wang, and Wenbin Zhang. 2024 · 2024
Closest in time.
Large language models in radiology: fundamentals, applications, ethical considerations, risks, and future directions
Tugba Akinci D’Antonoli, Arnaldo Stanzione, Christian Bluethgen, Federica Vernuccio, Lorenzo Ugga, Michail E Klontzas, Renato Cuocolo, Roberto Cannella, and Burak Koçak. 2024 · 2024
Closest in time.
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
Masahiro Kaneko, Danushka Bollegala, Naoaki Okazaki, and Timothy Baldwin. 2024 · 2024
Closest in time.
Gpt-4 passes the bar exam
Daniel Martin Katz, Michael James Bommarito, Shang Gao, and Pablo Arredondo. 2024 · 2024
Closest in time.
Evaluating Biases in Context-Dependent Health Questions
Sharon Levy, Tahilin Sanchez Karver, William D Adler, Michelle R Kaufman, and Mark Dredze. 2024 · 2024
Closest in time.
Addressing cognitive bias in medical language models
Samuel Schmidgall, Carl Harris, Ime Essien, Daniel Olshvang, Tawsifur Rahman, Ji Woong Kim, Rojin Ziaei, Jason Eshraghian, Peter Abadir, and Rama Chellappa. 2024 · 2024
Closest in time.
Uncovering Stereotypes in Large Language Models: A Task Complexity-based Approach. In Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers) . 1841–1857
Hari Shrawgi, Prasanjit Rath, Tushar Singhal, and Sandipan Dandapat. 2024 · 2024
Closest in time.
Trustllm: Trustworthiness in large language models
Lichao Sun, Yue Huang, Haoran Wang, Siyuan Wu, Qihui Zhang, Chujie Gao, Yixin Huang, Wenhan Lyu, Yixuan Zhang, Xiner Li, et al · 2024
Closest in time.
History, Development, and Principles of Large Language Models-An Introductory Survey
Zichong Wang, Zhibo Chu, Thang Viet Doan, Shiwen Ni, Min Yang, and Wenbin Zhang. 2024b · 2024
Closest in time.
Toward Fair Graph Neural Networks via Real Counterfactual Samples
Zichong Wang, Meikang Qiu, Min Chen, Malek Ben Salem, Xin Yao, and Wenbin Zhang. 2024d · 2024
Closest in time.
Group Fairness with Individual and Censorship Constraints. In 27th European Conference on Artificial Intelligence
Zichong Wang and Wenbin Zhang. 2024 · 2024
Closest in time.
Doremi: Optimizing data mixtures speeds up language model pretraining
Sang Michael Xie, Hieu Pham, Xuanyi Dong, Nan Du, Hanxiao Liu, Yifeng Lu, Percy S Liang, Quoc V Le, Tengyu Ma, and Adams Wei Yu. 2024 · 2024
Closest in time.
A Comprehensive Survey of Image and Video Generative AI: Recent Advances, Variants, and Applications
Shamim Yazdani, Nripsuta Saxena, Zichong Wang, Yanzhao Wu, and Wenbin Zhang. 2024 · 2024
Closest in time.
Improving Fairness in Machine Learning Software via Counterfactual Fairness Thinking. In Proceedings of the 2024 IEEE/ACM 46th International Conference on Software Engineering: Companion Proceedings . 420–421
Zhipeng Yin, Zichong Wang, and Wenbin Zhang. 2024 · 2024
Closest in time.
Fairness with Censorship: Bridging the Gap between Fairness Research and Real-World Deployment. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 38. 22685–22685
Wenbin Zhang. 2024 · 2024
Closest in time.
Graph Fairness via Authentic Counterfactuals: Tackling Structural and Causal Challenges
Zichong Wang, Zhipeng Yin, Fang Liu, Zhen Liu, Christine Lisetti, Rui Yu, Shaowei Wang, Jun Liu, Sukumar Ganapati, Shuigeng Zhou, and Wenbin Zhang. 2025b · 2025
Closest in time.
FG-SMOTE: Towards Fair Node Classification with Graph Neural Network
Zichong Wang, Zhipeng Yin, Yuying Zhang, Liping Yang, Tingting Zhang, Niki Pissinou, Yu Cai, Shu Hu, Yun Li, Liang Zhao, and Wenbin Zhang. 2025c · 2025
Closest in time.