Fetching the paper…
Reading the bibliography…
Large language models (LLMs) offer significant potential as tools to support an expanding range of decision-making tasks.
Reducing sentiment bias in language models via counterfactual evaluation
Po-Sen Huang, Huan Zhang, Ray Jiang, Robert Stanforth, Johannes Welbl, Jack Rae, Vishal Maini, Dani Yogatama, and Pushmeet Kohli. 2019 · 1911
Earlier work this paper cites.
Forming impressions of personality
Solomon E Asch. 1946 · 1946
Earlier work this paper cites.
Judgment under uncertainty: Heuristics and biases: Biases in judgments reveal some heuristics of thinking under uncertainty
Amos Tversky and Daniel Kahneman. 1974 · 1974
Earlier work this paper cites.
Illusory correlation in interpersonal perception: A cognitive basis of stereotypic judgments
David L Hamilton and Robert K Gifford. 1976 · 1976
Earlier work this paper cites.
A two-process account of long-term serial position effects
Arthur M Glenberg, Margaret M Bradley, Jennifer A Stevenson, Thomas A Kraus, Marilyn J Tkachuk, Ann L Gretz, Joel H Fish, and BettyAnn M Turpin. 1980 · 1980
Earlier work this paper cites.
The framing of decisions and the psychology of choice
Amos Tversky and Daniel Kahneman. 1981 · 1981
Earlier work this paper cites.
Judgment under uncertainty: Heuristics and biases
Daniel Kahneman, Paul Slovic, and Amos Tversky. 1982 · 1982
Earlier work this paper cites.
Status quo bias in decision making
William Samuelson and Richard Zeckhauser. 1988 · 1988
Earlier work this paper cites.
Stereoset: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2020 · 2004
Earlier work this paper cites.
Beyond accuracy: Behavioral testing of nlp models with checklist
Marco Tulio Ribeiro, Tongshuang Wu, Carlos Guestrin, and Sameer Singh. 2020 · 2005
Earlier work this paper cites.
Imagery and interpretations in social phobia: Support for the combined cognitive biases hypothesis
Colette R Hirsch, David M Clark, and Andrew Mathews. 2006 · 2006
Earlier work this paper cites.
Efficacy of bias awareness in debiasing oil and gas judgments
Matthew B Welsh, Steve H Begg, and Reidar B Bratvold. 2007 · 2007
Earlier work this paper cites.
The combined cognitive bias hypothesis in depression
Jonas Everaert, Ernst HW Koster, and Nazanin Derakshan. 2012 · 2012
Earlier work this paper cites.
Bias dilemma: de-biasing and the consequent introduction of new biases
Jiulin Teng. 2013 · 2013
Earlier work this paper cites.
Debiasing through raising awareness reduces the anchoring bias
Carolyn Mair, Martin Shepperd, et al. 2014 · 2014
Earlier work this paper cites.
The evolution of cognitive bias
Martie G Haselton, Daniel Nettle, and Paul W Andrews. 2015 · 2015
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Cited alongside, same era.
Gender bias in coreferencliang2021towardse resolution: Evaluation and debiasing methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2018 · 2018
Cited alongside, same era.
50 years of test (un) fairness: Lessons for machine learning
Ben Hutchinson and Margaret Mitchell. 2019 · 2019
Cited alongside, same era.
Bias in word embeddings
Orestis Papakyriakopoulos, Simon Hegelich, Juan Carlos Medina Serrano, and Fabienne Marco. 2020 · 2020
Cited alongside, same era.
Investigating gender bias in language models using causal mediation analysis
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov, Sharon Qian, Daniel Nevo, Yaron Singer, and Stuart Shieber. 2020 · 2020
Cited alongside, same era.
The combined cognitive bias hypothesis in anxiety: A systematic review and meta-analysis
Chantel J Leung, Jenny Yiend, Antonella Trotta, and Tatia MC Lee. 2022 · 2022
Later among the works it cites.
Pre-trained language models for interactive decision-making
Shuang Li, Xavier Puig, Chris Paxton, Yilun Du, Clinton Wang, Linxi Fan, Tao Chen, De-An Huang, Ekin Akyürek, Anima Anandkumar, et al. 2022 · 2022
Later among the works it cites.
Large pre-trained language models contain human-like biases of what is right and wrong to do
Patrick Schramowski, Cigdem Turan, Nico Andersen, Constantin A Rothkopf, and Kristian Kersting. 2022 · 2022
Later among the works it cites.
Counterfactually augmented data and unintended bias: The case of sexism and hate speech detection
Indira Sen, Mattia Samory, Claudia Wagner, and Isabelle Augenstein. 2022 · 2022
Later among the works it cites.
A study of implicit bias in pretrained language models against people with disabilities
Pranav Narayanan Venkit, Mukund Srinath, and Shomir Wilson. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Persistent anti-muslim bias in large language models
Abubakar Abid, Maheen Farooqi, and James Zou. 2021 · 2021
Cited alongside, same era.
Judicial impartiality: Cognitive and social biases in judicial decision making
Australian Law Reform Commission et al. 2021 · 2021
Cited alongside, same era.
Detect and perturb: Neutral rewriting of biased and sensitive text via gradient-based decoding
Zexue He, Bodhisattwa Prasad Majumder, and Julian McAuley. 2021 · 2021
Cited alongside, same era.
Bias out-of-the-box: An empirical analysis of intersectional occupational biases in popular generative language models
Hannah Rose Kirk, Yennie Jun, Filippo Volpin, Haider Iqbal, Elias Benussi, Frederic Dreyer, Aleksandar Shtedritski, and Yuki Asano. 2021 · 2021
Cited alongside, same era.
Towards understanding and mitigating social biases in language models
Paul Pu Liang, Chiyu Wu, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2021 · 2021
Cited alongside, same era.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in nlp
Timo Schick, Sahana Udupa, and Hinrich Schütze. 2021 · 2021
Cited alongside, same era.
Double perturbation: On the robustness of robustness and counterfactual bias evaluation
Chong Zhang, Jieyu Zhao, Huan Zhang, Kai-Wei Chang, and Cho-Jui Hsieh. 2021 · 2021
Cited alongside, same era.
Bias beyond English: Counterfactual tests for bias in sentiment analysis in four languages
Seraphina Goldfarb-Tarrant, Adam Lopez, Roi Blanco, and Diego Marcheggiani. 2023 · 2023
Later among the works it cites.
MathPrompter: Mathematical reasoning using large language models
Shima Imani, Liang Du, and Harsh Shrivastava. 2023 · 2023
Later among the works it cites.
Instructed to bias: Instruction-tuned language models exhibit emergent cognitive bias
Itay Itzhak, Gabriel Stanovsky, Nir Rosenfeld, and Yonatan Belinkov. 2023 · 2023
Later among the works it cites.
Gender bias and stereotypes in large language models
Hadas Kotek, Rikker Dockum, and David Q Sun. 2023 · 2023
Later among the works it cites.
Prompted LLMs as chatbot modules for long open-domain conversation
Gibbeum Lee, Volker Hartmann, Jongho Park, Dimitris Papailiopoulos, and Kangwook Lee. 2023 · 2023
Later among the works it cites.
Mind the biases: Quantifying cognitive biases in language model prompting
Ruixi Lin and Hwee Tou Ng. 2023 · 2023
Later among the works it cites.
Supporting human-ai collaboration in auditing llms with llms
Charvi Rastogi, Marco Tulio Ribeiro, Nicholas King, Harsha Nori, and Saleema Amershi. 2023 · 2023
Later among the works it cites.
Challenging the appearance of machine intelligence: Cognitive bias in llms
Alaina N Talboy and Elizabeth Fuller. 2023 · 2023
Later among the works it cites.
Fairpy: A toolkit for evaluation of social biases and their mitigation in large language models
Hrishikesh Viswanath and Tianyi Zhang. 2023 · 2023
Later among the works it cites.
Zero-shot cross-lingual summarization via large language models
Jiaan Wang, Yunlong Liang, Fandong Meng, Beiqi Zou, Zhixu Li, Jianfeng Qu, and Jie Zhou. 2023 · 2023
Later among the works it cites.