Fetching the paper…
Reading the bibliography…
Drawing parallels between human cognition and artificial intelligence, we explored how large language models (LLMs) internalize identities imposed by targeted prompts.
The nature of prejudice
Gordon Willard Allport, Kenneth Clark, and Thomas Pettigrew · 1954
Earlier work this paper cites.
Experiments in intergroup discrimination
Henri Tajfel · 1970
Earlier work this paper cites.
Social categorization and intergroup behaviour
Henri Tajfel, Michael G Billig, Robert P Bundy, and Claude Flament · 1971
Earlier work this paper cites.
Social identity and intergroup behaviour
Henri Tajfel · 1974
Earlier work this paper cites.
Social comparison and social identity: Some prospects for intergroup behaviour
John C Turner · 1975
Earlier work this paper cites.
An integrative theory of intergroup conflict
Henri Tajfel, John C Turner, William G Austin, and Stephen Worchel · 1979
Earlier work this paper cites.
Attitude polarization: Effects of group membership
Diane Mackie and Joel Cooper · 1984
Earlier work this paper cites.
Social identification effects in group polarization
Diane M Mackie · 1986
Earlier work this paper cites.
Social identity theory: Constructive and critical advances
Dominic Ed Abrams and Michael A Hogg · 1990
Earlier work this paper cites.
Knowing what to think by knowing who you are: Self-categorization and the nature of norm formation, conformity and group polarization
Dominic Abrams, Margaret Wetherell, Sandra Cochrane, Michael A Hogg, and John C Turner · 1990
Earlier work this paper cites.
Polarized norms and social frames of reference: A test of the self-categorization theory of group polarization
Michael A Hogg, John C Turner, and Barbara Davidson · 1990
Earlier work this paper cites.
Social discrimination and tolerance in intergroup relations: Reactions to intergroup difference
Amelie Mummendey and Michael Wenzel · 1999
Earlier work this paper cites.
The ambivalence toward men inventory: Differentiating hostile and benevolent beliefs about men
Peter Glick and Susan T Fiske · 1999
Earlier work this paper cites.
Negativity bias, negativity dominance, and contagion
Paul Rozin and Edward B Royzman · 2001
Earlier work this paper cites.
Social identity
Michael A Hogg · 2003
Earlier work this paper cites.
Psychological motives and political orientation–the left, the right, and the rigid: comment on jost et al
Jeff Greenberg and Eva Jonas · 2003
Earlier work this paper cites.
Political psychology: Key readings
John T Jost and Jim Sidanius · 2004
Earlier work this paper cites.
Dynamic remodeling of in-group bias during the 2008 presidential election
David G Rand, Thomas Pfeiffer, Anna Dreber, Rachel W Sheketoff, Nils C Wernerfelt, and Yochai Benkler · 2009
Earlier work this paper cites.
Intergroup bias
John F Dovidio and Samuel L Gaertner · 2010
Earlier work this paper cites.
The social psychology of intergroup relations
Nicole Tausch, Katharina Schmid, and Miles Hewstone · 2010
Earlier work this paper cites.
A cross-cutting calm: How social sorting drives affective polarization
Lilliana Mason · 2016
Earlier work this paper cites.
Social identity theory
Michael A Hogg · 2016
Earlier work this paper cites.
From primed concepts to action: A meta-analysis of the behavioral effects of incidentally presented words
Evan Weingarten, Qijia Chen, Maxwell McAdams, Jessica Yi, Justin Hepler, and Dolores Albarracín · 2016
Earlier work this paper cites.
One tribe to bind them all: How our social group attachments strengthen partisanship
Lilliana Mason and Julie Wronski · 2018
Earlier work this paper cites.
The ambivalent sexism inventory: Differentiating hostile and benevolent sexism
Peter Glick and Susan T Fiske · 2018
Earlier work this paper cites.
The spread of low-credibility content by social bots
Chengcheng Shao, Giovanni Luca Ciampaglia, Onur Varol, Kai-Cheng Yang, Alessandro Flammini, and Filippo Menczer · 2018
Cited alongside, same era.
Compassionate democrats and tough republicans: How ideology shapes partisan stereotypes
Scott Clifford · 2020
Cited alongside, same era.
Analyzing the impact of filter bubbles on social network polarization
Uthsav Chitra and Christopher Musco · 2020
Cited alongside, same era.
Mitigating political bias in language models through reinforced calibration
Ruibo Liu, Chenyan Jia, Jason Wei, Guangxuan Xu, Lili Wang, and Soroush Vosoughi · 2021
Cited alongside, same era.
Communitylm: Probing partisan worldviews from language models
Hang Jiang, Doug Beeferman, Brandon Roy, and Deb Roy · 2022
Cited alongside, same era.
Moral mimicry: Large language models produce moral rationalizations tailored to political identity
Reducing negative effects of the biases of language models in zero-shot setting
Xiaosu Wang, Yun Xiong, Beichen Kang, Yao Zhang, Philip S Yu, and Yangyong Zhu · 2023
Later among the works it cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gabriel Simmons · 2022
Cited alongside, same era.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Cited alongside, same era.
Jochen Hartmann, Jasper Schwenzow, and Maximilian Witte · 2023
Cited alongside, same era.
Is chat gpt biased against conservatives? an empirical study
Robert W McGee · 2023
Cited alongside, same era.
Whose opinions do language models reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee, Percy Liang, and Tatsunori Hashimoto · 2023
Cited alongside, same era.
“kelly is a warm person, joseph is a role model”: Gender biases in llm-generated reference letters
Yixin Wan, George Pu, Jiao Sun, Aparna Garimella, Kai-Wei Chang, and Nanyun Peng · 2023
Cited alongside, same era.
Large language models propagate race-based medicine
Jesutofunmi A Omiye, Jenna C Lester, Simon Spichak, Veronica Rotemberg, and Roxana Daneshjou · 2023
Cited alongside, same era.
What is the difference between a democrat and a republican?, 2023
Merriam-Webster · 2023
Later among the works it cites.
Human heuristics for ai-generated language are flawed
Maurice Jakesch, Jeffrey T Hancock, and Mor Naaman · 2023
Later among the works it cites.
Social bot detection in the age of chatgpt: Challenges and opportunities
Emilio Ferrara · 2023
Later among the works it cites.
Beyond digital" echo chambers": The role of viewpoint diversity in political discussion
Rishav Hada, Amir Ebrahimi Fard, Sarah Shugars, Federico Bianchi, Patricia Rossini, Dirk Hovy, Rebekah Tromble, and Nava Tintarev · 2023
Later among the works it cites.
Testing theory of mind in large language models and humans
James WA Strachan, Dalila Albergo, Giulia Borghini, Oriana Pansardi, Eugenio Scaliti, Saurabh Gupta, Krati Saxena, Alessandro Rufo, Stefano Panzeri, Guido Manzi, et al · 2024
Closest in time.
The benefits, risks and bounds of personalizing the alignment of large language models to individuals
Hannah Rose Kirk, Bertie Vidgen, Paul Röttger, and Scott A Hale · 2024
Closest in time.
Risk and prosocial behavioural cues elicit human-like response patterns from ai chatbots
Yukun Zhao, Zhen Huang, Martin Seligman, and Kaiping Peng · 2024
Closest in time.
Deception abilities emerged in large language models
Thilo Hagendorff · 2024
Closest in time.
Bias of ai-generated content: an examination of news produced by large language models
Xiao Fang, Shangkun Che, Minjia Mao, Hongzhe Zhang, Ming Zhao, and Xiaohang Zhao · 2024
Closest in time.
Evaluating large language model biases in persona-steered generation
Andy Liu, Mona Diab, and Daniel Fried · 2024
Closest in time.
A survey on evaluation of large language models
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, et al · 2024
Closest in time.
Cognitive bias in high-stakes decision-making with llms
Jessica Echterhoff, Yao Liu, Abeer Alessa, Julian McAuley, and Zexue He · 2024
Closest in time.
Thinking fair and slow: On the efficacy of structured prompts for debiasing language models
Shaz Furniturewala, Surgan Jandial, Abhinav Java, Pragyan Banerjee, Simra Shahid, Sumit Bhatia, and Kokil Jaidka · 2024
Closest in time.
More human than human: measuring chatgpt political bias
Fabio Motoki, Valdemar Pinho Neto, and Victor Rodrigues · 2024
Closest in time.
Llmrec: Large language models with graph augmentation for recommendation
Wei Wei, Xubin Ren, Jiabin Tang, Qinyong Wang, Lixin Su, Suqi Cheng, Junfeng Wang, Dawei Yin, and Chao Huang · 2024
Closest in time.
Once: Boosting content-based recommendation with both open-and closed-source large language models
Qijiong Liu, Nuo Chen, Tetsuya Sakai, and Xiao-Ming Wu · 2024
Closest in time.
Evaluating the persuasive influence of political microtargeting with large language models
Kobi Hackenburg and Helen Margetts · 2024
Closest in time.
Ai models collapse when trained on recursively generated data
Ilia Shumailov, Zakhar Shumaylov, Yiren Zhao, Nicolas Papernot, Ross Anderson, and Yarin Gal · 2024
Closest in time.
Temporal blind spots in large language models
Jonas Wallat, Adam Jatowt, and Avishek Anand · 2024
Closest in time.
Temporalmed: Advancing medical dialogues with time-aware responses in large language models
Yuyan Chen, Jin Zhao, Zhihao Wen, Zhixu Li, and Yanghua Xiao · 2024
Closest in time.
The art of refusal: A survey of abstention in large language models
Bingbing Wen, Jihan Yao, Shangbin Feng, Chenjun Xu, Yulia Tsvetkov, Bill Howe, and Lucy Lu Wang · 2024
Closest in time.