Fetching the paper…
Reading the bibliography…
With large language models (LLMs) like GPT-4 appearing to behave increasingly human-like in text-based interactions, it has become popular to attempt to evaluate personality traits of LLMs using questionnaires originally developed for humans.
Construct validity in psychological tests
Lee J Cronbach and Paul E Meehl · 1955
Earlier work this paper cites.
The proof and measurement of association between two things
Charles Spearman · 1961
Earlier work this paper cites.
Structural equations with latent variables
Kenneth A Bollen · 1989
Earlier work this paper cites.
Big five inventory
Oliver P John, Eileen M Donahue, and Robert L Kentle · 1991
Earlier work this paper cites.
Measurement invariance, factor analysis and factorial invariance
William Meredith · 1993
Earlier work this paper cites.
Cutoff criteria for fit indexes in covariance structure analysis: Conventional criteria versus new alternatives
Li-tze Hu and Peter M. Bentler · 1999
Earlier work this paper cites.
Test theory: A unified treatment. nueva york, 1999
R McDonald · 1999
Earlier work this paper cites.
True score theory: The traditional method
Howard Wainer and David Thissen · 2001
Earlier work this paper cites.
Methods of multivariate analysis. 2002
AC Rencher · 2002
Earlier work this paper cites.
Cronbach’s α \alpha , Revelle’s β \beta , and McDonald’s ω \omega H: Their relations with each other and two alternative conceptualizations of reliability
Richard E Zinbarg, William Revelle, Iftah Yovel, and Wen Li · 2005
Earlier work this paper cites.
Standards for educational and psychological testing, 2014
American Educational Research Association, American Psychological Association, and National Council on Measurement in Education · 2014
Earlier work this paper cites.
Confirmatory factor analysis for applied research
Timothy A Brown · 2015
Earlier work this paper cites.
Die deutsche Version des Big Five Inventory 2 (BFI-2)
Daniel Danner, Beatrice Rammstedt, Matthias Bluemke, Lisa Treiber, Sabrina Berres, Christopher J. Soto, and Oliver P. John · 2016
Earlier work this paper cites.
The next big five inventory (bfi-2): Developing and assessing a hierarchical model with 15 facets to enhance bandwidth, fidelity, and predictive power
Christopher J Soto and Oliver P John · 2017
Cited alongside, same era.
Sanity checks for saliency maps
Julius Adebayo, Justin Gilmer, Michael Muelly, Ian Goodfellow, Moritz Hardt, and Been Kim · 2018
Cited alongside, same era.
Personalizing dialogue agents: I have a dog, do you have pets too?
Saizheng Zhang, Emily Dinan, Jack Urbanek, Arthur Szlam, Douwe Kiela, and Jason Weston · 2018
Cited alongside, same era.
Is gpt-3 a psychopath? evaluating large language models from a psychological perspective
Xingxuan Li, Yutong Li, Linlin Liu, Lidong Bing, and Shafiq Joty · 2022
Cited alongside, same era.
Shengyu Mao, Ningyu Zhang, Xiaohan Wang, Mengru Wang, Yunzhi Yao, Yong Jiang, Pengjun Xie, Fei Huang, and Huajun Chen · 2023
Closest in time.
Do llms possess a personality? making the mbti test an amazing evaluation for large language models
Keyu Pan and Yawen Zeng · 2023
Closest in time.
Who is chatgpt? benchmarking llms’ psychological portrayal using psychobench
Jen-tse Huang, Wenxuan Wang, Eric John Li, Man Ho Lam, Shujie Ren, Youliang Yuan, Wenxiang Jiao, Zhaopeng Tu, and Michael R Lyu · 2023
Closest in time.
Can ai have a personality?
Umarpreet Singh and Parham Aarabhi · 2023
Closest in time.
On the humanity of conversational ai: Evaluating the psychological portrayal of llms
Jen-tse Huang, Wenxuan Wang, Eric John Li, Man Ho LAM, Shujie Ren, Youliang Yuan, Wenxiang Jiao, Zhaopeng Tu, and Michael Lyu · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Saketh Reddy Karra, Son The Nguyen, and Theja Tulabandhula · 2022
Cited alongside, same era.
Factor structure, gender invariance, measurement properties, and short forms of the spanish adaptation of the big five inventory-2
David Gallardo-Pujol, Víctor Rouco, Anna Cortijos-Bernabeu, Luis Oceja, Christopher J Soto, and Oliver P John · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
Personality traits in large language models
Mustafa Safdari, Greg Serapio-García, Clément Crepy, Stephen Fitz, Peter Romero, Luning Sun, Marwa Abdulhai, Aleksandra Faust, and Maja Matarić · 2023
Cited alongside, same era.
Yang Lu, Jordan Yu, and Shou-Hsuan Stephen Huang · 2023
Cited alongside, same era.
Evaluating and inducing personality in pre-trained language models
Guangyuan Jiang, Manjie Xu, Song-Chun Zhu, Wenjuan Han, Chi Zhang, and Yixin Zhu · 2023
Cited alongside, same era.
Hang Jiang, Xiajie Zhang, Xubo Cao, Jad Kabbara, and Deb Roy · 2023
Cited alongside, same era.
Ai psychometrics: Assessing the psychological profiles of large language models through psychometric inventories
Max Pellert, Clemens M Lechner, Claudia Wagner, Beatrice Rammstedt, and Markus Strohmaier · 2023
Cited alongside, same era.
Closest in time.
Manipulating the perceived personality traits of language models
Graham Caron and Shashank Srivastava · 2023
Closest in time.
Administering ipip measures, with a 50-item sample questionnaire
International Personality Item Pool · 2023
Closest in time.
Using cognitive psychology to understand gpt-3
Marcel Binz and Eric Schulz · 2023
Closest in time.
Whose opinions do language models reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee, Percy Liang, and Tatsunori Hashimoto · 2023
Closest in time.
A turing test of whether ai chatbots are behaviorally similar to humans
Qiaozhu Mei, Yutong Xie, Walter Yuan, and Matthew O Jackson · 2024
Closest in time.
On the humanity of conversational AI: Evaluating the psychological portrayal of LLMs
Anonymous · 2024
Closest in time.
Limited ability of llms to simulate human psychological behaviours: a psychometric analysis
Nikolay B Petrov, Gregory Serapio-García, and Jason Rentfrow · 2024
Closest in time.