Fetching the paper…
Reading the bibliography…
As large language models (LLMs) are increasingly used in human-centered tasks, assessing their psychological traits is crucial for understanding their social impact and ensuring trustworthy AI alignment.
The Myers-Briggs Type Indicator
IB Myers. 1962 · 1962
Earlier work this paper cites.
Myers-Briggs type indicator
Katharine C Briggs. 1976 · 1976
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff. 1978 · 1978
Earlier work this paper cites.
beliefs about beliefs: Representation and constraining function of wrong beliefs in young children’s understanding of deception
Heinz Wimmer and Josef Perner. 1983 · 1983
Earlier work this paper cites.
Does the autistic child have a “theory of mind”?
Simon Baron-Cohen, Alan M Leslie, and Uta Frith. 1985 · 1985
Earlier work this paper cites.
A guide to the development and use of the Myers-Briggs type indicator: Manual
Isabel Briggs Myers. 1985 · 1985
Earlier work this paper cites.
“John thinks that Mary thinks that…” attribution of second-order beliefs by 5-to 10-year-old children
Josef Perner and Heinz Wimmer. 1985 · 1985
Earlier work this paper cites.
Three-year-olds’ difficulty with false belief: The case for a conceptual deficit
Josef Perner, Susan R Leekam, and Heinz Wimmer. 1987 · 1987
Earlier work this paper cites.
Development and validation of brief measures of positive and negative affect: the PANAS scales
David Watson, Lee Anna Clark, and Auke Tellegen. 1988 · 1988
Earlier work this paper cites.
Big five inventory
Oliver P John, Eileen M Donahue, and Robert L Kentle. 1991 · 1991
Earlier work this paper cites.
The aggression questionnaire
Arnold H Buss and Mark Perry. 1992 · 1992
Earlier work this paper cites.
An advanced test of theory of mind: Understanding of story characters’ thoughts and feelings by able autistic, mentally handicapped, and normal children and adults
Francesca GE Happé. 1994 · 1994
Earlier work this paper cites.
Myers-Briggs Type Indicator (MBTI): Some psychometric limitations
Gregory J. Boyle. 1995 · 1995
Earlier work this paper cites.
Theory-of-mind deficits and causal attributions
Peter Kinderman, Robin Dunbar, and Richard P Bentall. 1998 · 1998
Earlier work this paper cites.
Recognition of faux pas by normally developing children and children with Asperger syndrome or high-functioning autism
Simon Baron-Cohen, Michelle O’riordan, Valerie Stone, Rosie Jones, and Kate Plaisted. 1999 · 1999
Earlier work this paper cites.
A broad-bandwidth, public domain, personality inventory measuring the lower-level facets of several five-factor models
Lewis R. Goldberg. 1999 · 1999
Earlier work this paper cites.
The Big-Five trait taxonomy: History, measurement, and theoretical perspectives
Oliver P John, Sanjay Srivastava, et al · 1999
Earlier work this paper cites.
The relationship between psychometric and self-estimated intelligence, creativity, personality and academic achievement
Adrian Furnham, Jane Zhang, and Tomas Chamorro-Premuzic. 2005 · 2005
Earlier work this paper cites.
Ascertaining the validity of individual protocols from web-based personality inventories
John A Johnson. 2005 · 2005
Earlier work this paper cites.
The strange stories test: A replication study of children and adolescents with Asperger syndrome
Nils Kaland, Annette Møller-Nielsen, Lars Smith, Erik Lykke Mortensen, Kirsten Callesen, and Dorte Gottlieb. 2005 · 2005
Earlier work this paper cites.
Assessing the Convergent and Discriminant Validity of Goldberg’s International Personality Item Pool
Lim Beng-Chong and E. Ployhart Robert. 2006 · 2006
Earlier work this paper cites.
The international personality item pool and the future of public-domain personality measures
Lewis R Goldberg, John A Johnson, Herbert W Eber, Robert Hogan, Michael C Ashton, C Robert Cloninger, and Harrison G Gough. 2006 · 2006
Earlier work this paper cites.
The generalizability of the buss–perry aggression questionnaire
József Gerevich, Erika Bácskai, and PÁL Czobor. 2007 · 2007
Earlier work this paper cites.
The HEXACO–60: A short measure of the major dimensions of personality
Michael C Ashton and Kibeom Lee. 2009 · 2009
Earlier work this paper cites.
Cualidades paramétricas del cuestionario de agresión (AQ) de Buss y Perry en estudiantes universitarios de la ciudad de Medellín (Colombia)
Diego Castrillón M, Paola Ortiz T, and Fernando Vieco G. 2009 · 2009
Earlier work this paper cites.
Theory of mind: An overview and behavioral perspective
Henry D Schlinger. 2009 · 2009
Earlier work this paper cites.
The American Idol Effect: Are students good judges of their creativity across domains?
James C Kaufman, Michelle L Evans, and John Baer. 2010 · 2010
Earlier work this paper cites.
The Positive and Negative Affect Schedule: Psychometric Properties of the Korean Version
Young-Jin Lim, Bum-Hee Yu, Doh-Kwan Kim, and Ji-Hae Kim. 2010 · 2010
Earlier work this paper cites.
It doesn’t hurt to ask… But sometimes it hurts to believe: Polish students’ creative self-efficacy and its predictors
Maciej Karwowski. 2011 · 2011
Earlier work this paper cites.
Short assessment of the Big Five: Robust across survey methods except telephone interviewing
Frieder R Lang, Dennis John, Oliver Lüdtke, Jürgen Schupp, and Gert G Wagner. 2011 · 2011
Earlier work this paper cites.
The Buss-Perry Aggression Questionnaire: construct validity and gender invarianceamong Argentinean adolescents
Cecilia Reyna, Anahi Sanchez, Maria Gabriela Lello Ivacevich, and Silvina Brussino. 2011 · 2011
Earlier work this paper cites.
Reliability and Concurrent Validity of the International Personality item Pool (IPIP) Big-five Factor Markers in Nigeria
Akinsulore A., Fatoye O., Awaa O., Aloba Olutayo, Mapayi B., and Ibigbami O. 2012 · 2012
Earlier work this paper cites.
Structural validity and reliability of the Positive and Negative Affect Schedule (PANAS): Evidence from a large Brazilian community sample
Hudson W. de Carvalho, Sérgio B. Andreoli, Diogo R. Lara, Christopher J. Patrick, Maria Inês Quintana, Rodrigo A. Bressan, Marcelo F. de Melo, Jair de J. Mari, and Miguel R. Jorge. 2013 · 2012
Earlier work this paper cites.
Evaluarea personalită
Rusu S., P. Maricu · 2012
Earlier work this paper cites.
The Dark Triad of personality: A 10 year review
Adrian Furnham, Steven C Richards, and Delroy L Paulhus. 2013 · 2013
Earlier work this paper cites.
The brief aggression questionnaire: psychometric and behavioral evidence for an efficient measure of trait aggression
Gregory D. Webster, C. Nathan DeWall, Richard S. Pond, Jr., Timothy Deckman, Peter K. Jonason, Bonnie M. Le, Austin Lee Nichols, Tatiana Orozco Schember, Laura C. Crysel, Benjamin S. Crosier, C. Veronica Smith, E. Layne Paddock, John B. Nezlek, Lee A. Kirkpatrick, Angela D. Bryan, and Renée J. Bator. 2013 · 2013
Earlier work this paper cites.
Measuring thirty facets of the Five Factor Model with a 120-item public domain inventory: Development of the IPIP-NEO-120
John A Johnson. 2014 · 2014
Earlier work this paper cites.
Introducing the short dark triad (SD3) a brief measure of dark personality traits
Daniel N Jones and Delroy L Paulhus. 2014 · 2014
Earlier work this paper cites.
Differences in within- and between-person factor structure of positive and negative affect: Analysis of two intensive measurement studies using multilevel structural equation modeling
Jonathan Rush and Scott M. Hofer. 2014 · 2014
Earlier work this paper cites.
Psychometric Properties of the International Personality Item Pool Big-Five Personality Questionnaire for the Greek population
Ypofanti M., Zisi V., Zourbanos N., Mouchtouri B., Tzanne P., Theodorakis Y., and Lyrakos G. 2015 · 2015
Earlier work this paper cites.
Psychometric properties of the Positive and Negative Affect Schedule (PANAS) in a heterogeneous sample of substance users
Kelly Serafini, Bo Malin-Mayor, Charla Nich, Karen Hunkele, and Kathleen M. Carroll. 2016 · 2015
Earlier work this paper cites.
The lazy mindreader. A humanities perspective on mindreading and multiple-order intentionality
Max J van Duijn. 2016 · 2016
Earlier work this paper cites.
Using Item Response Theory to Develop a 60-Item Representation of the NEO PI–R Using the International Personality Item Pool: Development of the IPIP–NEO–60
Jessica L Maples-Keller, Rachel L Williamson, Chelsea E Sleep, Nathan T Carter, W Keith Campbell, and Joshua D Miller. 2019 · 2017
Earlier work this paper cites.
Measuring creative self-efficacy and creative personal identity
Maciej Karwowski, Izabella Lebuda, and Ewa Wiśniewska. 2018 · 2018
Earlier work this paper cites.
Assessing the Structure of the Five Factor Model of Personality (IPIP-NEO-120) in the Public Domain
Petri J. Kajonius and John A. Johnson. 2019 · 2019
Earlier work this paper cites.
Revisiting the evaluation of theory of mind through question answering. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . 5872–5877
Matthew Le, Y-Lan Boureau, and Maximilian Nickel. 2019 · 2019
Earlier work this paper cites.
Examining the crosscultural validity of the positive affect and negative affect schedule between an Asian (Singaporean) sample and a Western (American) sample
Sean T. H. Lee, Andree Hartanto, Jose C. Yong, Brandon Koh, and Angela K.y. Leung. 2019 · 2019
Earlier work this paper cites.
Meta-analytic investigations of the HEXACO Personality Inventory (-Revised)
Morten Moshagen, Isabel Thielmann, Benjamin E Hilbig, and Ingo Zettler. 2019 · 2019
Earlier work this paper cites.
Socialiqa: Commonsense reasoning about social interactions
Maarten Sap, Hannah Rashkin, Derek Chen, Ronan LeBras, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
CONTENT VALIDITY AND RELIABILITY OF BUSS AND PERRY AGRESSIVE QUESTIONNAIRE (BPAQ) INVENTORY
Ibrahim Abd Ghani and Norsayyadatina Che Rozubi. 2020 · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom B Brown. 2020 · 2020
Earlier work this paper cites.
Multigroup multilevel structure of the child and parent versions of the Positive and Negative Affect Schedule (PANAS) in adolescents with and without ADHD
Hana-May Eadeh, Rosanna Breaux, Joshua M. Langberg, Molly A. Nikolas, and Stephen P. Becker. 2020 · 2020
Earlier work this paper cites.
Questionário de Agressão de Buss-Perry versão reduzida (QA-R): análises estruturais
Tomaz Paiva Tamyres, Eduardo Pimentel Carlos, de Sousa Bezerra de Menezes Thaís, Costa A.C.R., das Graças Carvalho Costa Dinara, and H. Vasconcelos M. 2020 · 2020
Cited alongside, same era.
Buss-Perry Aggression Questionnaire: Factor Structure and Measurement Invariance Among Portuguese Male Perpetrators of Intimate Partner Violence
Olga Cunha, Manuela Peixoto, Ana Rita Cruz, and Rui Abrunhosa Gonçalves. 2021 · 2021
Cited alongside, same era.
Aggression Amongst Outpatients With Schizophrenia and Related Psychoses in a Tertiary Mental Health Institution
Anitha Jeyagurunathan, Jue Hua Lau, Edimansyah Abdin, Saleha Shafie, Sherilyn Chang, Ellaisha Samari, Laxman Cetty, Ker-Chiah Wei, Yee Ming Mok, Charmaine Tang, Swapna Verma, Siow Ann Chong, and Mythily Subramaniam. 2022 · 2021
Cited alongside, same era.
An Initial Analysis of Reliability and Validity of a Personality Instrument Using the Rasch Measurement Model
Mohamed N., Sulaiman W., Halim F., and Saidfudin Masodi Mohd. 2021 · 2021
Cited alongside, same era.
Identifying and manipulating the personality traits of language models
Evaluating Language Model Agency Through Negotiations. In The Twelfth International Conference on Learning Representations, ICLR 2024, Vienna, Austria, May 7-11, 2024 . OpenReview.net
Tim R. Davidson, Veniamin Veselovsky, Michal Kosinski, and Robert West. 2024 · 2024
Later among the works it cites.
Gtbench: Uncovering the strategic reasoning limitations of llms via game-theoretic evaluations
Jinhao Duan, Renming Zhang, James Diffenderfer, Bhavya Kailkhura, Lichao Sun, Elias Stengel-Eskin, Mohit Bansal, Tianlong Chen, and Kaidi Xu. 2024 · 2024
Later among the works it cites.
Can large language models serve as rational players in game theory? a systematic analysis. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 38. 17960–17967
Caoyun Fan, Jindou Chen, Yaohui Jin, and Hao He. 2024 · 2024
Later among the works it cites.
Large Language Models as Simulated Economic Agents: What Can We Learn from Homo Silicus?. In Proceedings of the 25th ACM Conference on Economics and Computation, EC 2024, New Haven, CT, USA, July 8-11, 2024 , Dirk Bergemann, Robert Kleinberg, and Daniela Sabán (Eds.). ACM, 614–615
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Graham Caron and Shashank Srivastava. 2022 · 2022
Cited alongside, same era.
PaLM: Scaling Language Modeling with Pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Cited alongside, same era.
Examining the within- and between-person structure of a short form of the positive and negative affect schedule: A multilevel and dynamic approach
Eric M. Cooke, Noémi K. Schuurman, and Yao Zheng. 2022 · 2022
Cited alongside, same era.
Estimating the Personality of White-Box Language Models
Saketh Reddy Karra, Son The Nguyen, and Theja Tulabandhula. 2022 · 2022
Cited alongside, same era.
Xingxuan Li, Yutong Li, Shafiq Joty, Linlin Liu, Fei Huang, Lin Qiu, and Lidong Bing. 2022 · 2022
Cited alongside, same era.
Who is GPT-3? An exploration of personality, values and demographics
Marilù Miotto, Nicola Rossberg, and Bennett Kleinberg. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
Development and Validation Regarding the Lithuanian Version of the Positive and Negative Affect Schedule (PANAS-X)
Karolina Petraškaitė and Neringa Grigutytė. 2022 · 2022
Cited alongside, same era.
Apostolos Filippas, John J. Horton, and Benjamin S. Manning. 2024 · 2024
Later among the works it cites.
Ivar Frisch and Mario Giulianelli. 2024 · 2024
Later among the works it cites.
Human-like Affective Cognition in Foundation Models
Kanishk Gandhi, Zoe Lynch, Jan-Philipp Fränken, Kayla Patterson, Sharon Wambu, Tobias Gerstenberg, Desmond C. Ong, and Noah D. Goodman. 2024 · 2024
Later among the works it cites.
AFSPP: Agent Framework for Shaping Preference and Personality with Large Language Models
Zihong He and Changwang Zhang. 2024a · 2024
Later among the works it cites.
AFSPP: Agent Framework for Shaping Preference and Personality with Large Language Models
Zihong He and Changwang Zhang. 2024b · 2024
Later among the works it cites.
Designing LLM-Agents with Personalities: A Psychometric Approach
Muhua Huang, Xijuan Zhang, Christopher Soto, and James Evans. 2024 · 2024
Later among the works it cites.
Perceptions to Beliefs: Exploring Precursory Inferences for Theory of Mind in Large Language Models. In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, EMNLP 2024, Miami, FL, USA, November 12-16, 2024 , Yaser Al-Onaizan, Mohit Bansal, and Yun-Nung Chen (Eds.). Association for Computational Linguistics, 19794–19809
Chani Jung, Dongkwan Kim, Jiho Jin, Jiseon Kim, Yeon Seonwoo, Yejin Choi, Alice Oh, and Hyunwoo Kim. 2024 · 2024
Later among the works it cites.
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
Elise Karinshak, Amanda Hu, Kewen Kong, Vishwanatha Rao, Jingren Wang, Jindong Wang, and Yi Zeng. 2024 · 2024
Later among the works it cites.
Exploring the frontiers of llms in psychological applications: A comprehensive review
Luoma Ke, Song Tong, Peng Cheng, and Kaiping Peng. 2024 · 2024
Later among the works it cites.
Driving generative agents with their personality
Lawrence J Klinkert, Stephanie Buongiorno, and Corey Clark. 2024 · 2024
Later among the works it cites.
Lucio La Cava and Andrea Tagarelli. 2024 · 2024
Later among the works it cites.
Seungbeen Lee, Seungwon Lim, Seungju Han, Giyeong Oh, Hyungjoo Chae, Jiwan Chung, Minju Kim, Beong-woo Kwak, Yeonsoo Lee, Dongha Lee, Jinyoung Yeo, and Youngjae Yu. 2024 · 2024
Later among the works it cites.
Can LLMs Mimic Human-Like Mental Accounting and Behavioral Biases?. In Proceedings of the 25th ACM Conference on Economics and Computation, EC 2024, New Haven, CT, USA, July 8-11, 2024 . ACM, 581
Yan Leng. 2024 · 2024
Later among the works it cites.
BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data
Wenkai Li, Jiarui Liu, Andy Liu, Xuhui Zhou, Mona Diab, and Maarten Sap. 2024c · 2024
Later among the works it cites.
Evaluating psychological safety of large language models. In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing . 1826–1843
Xingxuan Li, Yutong Li, Lin Qiu, Shafiq Joty, and Lidong Bing. 2024b · 2024
Later among the works it cites.
Quantifying AI Psychology: A Psychometrics Benchmark for Large Language Models
Yuan Li, Yue Huang, Hongyi Wang, Xiangliang Zhang, James Zou, and Lichao Sun. 2024a · 2024
Later among the works it cites.
Dynamic Generation of Personalities with Large Language Models
Jianzhi Liu, Hexiang Gu, Tianyu Zheng, Liuyu Xiang, Huijia Wu, Jie Fu, and Zhaofeng He. 2024 · 2024
Later among the works it cites.
Editing Personality For Large Language Models. In CCF International Conference on Natural Language Processing and Chinese Computing . Springer, 241–254
Shengyu Mao, Xiaohan Wang, Mengru Wang, Yong Jiang, Pengjun Xie, Fei Huang, and Ningyu Zhang. 2024 · 2024
Later among the works it cites.
LLMs with Personalities in Multi-issue Negotiation Games
Sean Noh and Ho-Chun Herbert Chang. 2024 · 2024
Later among the works it cites.
Balrog: Benchmarking agentic llm and vlm reasoning on games
Davide Paglieri, Bartłomiej Cupiał, Samuel Coward, Ulyana Piterbarg, Maciej Wolczyk, Akbir Khan, Eduardo Pignatelli, Łukasz Kuciński, Lerrel Pinto, Rob Fergus, et al · 2024
Later among the works it cites.
Diminished diversity-of-thought in a standard large language model
Peter S Park, Philipp Schoenegger, and Chongyang Zhu. 2024 · 2024
Later among the works it cites.
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis
Nikolay B Petrov, Gregory Serapio-García, and Jason Rentfrow. 2024 · 2024
Later among the works it cites.
EmoBench: Evaluating the Emotional Intelligence of Large Language Models. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2024, Bangkok, Thailand, August 11-16, 2024 , Lun-Wei Ku, Andre Martins, and Vivek Srikumar (Eds.). Association for Computational Linguistics, 5986–6004
Sahand Sabour, Siyang Liu, Zheyuan Zhang, June M. Liu, Jinfeng Zhou, Alvionna S. Sunaryo, Tatia M. C. Lee, Rada Mihalcea, and Minlie Huang. 2024 · 2024
Later among the works it cites.
Identifying Multiple Personalities in Large Language Models with External Evaluation
Xiaoyang Song, Yuta Adachi, Jessie Feng, Mouwei Lin, Linhao Yu, Frank Li, Akshat Gupta, Gopala Anumanchipalli, and Simerjot Kaur. 2024 · 2024
Later among the works it cites.
The Effects of Embodiment and Personality Expression on Learning in LLM-based Educational Agents
Sinan Sonlu, Bennie Bendiksen, Funda Durupinar, and Uğur Güdükbay. 2024 · 2024
Later among the works it cites.
LLMs Simulate Big5 Personality Traits: Further Evidence. In Proceedings of the 1st Workshop on Personalization of Generative AI Systems (PERSONALIZE 2024) . 83–87
Aleksandra Sorokovikova, Sharwin Rezagholi, Natalia Fedorova, and Ivan P. Yamshchikov. 2024 · 2024
Later among the works it cites.
Secret Keepers: The Impact of LLMs on Linguistic Markers of Personal Traits
Zhivar Sourati, Meltem Ozcan, Colin McDaniel, Alireza Ziabari, Nuan Wen, Ala Tak, Fred Morstatter, and Morteza Dehghani. 2024 · 2024
Later among the works it cites.
Large language models could change the future of behavioral healthcare: a proposal for responsible development and evaluation
Elizabeth C Stade, Shannon Wiltsey Stirman, Lyle H Ungar, Cody L Boland, H Andrew Schwartz, David B Yaden, João Sedoc, Robert J DeRubeis, Robb Willer, and Johannes C Eichstaedt. 2024 · 2024
Later among the works it cites.
The Personification of ChatGPT (GPT-4)—Understanding Its Personality and Adaptability
Leandro Stöckli, Luca Joho, Felix Lehner, and Thomas Hanne. 2024 · 2024
Later among the works it cites.
Testing theory of mind in large language models and humans
James WA Strachan, Dalila Albergo, Giulia Borghini, Oriana Pansardi, Eugenio Scaliti, Saurabh Gupta, Krati Saxena, Alessandro Rufo, Stefano Panzeri, Guido Manzi, et al · 2024
Later among the works it cites.
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models
Fiona Anting Tan, Gerard Christopher Yeo, Kokil Jaidka, Fanyou Wu, Weijie Xu, Vinija Jain, Aman Chadha, Yang Liu, and See-Kiong Ng. 2024 · 2024
Later among the works it cites.
Do LLMs Exhibit Human-like Response Biases? A Case Study in Survey Design
Lindia Tjuatja, Valerie Chen, Tongshuang Wu, Ameet Talwalkwar, and Graham Neubig. 2024 · 2024
Later among the works it cites.
Oguzhan Topsakal, Colby Jacob Edell, and Jackson Bailey Harper. 2024 · 2024
Later among the works it cites.
Self-assessment, Exhibition, and Recognition: a Review of Personality in Large Language Models
Zhiyuan Wen, Yu Yang, Jiannong Cao, Haoming Sun, Ruosong Yang, and Shuaiqi Liu. 2024 · 2024
Later among the works it cites.
Controllm: Crafting diverse personalities for language models
Yixuan Weng, Shizhu He, Kang Liu, Shengping Liu, and Jun Zhao. 2024 · 2024
Later among the works it cites.
Can Large Language Model Agents Simulate Human Trust Behavior?. In The Thirty-eighth Annual Conference on Neural Information Processing Systems
Chengxing Xie, Canyu Chen, Feiran Jia, Ziyu Ye, Shiyang Lai, Kai Shu, Jindong Gu, Adel Bibi, Ziniu Hu, David Jurgens, et al · 2024
Later among the works it cites.
Yang Yan, Lizhi Ma, Anqi Li, Jingsong Ma, and Zhenzhong Lan. 2024 · 2024
Later among the works it cites.
Humanity in AI: Detecting the Personality of Large Language Models
Baohua Zhan, Yongyi Huang, Wenyao Cui, Huaping Zhang, and Jianyun Shang. 2024 · 2024
Later among the works it cites.
The Better Angels of Machine Personality: How Personality Relates to LLM Safety
Jie Zhang, Dongrui Liu, Chen Qian, Ziyue Gan, Yong Liu, Yu Qiao, and Jing Shao. 2024a · 2024
Later among the works it cites.
Yadong Zhang, Shaoguang Mao, Tao Ge, Xun Wang, Yan Xia, Man Lan, and Furu Wei. 2024b · 2024
Later among the works it cites.
Claude AI (Version 3.7 Sonnet)
Anthropic. 2025 · 2025
Closest in time.
Evaluating Personality Traits in Large Language Models: Insights from Psychological Questionnaires
Pranav Bhandari, Usman Naseem, Amitava Datta, Nicolas Fay, and Mehwish Nasim. 2025 · 2025
Closest in time.
Language Models Predict Empathy Gaps Between Social In-groups and Out-groups
Yu Hou, Hal Daumé III, and Rachel Rudinger. 2025 · 2025
Closest in time.
PokéChamp: an Expert-level Minimax Language Agent
Seth Karten, Andy Luu Nguyen, and Chi Jin. 2025 · 2025
Closest in time.
Exploring the Potential of Large Language Models to Simulate Personality
Maria Molchanova, Anna Mikhailova, Anna Korzanova, Lidiia Ostyakova, and Alexandra Dolidze. 2025 · 2025
Closest in time.
Hainiu Xu, Siya Qi, Jiazheng Li, Yuxiang Zhou, Jinhua Du, Caroline Catmur, and Yulan He. 2025 · 2025
Closest in time.
A survey of table reasoning with large language models
Xuanliang Zhang, Dingzirui Wang, Longxu Dou, Qingfu Zhu, and Wanxiang Che. 2025 · 2025
Closest in time.
LMLPA: Language Model Linguistic Personality Assessment
Jingyao Zheng, Xian Wang, Simo Hosio, Xiaoxian Xu, and Lik-Hang Lee. 2025 · 2025
Closest in time.