Fetching the paper…
Reading the bibliography…
Large language models (LLMs) are increasingly deployed in socially grounded applications, where success requires interpreting context, inferring others' mental states, and reasoning about unreliable information.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff. 1978 · 1978
Earlier work this paper cites.
Counterfactual thinking
Neal J. Roese. 1997 · 1997
Earlier work this paper cites.
What is agency?
Mustafa Emirbayer and Ann Mische. 1998 · 1998
Earlier work this paper cites.
Social cognition: Making sense of people
Ziva Kunda. 1999 · 1999
Earlier work this paper cites.
Social cognition in humans
Chris D Frith and Uta Frith. 2007 · 2007
Earlier work this paper cites.
Multiagent systems: Algorithmic, game-theoretic, and logical foundations
Yoav Shoham and Kevin Leyton-Brown. 2008 · 2008
Earlier work this paper cites.
An open review of openreview: A critical analysis of the machine learning conference review process
David Tran, Alex Valtchanov, Keshav Ganapathy, Raymond Feng, Eric Slud, Micah Goldblum, and Tom Goldstein. 2020 · 2010
Earlier work this paper cites.
The Resistance: Avalon
Don Eskridge. 2012 · 2012
Earlier work this paper cites.
How noisy social media text, how diffrnt social media sources?
Timothy Baldwin, Paul Cook, Marco Lui, Andrew MacKinlay, and Li Wang. 2013 · 2013
Earlier work this paper cites.
Social physics: How good ideas spread-the lessons from a new science
Alex Pentland. 2014 · 2014
Earlier work this paper cites.
Making minds: How theory of mind develops
Henry M Wellman. 2014 · 2014
Earlier work this paper cites.
The science of fake news
David MJ Lazer, Matthew A Baum, Yochai Benkler, Adam J Berinsky, Kelly M Greenhill, Filippo Menczer, Miriam J Metzger, Brendan Nyhan, Gordon Pennycook, David Rothschild, et al. 2018 · 2018
Earlier work this paper cites.
Evaluating theory of mind in question answering
Aida Nematzadeh, Kaylee Burns, Erin Grant, Alison Gopnik, and Thomas L Griffiths. 2018 · 2018
Earlier work this paper cites.
Social-ecological systems as complex adaptive systems: Organizing principles for advancing research methods and approaches
Rika Preiser, Reinette Biggs, Alta De Vos, and Carl Folke. 2018 · 2018
Earlier work this paper cites.
Fusing heterogeneous data: A case for remote sensing and social media
Han Wang, Erik Skau, Hamid Krim, and Guido Cervone. 2018 · 2018
Earlier work this paper cites.
Superhuman AI for multiplayer poker
Noam Brown and Tuomas Sandholm. 2019 · 2019
Earlier work this paper cites.
Social IQa: Commonsense reasoning about social interactions
Maarten Sap, Hannah Rashkin, Derek Chen, Ronan Le Bras, and Yejin Choi. 2019b · 2019
Earlier work this paper cites.
HellaSwag: Can a machine really finish your sentence?
Rowan Zellers, Ari Holtzman, Yonatan Bisk, Ali Farhadi, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
PIQA: Reasoning about physical commonsense in natural language
Yonatan Bisk, Rowan Zellers, Ronan Le Bras, Jianfeng Gao, and Yejin Choi. 2020 · 2020
Earlier work this paper cites.
GoEmotions: A dataset of fine-grained emotions
Dorottya Demszky, Dana Movshovitz-Attias, Jeongwoo Ko, Alan Cowen, Gaurav Nemade, and Sujith Ravi. 2020 · 2020
Earlier work this paper cites.
Social chemistry 101: Learning to reason about social and moral norms
Maxwell Forbes, Jena D Hwang, Vered Shwartz, Maarten Sap, and Yejin Choi. 2020 · 2020
Earlier work this paper cites.
CommonGen: A constrained text generation challenge for generative commonsense reasoning
Bill Yuchen Lin, Wangchunshu Zhou, Ming Shen, Pei Zhou, Chandra Bhagavatula, Yejin Choi, and Xiang Ren. 2020 · 2020
Earlier work this paper cites.
TextAttack: A framework for adversarial attacks, data augmentation, and adversarial training in NLP
John X Morris, Eli Lifland, Jin Yong Yoo, Jake Grigsby, Di Jin, and Yanjun Qi. 2020 · 2020
Earlier work this paper cites.
GLUCOSE: GeneraLized and COntextualized story explanations
Nasrin Mostafazadeh, Aditya Kalyanpur, Lori Moon, David Buchanan, Lauren Berkowitz, Or Biran, and Jennifer Chu-Carroll. 2020 · 2020
Earlier work this paper cites.
Beyond accuracy: Behavioral testing of NLP models with CheckList
Marco Tulio Ribeiro, Tongshuang Wu, Carlos Guestrin, and Sameer Singh. 2020 · 2020
Earlier work this paper cites.
Aligning ai with shared human values
Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch, Jerry Li, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
CREAK: A dataset for commonsense reasoning over entity knowledge
Yasumasa Onoe, Michael JQ Zhang, Eunsol Choi, and Greg Durrett. 2021 · 2021
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Cited alongside, same era.
Mastering the game of no-press diplomacy via human-regularized reinforcement learning and planning
Anton Bakhtin, David J Wu, Adam Lerer, Jonathan Gray, Athul Paul Jacob, Gabriele Farina, Alexander H Miller, and Noam Brown. 2023 · 2023
Cited alongside, same era.
Understanding social reasoning in language models with language models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman. 2023 · 2023
Cited alongside, same era.
What can large language models do in chemistry? a comprehensive benchmark on eight tasks
Taicheng Guo, Kehan Guo, Bozhao Nan, Zhenwen Liang, Zhichun Guo, Nitesh V. Chawla, Olaf Wiest, and Xiangliang Zhang. 2023 · 2023
Cited alongside, same era.
Pre-trained multimodal large language model enhances dermatological diagnosis using SkinGPT-4
Juexiao Zhou, Xiaonan He, Liyuan Sun, Jiannan Xu, Xiuying Chen, Yuetan Chu, Longxi Zhou, Xingyu Liao, Bin Zhang, Shawn Afvari, et al. 2024 · 2024
Later among the works it cites.
Cross-cultural transfer of commonsense reasoning in LLMs: Evidence from the Arab world
Saeed Almheiri, Rania Elbadry, Mena Attia, Chenxi Wang, Preslav Nakov, Timothy Baldwin, and Fajri Koto. 2025 · 2025
Closest in time.
Introducing Claude Sonnet 4.5
Anthropic. 2025 · 2025
Closest in time.
Evaluate bias without manual test sets: A concept representation perspective for LLMs
Lang Gao, Kaiyang Wan, Wei Liu, Chenxi Wang, Zirui Song, Zixiang Xu, Yanbo Wang, Veselin Stoyanov, and Xiuying Chen. 2025 · 2025
Closest in time.
Gemini 2.5 Pro Experimental (gemini-2.5-pro-exp-03-25)
Google DeepMind. 2025 · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FANToM: A benchmark for stress-testing machine theory of mind in interactions
Hyunwoo Kim, Melanie Sclar, Xuhui Zhou, Ronan Le Bras, Gunhee Kim, Yejin Choi, and Maarten Sap. 2023 · 2023
Cited alongside, same era.
Theory of mind for multi-agent collaboration via large language models
Huao Li, Yu Quan Chong, Simon Stepputtis, Joseph Campbell, Dana Hughes, Michael Lewis, and Katia Sycara. 2023 · 2023
Cited alongside, same era.
AvalonBench: Evaluating LLMs playing the game of avalon
Jonathan Light, Min Cai, Sheng Shen, and Ziniu Hu. 2023 · 2023
Cited alongside, same era.
Self-refine: Iterative refinement with self-feedback
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, et al. 2023 · 2023
Cited alongside, same era.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2023 · 2023
Cited alongside, same era.
Direct preference optimization: Your language model is secretly a reward model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D Manning, Stefano Ermon, and Chelsea Finn. 2023 · 2023
Cited alongside, same era.
Jaroslaw Szumega, Lamine Bougueroua, Blerina Gkotse, Pierre Jouvelot, and Federico Ravotti. 2023 · 2023
Cited alongside, same era.
DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Peiyi Wang, Qihao Zhu, Runxin Xu, Ruoyu Zhang, Shirong Ma, Xiao Bi, et al. 2025 · 2025
Closest in time.
ASTRO: Automatic strategy optimization for non-cooperative dialogues
Yikuan Hu, Chen Huang, and Wenqiang Lei. 2025b · 2025
Closest in time.
On the trustworthiness of generative foundation models: Guideline, assessment, and perspective
Yue Huang, Chujie Gao, Siyuan Wu, Haoran Wang, Xiangqi Wang, Yujun Zhou, Yanbo Wang, Jiayi Ye, Jiawen Shi, Qihui Zhang, et al. 2025 · 2025
Closest in time.
A personalized conversational benchmark: Towards simulating personalized conversations
Li Li, Peilin Cai, Ryan A. Rossi, Franck Dernoncourt, Branislav Kveton, Junda Wu, Tong Yu, Linxin Song, Tiankai Yang, Yuehan Qin, Nesreen K. Ahmed, Samyadeep Basu, Subhojyoti Mukherjee, Ruiyi Zhang, Zhengmian Hu, Bo Ni, Yuxiao Zhou, Zichao Wang, Yue Huang, and 5 others. 2025 · 2025
Closest in time.
The stepwise deception: Simulating the evolution from true news to fake news with LLM agents
Yuhan Liu, Zirui Song, Juntian Zhang, Xiaoqing Zhang, Xiuying Chen, and Rui Yan. 2025b · 2025
Closest in time.
OpenAI o3-mini
OpenAI. 2025 · 2025
Closest in time.
QwQ-32B: Embracing the power of reinforcement learning
Qwen Team. 2025 · 2025
Closest in time.
QUITE: A query rewrite system beyond rules with LLM agents
Yuyang Song, Hanxu Yan, Jiale Lao, Yibo Wang, Yufei Li, Yuanchun Zhou, Jianguo Wang, and Mingjie Tang. 2025 · 2025
Closest in time.
DebateBench: A challenging long context reasoning benchmark for large language models
Utkarsh Tiwari, Aryan Seth, Adi Mukherjee, Kaavya Mer, Kavish, and Dhruv Kumar. 2025 · 2025
Closest in time.
Word form matters: LLMs’ semantic reconstruction under typoglycemia
Chenxi Wang, Tianle Gu, Zhongyu Wei, Lang Gao, Zirui Song, and Xiuying Chen. 2025a · 2025
Closest in time.
Under the shadow of Babel: How language shapes reasoning in LLMs
Chenxi Wang, Yixuan Zhang, Lang Gao, Zixiang Xu, Zirui Song, Yanbo Wang, and Xiuying Chen. 2025c · 2025
Closest in time.
Exploring large language models for word games: Who is the spy?
Chentian Wei, Jiewei Chen, and Jinzhu Xu. 2025 · 2025
Closest in time.
Towards dynamic theory of mind: Evaluating LLM adaptation to temporal evolution of human states
Yang Xiao, Jiashuo Wang, Qiancheng Xu, Changhe Song, Chunpu Xu, Yi Cheng, Wenjie Li, and Pengfei Liu. 2025 · 2025
Closest in time.
Cross-lingual pitfalls: Automatic probing cross-lingual weakness of multilingual large language models
Zixiang Xu, Yanbo Wang, Yue Huang, Xiuying Chen, Jieyu Zhao, Meng Jiang, and Xiangliang Zhang. 2025 · 2025
Closest in time.
Multimodal generative engine optimization: Rank manipulation for vision–language model rankers
Yixuan Du, Chenxiao Yu, Haoyan Xu, Ziyi Wang, Yue Zhao, and Xiyang Hu. 2026 · 2026
Closest in time.
ProbeLLM: Automating principled diagnosis of LLM failures
Yue Huang, Zhengzhe Jiang, Yuchen Ma, Yu Jiang, Xiangqi Wang, Yujun Zhou, Yuexing Hao, Kehan Guo, Pin-Yu Chen, Marzyeh Ghassemi, Stefan Feuerriegel, and Xiangliang Zhang. 2026 · 2026
Closest in time.
RiskLab: A controlled toolkit for probing emergent risks in LLM-based multi-agent systems
Yu Jiang, Wenjie Wang, Yue Huang, Yanbo Wang, Zhenhong Zhou, Xiuying Chen, Yang Liu, Pin-Yu Chen, Wei Wang, and Xiangliang Zhang. 2026 · 2026
Closest in time.
“Someone Hid It!”: Query-agnostic black-box attacks on LLM-based retrieval
Jiate Li, Defu Cao, Li Li, Wei Yang, Yuehan Qin, Chenxiao Yu, Tiannuo Yang, Ryan A. Rossi, Yan Liu, Xiyang Hu, and Yue Zhao. 2026 · 2026
Closest in time.
Large-scale terminal agentic trajectory generation from dockerized environments
Siwei Wu, Yizhi Li, Yuyang Song, Wei Zhang, Yang Wang, Riza Batista-Navarro, Xian Yang, Mingjie Tang, Bryan Dai, Jian Yang, and Chenghua Lin. 2026 · 2026
Closest in time.
No attacker needed: Unintentional cross-user contamination in shared-state LLM agents
Tiankai Yang, Jiate Li, Yi Nian, Shen Dong, Ruiyao Xu, Ryan A. Rossi, Kaize Ding, and Yue Zhao. 2026 · 2026
Closest in time.
Tracing moral foundations in large language models
Chenxiao Yu, Bowen Yi, Farzan Karimi-Malekabadi, Suhaib Abdurahman, Jinyi Ye, Shrikanth Narayanan, Yue Zhao, and Morteza Dehghani. 2026 · 2026
Closest in time.