Fetching the paper…
Reading the bibliography…
While humans naturally develop theory of mind (ToM), the capability to understand other people's mental states and beliefs, state-of-the-art large language models (LLMs) underperform on simple ToM benchmarks.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff. 1978 · 1978
Earlier work this paper cites.
Beliefs about beliefs: Representation and constraining function of wrong beliefs in young children’s understanding of deception
Heinz Wimmer and Josef Perner. 1983 · 1983
Earlier work this paper cites.
Does the autistic child have a theory of mind?
Simon Baron-Cohen, Alan M. Leslie, and Uta Frith. 1985 · 1985
Earlier work this paper cites.
Temperament and the development of self-regulation
Mary K Rothbart and Michael I Posner. 1985 · 1985
Earlier work this paper cites.
Young children understand that looking leads to knowing (so long as they are looking into a single barrel)
Chris Pratt and Peter Bryant. 1990 · 1990
Earlier work this paper cites.
The ‘seeing-leads-to-knowing’ deficit in autism: The pratt and bryant probe
Simon Baron-Cohen and Frances Goodhart. 1994 · 1994
Earlier work this paper cites.
Individual differences in inhibitory control and children’s theory of mind
S M Carlson and L J Moses. 2001 · 2001
Earlier work this paper cites.
How specific is the relation between executive function and theory of mind? contributions of inhibitory control and working memory
Stephanie M Carlson, Louis J Moses, and Casey Breton. 2002 · 2002
Earlier work this paper cites.
Bayesian theory of mind: Modeling joint belief-desire attribution
Chris Baker, Rebecca Saxe, and Joshua Tenenbaum. 2011 · 2011
Earlier work this paper cites.
Revisiting the evaluation of theory of mind through question answering
Matthew Le, Y-Lan Boureau, and Maximilian Nickel. 2019 · 2019
Earlier work this paper cites.
Foundations of theory of mind and its development in early childhood
Hannes Rakoczy. 2022 · 2022
Earlier work this paper cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, brian ichter, Fei Xia, Ed H. Chi, Quoc V Le, and Denny Zhou. 2022 · 2022
Cited alongside, same era.
Understanding social reasoning in language models with language models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman. 2023 · 2023
Cited alongside, same era.
FANToM: A benchmark for stress-testing machine theory of mind in interactions
Hyunwoo Kim, Melanie Sclar, Xuhui Zhou, Ronan Bras, Gunhee Kim, Yejin Choi, and Maarten Sap. 2023 · 2023
Cited alongside, same era.
Towards a holistic landscape of situated theory of mind in large language models
Ziqiao Ma, Jacob Sansom, Run Peng, and Joyce Chai. 2023b · 2023
Cited alongside, same era.
Openai json mode
OpenAI. 2023 · 2023
Cited alongside, same era.
Minding language models’ (lack of) theory of mind: A plug-and-play multi-character belief tracker
Think twice: Perspective-taking improves large language models’ theory-of-mind capabilities
Alex Wilf, Sihyun Shawn Lee, Paul Pu Liang, and Louis-Philippe Morency. 2023 · 2023
Later among the works it cites.
Hi-ToM: A benchmark for evaluating higher-order theory of mind reasoning in large language models
Yufan Wu, Yinghui He, Yilin Jia, Rada Mihalcea, Yulong Chen, and Naihao Deng. 2023 · 2023
Later among the works it cites.
How far are large language models from agents with theory-of-mind?
Pei Zhou, Aman Madaan, Srividya Pranavi Potharaju, Aditya Gupta, Kevin R. McKee, Ari Holtzman, Jay Pujara, Xiang Ren, Swaroop Mishra, Aida Nematzadeh, Shyam Upadhyay, and Manaal Faruqui. 2023 · 2023
Later among the works it cites.
Llama 3 model card
AI@Meta. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Melanie Sclar, Sachin Kumar, Peter West, Alane Suhr, Yejin Choi, and Yulia Tsvetkov. 2023 · 2023
Cited alongside, same era.
Do large language models know what humans know?
Sean Trott, Cameron Jones, Tyler Chang, James Michaelov, and Benjamin Bergen. 2023 · 2023
Cited alongside, same era.
Theory of mind in large language models: Examining performance of 11 state-of-the-art models vs. children aged 7-10 on advanced tests
Max van Duijn, Bram van Dijk, Tom Kouwenhoven, Werner de Valk, Marco Spruit, and Peter van der Putten. 2023 · 2023
Cited alongside, same era.
System 2 attention (is something you might need too)
Jason Weston and Sainbayar Sukhbaatar. 2023 · 2023
Cited alongside, same era.
Can a machine know that we know what it knows?
Oliver Whang. 2023 · 2023
Cited alongside, same era.
ToMChallenges: A principle-guided dataset and diverse evaluation tasks for exploring theory of mind
Xiaomeng Ma, Lingyu Gao, and Qihui Xu. 2023a
Cited in the paper.
Gemini-Team. 2024 · 2024
Closest in time.
Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, Gianna Lengyel, Guillaume Bour, Guillaume Lample, Lélio Renard Lavaud, Lucile Saulnier, Marie-Anne Lachaux, Pierre Stock, Sandeep Subramanian, Sophia Yang, Szymon Antoniak, Teven Le Scao, Théophile Gervet, Thibaut Lavril, Thomas Wang, Timothée Lacroix, and William El Sayed. 2024 · 2024
Closest in time.
Clever hans or neural theory of mind? stress testing social reasoning in large language models
Natalie Shapira, Mosh Levy, Seyed Hossein Alavi, Xuhui Zhou, Yejin Choi, Yoav Goldberg, Maarten Sap, and Vered Shwartz. 2024 · 2024
Closest in time.
Llms achieve adult human performance on higher-order theory of mind tasks
Winnie Street, John Oliver Siy, Geoff Keeling, Adrien Baranes, Benjamin Barnett, Michael McKibben, Tatenda Kanyere, Alison Lentz, Blaise Aguera y Arcas, and Robin I. M. Dunbar. 2024 · 2024
Closest in time.
Tom-lm: Delegating theory of mind reasoning to external symbolic executors in large language models
Weizhi Tang and Vaishak Belle. 2024 · 2024
Closest in time.
Hainiu Xu, Runcong Zhao, Lixing Zhu, Jinhua Du, and Yulan He. 2024 · 2024
Closest in time.