Fetching the paper…
Reading the bibliography…
Theory of Mind (ToM) can be used to assess the capabilities of Large Language Models (LLMs) in complex scenarios where social reasoning is required.
A note on "iterated knowings"
James Cargile. 1970 · 1970
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff. 1978 · 1978
Earlier work this paper cites.
Ascribing Mental Qualities To Machines
John McCarthy. 1979 · 1979
Earlier work this paper cites.
Beliefs about beliefs: Representation and constraining function of wrong beliefs in young children’s understanding of deception
Heinz Wimmer and Josef Perner. 1983 · 1983
Earlier work this paper cites.
Does the autistic child have a “theory of mind”?
Simon Baron-Cohen, Alan M. Leslie, and Uta Frith. 1985 · 1985
Earlier work this paper cites.
The intentional stance in theory and practice
Daniel C. Dennett. 1988 · 1988
Earlier work this paper cites.
Cognitive load during problem solving: Effects on learning
John Sweller. 1988 · 1988
Earlier work this paper cites.
The theory theory
Alison Gopnik and Henry M. Wellman. 1994 · 1994
Earlier work this paper cites.
Cognitive load theory, learning difficulty, and instructional design
John Sweller. 1994 · 1994
Earlier work this paper cites.
Empathy: Its ultimate and proximate bases
Stephanie D Preston and Frans BM De Waal. 2002 · 2002
Earlier work this paper cites.
Theory of Mind for a Humanoid Robot
Brian Scassellati. 2002 · 2002
Earlier work this paper cites.
Automatic cognitive load detection from speech features
Bo Yin, Natalie Ruiz, Fang Chen, and M. Asif Khawaja. 2007 · 2007
Earlier work this paper cites.
The cultural origins of human cognition
Michael Tomasello. 2009 · 2009
Earlier work this paper cites.
Element Interactivity and Intrinsic, Extraneous, and Germane Cognitive Load
John Sweller. 2010 · 2010
Earlier work this paper cites.
Bayesian Theory of Mind: Modeling Joint Belief-Desire Attribution
Chris Baker, Rebecca Saxe, and Joshua Tenenbaum. 2011 · 2011
Earlier work this paper cites.
Thinking, Fast and Slow
Daniel Kahneman. 2011 · 2011
Earlier work this paper cites.
Measuring Cognitive Load
John Sweller, Paul Ayres, and Slava Kalyuga. 2011 · 2011
Earlier work this paper cites.
Folk psychology and the explanation of human behavior 1
Paul M Churchland. 2013 · 2013
Earlier work this paper cites.
The Development of Theory of Mind: Historical Reflections
Henry M. Wellman. 2017 · 2017
Cited alongside, same era.
David Ha and Jürgen Schmidhuber. 2018 · 2018
Cited alongside, same era.
Machine theory of mind
Neil Rabinowitz, Frank Perbet, Francis Song, Chiyuan Zhang, SM Ali Eslami, and Matthew Botvinick. 2018 · 2018
Cited alongside, same era.
Revisiting the Evaluation of Theory of Mind through Question Answering
Matthew Le, Y-Lan Boureau, and Maximilian Nickel. 2019 · 2019
Cited alongside, same era.
Social IQa: Commonsense Reasoning about Social Interactions
Maarten Sap, Hannah Rashkin, Derek Chen, Ronan Le Bras, and Yejin Choi. 2019 · 2019
Cited alongside, same era.
Implicit representations of meaning in neural language models
Theory of mind may have spontaneously emerged in large language models
Michal Kosinski. 2023 · 2023
Later among the works it cites.
Emergent world representations: Exploring a sequence model trained on a synthetic task
Kenneth Li, Aspen K Hopkins, David Bau, Fernanda Viégas, Hanspeter Pfister, and Martin Wattenberg. 2023 · 2023
Later among the works it cites.
Boosting theory-of-mind performance in large language models via prompting
Shima Rahimi Moghaddam and Christopher J. Honey. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Belinda Z. Li, Maxwell Nye, and Jacob Andreas. 2021 · 2021
Cited alongside, same era.
Show your work: Scratchpads for intermediate computation with language models
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, et al. 2021 · 2021
Cited alongside, same era.
Learning chess blindfolded: Evaluating language models on state tracking
Shubham Toshniwal, Sam Wiseman, Karen Livescu, and Kevin Gimpel. 2021 · 2021
Cited alongside, same era.
Baby Intuitions Benchmark (BIB): Discerning the goals, preferences, and actions of others
Kanishk Gandhi, Gala Stojnic, Brenden M. Lake, and Moira R. Dillon. 2022 · 2022
Cited alongside, same era.
Multi-agent deep reinforcement learning: a survey
Sven Gronauer and Klaus Diepold. 2022 · 2022
Cited alongside, same era.
Neural theory-of-mind? on the limits of social intelligence in large lms
Maarten Sap, Ronan LeBras, Daniel Fried, and Yejin Choi. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Cited alongside, same era.
Joon Sung Park, Joseph C. O’Brien, Carrie J. Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein. 2023 · 2023
Later among the works it cites.
MindGames: Targeting theory of mind in large language models with dynamic epistemic modal logic
Damien Sileo and Antoine Lernould. 2023 · 2023
Later among the works it cites.
Large language models fail on trivial alterations to theory-of-mind tasks
Tomer Ullman. 2023 · 2023
Later among the works it cites.
Lionel Wong, Gabriel Grand, Alexander K Lew, Noah D Goodman, Vikash K Mansinghka, Jacob Andreas, and Joshua B Tenenbaum. 2023 · 2023
Later among the works it cites.
Llama 3 model card
AI@Meta. 2024 · 2024
Closest in time.
Tombench: Benchmarking theory of mind in large language models
Zhuang Chen, Jincenzi Wu, Jinfeng Zhou, Bosi Wen, Guanqun Bi, Gongyao Jiang, Yaru Cao, Mengting Hu, Yunghwei Lai, Zexuan Xiong, et al. 2024 · 2024
Closest in time.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al. 2024 · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al. 2024 · 2024
Closest in time.
Code pretraining improves entity tracking abilities of language models
Najoung Kim, Sebastian Schuster, and Shubham Toshniwal. 2024 · 2024
Closest in time.
Dissociating language and thought in large language models
Kyle Mahowald, Anna A Ivanova, Idan A Blank, Nancy Kanwisher, Joshua B Tenenbaum, and Evelina Fedorenko. 2024 · 2024
Closest in time.
Clever hans or neural theory of mind? stress testing social reasoning in large language models
Natalie Shapira, Mosh Levy, Seyed Hossein Alavi, Xuhui Zhou, Yejin Choi, Yoav Goldberg, Maarten Sap, and Vered Shwartz. 2024 · 2024
Closest in time.
Testing theory of mind in large language models and humans
James WA Strachan, Dalila Albergo, Giulia Borghini, Oriana Pansardi, Eugenio Scaliti, Saurabh Gupta, Krati Saxena, Alessandro Rufo, Stefano Panzeri, Guido Manzi, et al. 2024 · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan. 2024 · 2024
Closest in time.