Fetching the paper…
Reading the bibliography…
Understanding people's social interactions in complex real-world scenarios often relies on intricate mental reasoning.
Does the autistic child have a “theory of mind”?
Baron-Cohen, S.; Leslie, A. M.; and Frith, U. 1985 · 1985
Earlier work this paper cites.
Meta-analysis of theory-of-mind development: The truth about false belief
Wellman, H. M.; Cross, D.; and Watson, J. 2001 · 2001
Earlier work this paper cites.
Preschool emotional competence: Pathway to social competence?
Denham, S. A.; Blair, K. A.; DeMulder, E.; Levitas, J.; Sawyer, K.; Auerbach-Major, S.; and Queenan, P. 2003 · 2003
Earlier work this paper cites.
A Framework for Sequential Planning in Multi-Agent Settings
Gmytrasiewicz, P. J.; and Doshi, P. 2005 · 2005
Earlier work this paper cites.
Social evaluation by preverbal infants
Hamlin, J. K.; Wynn, K.; and Bloom, P. 2007 · 2007
Earlier work this paper cites.
Help or hinder: Bayesian models of social goal inference
Ullman, T.; Baker, C.; Macindoe, O.; Evans, O.; Goodman, N.; and Tenenbaum, J. 2009 · 2009
Earlier work this paper cites.
Watch-And-Help: A Challenge for Social Perception and Human-AI Collaboration
Puig, X.; Shu, T.; Li, S.; Wang, Z.; Tenenbaum, J. B.; Fidler, S.; and Torralba, A. 2020 · 2010
Earlier work this paper cites.
Rational quantitative attribution of beliefs, desires and percepts in human mentalizing
Baker, C. L.; Jara-Ettinger, J.; Saxe, R.; and Tenenbaum, J. B. 2017 · 2017
Earlier work this paper cites.
Rabinowitz, N. C.; Perbet, F.; Song, H. F.; Zhang, C.; Eslami, S. M. A.; and Botvinick, M. 2018 · 2018
Earlier work this paper cites.
IPOMDP-Net: A Deep Neural Network for Partially Observable Multi-Agent Planning Using Interactive POMDPs
Han, Y.; and Gmytrasiewicz, P. 2019 · 2019
Earlier work this paper cites.
Revisiting the evaluation of theory of mind through question answering
Le, M.; Boureau, Y.-L.; and Nickel, M. 2019 · 2019
Earlier work this paper cites.
Adventures in Flatland: Perceiving Social Interactions Under Physical Dynamics
Shu, T.; Kryven, M.; Ullman, T. D.; and Tenenbaum, J. 2020 · 2020
Earlier work this paper cites.
Online bayesian goal inference for boundedly rational planning agents
Zhi-Xuan, T.; Mann, J.; Silver, T.; Tenenbaum, J.; and Mansinghka, V. 2020 · 2020
Earlier work this paper cites.
Exploring RoBERTa’s theory of mind through textual entailment
Cohen, M. 2021 · 2021
Earlier work this paper cites.
Baby Intuitions Benchmark (BIB): Discerning the goals, preferences, and actions of others
Gandhi, K.; Stojnic, G.; Lake, B. M.; and Dillon, M. R. 2021 · 2021
Earlier work this paper cites.
Agent: A benchmark for core psychological reasoning
Shu, T.; Bhandwaldar, A.; Gan, C.; Smith, K.; Liu, S.; Gutfreund, D.; Spelke, E.; Tenenbaum, J.; and Ullman, T. 2021 · 2021
Earlier work this paper cites.
Social interactions as recursive mdps
Tejwani, R.; Kuo, Y.-L.; Shu, T.; Katz, B.; and Barbu, A. 2021 · 2021
Earlier work this paper cites.
Large Language Models are Zero-Shot Reasoners
Kojima, T.; Gu, S. S.; Reid, M.; Matsuo, Y.; and Iwasawa, Y. 2022 · 2022
Cited alongside, same era.
Sparks of artificial general intelligence: Early experiments with gpt-4
Bubeck, S.; Chandrasekaran, V.; Eldan, R.; Gehrke, J.; Horvitz, E.; Kamar, E.; Lee, P.; Lee, Y. T.; Li, Y.; Lundberg, S.; et al. 2023 · 2023
Cited alongside, same era.
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Chen, Z.; Wu, J.; Wang, W.; Su, W.; Chen, G.; Xing, S.; Zhong, M.; Zhang, Q.; Zhu, X.; Lu, L.; Li, B.; Luo, P.; Lu, T.; Qiao, Y.; and Dai, J. 2023 · 2023
Cited alongside, same era.
Hi-tom: A benchmark for evaluating higher-order theory of mind reasoning in large language models
He, Y.; Wu, Y.; Jia, Y.; Mihalcea, R.; Chen, Y.; and Deng, N. 2023 · 2023
Cited alongside, same era.
VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Cheng, Z.; Leng, S.; Zhang, H.; Xin, Y.; Li, X.; Chen, G.; Zhu, Y.; Zhang, W.; Luo, Z.; Zhao, D.; and Bing, L. 2024 · 2024
Closest in time.
Understanding social reasoning in language models with language models
Gandhi, K.; Fränken, J.-P.; Gerstenberg, T.; and Goodman, N. 2024 · 2024
Closest in time.
TimeToM: Temporal Space is the Key to Unlocking the Door of Large Language Models’ Theory-of-Mind
Hou, G.; Zhang, W.; Shen, Y.; Wu, L.; and Lu, W. 2024 · 2024
Closest in time.
Ivanova, A. A.; Sathe, A.; Lipkin, B.; Kumar, U.; Radkani, S.; Clark, T. H.; Kauf, C.; Hu, J.; Pramod, R. T.; Grand, G.; Paulun, V.; Ryskina, M.; Akyürek, E.; Wilcox, E.; Rashid, N.; Choshen, L.; Levy, R.; Fedorenko, E.; Tenenbaum, J.; and Andreas, J. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kim, H.; Sclar, M.; Zhou, X.; Bras, R. L.; Kim, G.; Choi, Y.; and Sap, M. 2023 · 2023
Cited alongside, same era.
Theory of Mind May Have Spontaneously Emerged in Large Language Models
Kosinski, M. 2023 · 2023
Cited alongside, same era.
Visual Instruction Tuning
Liu, H.; Li, C.; Wu, Q.; and Lee, Y. J. 2023 · 2023
Cited alongside, same era.
Relational visual representations underlie human social interaction recognition
Malik, M.; and Isik, L. 2023 · 2023
Cited alongside, same era.
OpenAI. 2023 · 2023
Cited alongside, same era.
Puig, X.; Shu, T.; Tenenbaum, J. B.; and Torralba, A. 2023 · 2023
Cited alongside, same era.
MultiVENT: Multilingual Videos of Events with Aligned Natural Text
Sanders, K.; Etter, D.; Kriz, R.; and Van Durme, B. 2023 · 2023
Cited alongside, same era.
Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks
Ullman, T. 2023 · 2023
Cited alongside, same era.
Neural Amortized Inference for Nested Multi-agent Reasoning
Jha, K.; Le, T. A.; Jin, C.; Kuo, Y.-L.; Tenenbaum, J. B.; and Shu, T. 2024 · 2024
Closest in time.
Mmtom-qa: Multimodal theory of mind question answering
Jin, C.; Wu, Y.; Cao, J.; Xiang, J.; Kuo, Y.-L.; Hu, Z.; Ullman, T.; Torralba, A.; Tenenbaum, J. B.; and Shu, T. 2024 · 2024
Closest in time.
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
Pandya, P.; Talwarr, A. S.; Gupta, V.; Kataria, T.; Gupta, V.; and Roth, D. 2024 · 2024
Closest in time.
Perception test: A diagnostic benchmark for multimodal video models
Patraucean, V.; Smaira, L.; Gupta, A.; Recasens, A.; Markeeva, L.; Banarse, D.; Koppula, S.; Malinowski, M.; Yang, Y.; Doersch, C.; et al. 2024 · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reid, M.; Savinov, N.; Teplyashin, D.; Lepikhin, D.; Lillicrap, T.; Alayrac, J.-b.; Soricut, R.; Lazaridou, A.; Firat, O.; Schrittwieser, J.; et al. 2024 · 2024
Closest in time.
EmoBench: Evaluating the Emotional Intelligence of Large Language Models
Sabour, S.; Liu, S.; Zhang, Z.; Liu, J. M.; Zhou, J.; Sunaryo, A. S.; Li, J.; Lee, T. M. C.; Mihalcea, R.; and Huang, M. 2024 · 2024
Closest in time.
Views Are My Own, But Also Yours: Benchmarking Theory of Mind using Common Ground
Soubki, A.; Murzaku, J.; Jordehi, A. Y.; Zeng, P.; Markowska, M.; Mirroshandel, S. A.; and Rambow, O. 2024 · 2024
Closest in time.
A Bayesian theory of mind approach to modeling cooperation and communication
Stacy, S.; Gong, S.; Parab, A.; Zhao, M.; Jiang, K.; and Gao, T. 2024 · 2024
Closest in time.
MTVQA: Benchmarking Multilingual Text-Centric Visual Question Answering
Tang, J.; Liu, Q.; Ye, Y.; Lu, J.; Wei, S.; Lin, C.; Li, W.; Mahmood, M. F. F. B.; Feng, H.; Zhao, Z.; et al. 2024 · 2024
Closest in time.
Theory of Mind abilities of Large Language Models in Human-Robot Interaction: An Illusion?
Verma, M.; Bhambri, S.; and Kambhampati, S. 2024 · 2024
Closest in time.
Xu, H.; Zhao, R.; Zhu, L.; Du, J.; and He, Y. 2024 · 2024
Closest in time.
SEED-Story: Multimodal Long Story Generation with Large Language Model
Yang, S.; Ge, Y.; Li, Y.; Chen, Y.; Ge, Y.; Shan, Y.; and Chen, Y. 2024 · 2024
Closest in time.
Zhang, J.; Huang, W.; Ma, Z.; Michel, O.; He, D.; Gupta, T.; Ma, W.-C.; Farhadi, A.; Kembhavi, A.; and Krishna, R. 2024 · 2024
Closest in time.