Fetching the paper…
Reading the bibliography…
Recent advances in Artificial Intelligence (AI) have yielded powerful computational models that, by learning from vast amounts of human-generated data, are increasingly posited as approximate models of human cognition.
Computing machinery and intelligence
Alan Turing · 1950
Earlier work this paper cites.
Construct validity in psychological tests
Lee J Cronbach and Paul E Meehl · 1955
Earlier work this paper cites.
The perceptron: a probabilistic model for information storage and organization in the brain
Frank Rosenblatt · 1958
Earlier work this paper cites.
Eliza—a computer program for the study of natural language communication between man and machine
Joseph Weizenbaum · 1966
Earlier work this paper cites.
Judgment under uncertainty: Heuristics and biases: Biases in judgments reveal some heuristics of thinking under uncertainty
Amos Tversky and Daniel Kahneman · 1974
Earlier work this paper cites.
Psychologism and behaviorism
Ned Block · 1981
Earlier work this paper cites.
Vision: A Computational Investigation into the Human Representation and Processing of Visual Information
David Marr · 1982
Earlier work this paper cites.
Beliefs about beliefs: Representation and constraining function of wrong beliefs in young children’s understanding of deception
Heinz Wimmer and Josef Perner · 1983
Earlier work this paper cites.
(1986) d. e. rumelhart, g. e. hinton, and r. j. williams, ”learning internal representations by error propagation,” parallel distributed processing: Explorations in the microstructures of cognition, vol. i, d. e. rumelhart and j. l. mcclelland (eds.) cambridge, ma: Mit press, pp. 318-362
David E. Rumelhart, Geoffrey E. Hinton, and Ronald J. Williams · 1988
Earlier work this paper cites.
Society of mind
Marvin Minsky · 1988
Earlier work this paper cites.
Cognitive modeling and intelligent tutoring
John R Anderson, C Franklin Boyle, Albert T Corbett, and Matthew W Lewis · 1990
Earlier work this paper cites.
Calibration and probability judgements: Conceptual and methodological issues
Gideon Keren · 1991
Earlier work this paper cites.
Validity of psychological assessment: Validation of inferences from persons’ responses and performances as scientific inquiry into score meaning
Samuel Messick · 1995
Earlier work this paper cites.
Measuring psychological uncertainty: Verbal versus numeric methods
Paul D Windschitl and Gary L Wells · 1996
Earlier work this paper cites.
Bayesian modeling of human concept learning
Joshua Tenenbaum · 1998
Earlier work this paper cites.
Two reasons to abandon the false belief task as a test of theory of mind
Paul Bloom and Tim P German · 2000
Earlier work this paper cites.
Social cognitive theory: An agentic perspective
Albert Bandura · 2001
Earlier work this paper cites.
Probabilistic models of language processing and acquisition
Nick Chater and Christopher D Manning · 2006
Earlier work this paper cites.
Uncertain Judgements: Eliciting Expert Probabilities
A. O’Hagan, C. E. Buck, A. Daneshkhah, J. R. Eiser, and P. H. Garthwaite et al · 2006
Earlier work this paper cites.
Before and below ‘theory of mind’: embodied simulation and the neural correlates of social cognition
Vittorio Gallese · 2007
Earlier work this paper cites.
Liberals and conservatives rely on different sets of moral foundations
Jesse Graham, Jonathan Haidt, and Brian A Nosek · 2009
Earlier work this paper cites.
Adaptive rationality: An evolutionary perspective on cognitive bias
Martie G Haselton, Gregory A Bryant, Andreas Wilke, David A Frederick, Andrew Galperin, Willem E Frankenhuis, and Tyler Moore · 2009
Earlier work this paper cites.
The weirdest people in the world?
Joseph Henrich, Steven J Heine, and Ara Norenzayan · 2010
Earlier work this paper cites.
Legibility and predictability of robot motion
Anca D Dragan, Kenton CT Lee, and Siddhartha S Srinivasa · 2013
Earlier work this paper cites.
Striking individual differences in color perception uncovered by ‘the dress’ photograph
Rosa Lafer-Sousa, Katherine L Hermann, and Bevil R Conway · 2015
Earlier work this paper cites.
Cultural differences in moral judgment and behavior, across and within societies
Jesse Graham, Peter Meindl, Erica Beall, Kate M Johnson, and Li Zhang · 2016
Earlier work this paper cites.
Building machines that learn and think like people
Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman · 2017
Earlier work this paper cites.
Evaluation in artificial intelligence: from task-oriented to ability-oriented measurement
José Hernández-Orallo · 2017
Earlier work this paper cites.
Rational quantitative attribution of beliefs, desires and percepts in human mentalizing
Chris L Baker, Julian Jara-Ettinger, Rebecca Saxe, and Joshua B Tenenbaum · 2017
Earlier work this paper cites.
Measuring and understanding individual differences in cognition, 2018
Neeltje J Boogert, Joah R Madden, Julie Morand-Ferron, and Alex Thornton · 2018
Earlier work this paper cites.
Expert knowledge elicitation: Subjective but scientific
Anthony O’Hagan · 2018
Earlier work this paper cites.
It’s going to be okay: Measuring access to support in online communities
Zijian Wang and David Jurgens · 2018
Earlier work this paper cites.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, Mark K Ho, Tom Griffiths, and Sanjit et al Seshia · 2019
Earlier work this paper cites.
Human uncertainty makes classification more robust
Joshua C Peterson, Ruairidh M Battleday, Thomas L Griffiths, and Olga Russakovsky · 2019
Earlier work this paper cites.
Revisiting the evaluation of theory of mind through question answering
Matthew Le, Y-Lan Boureau, and Maximilian Nickel · 2019
Cited alongside, same era.
Uncovering the structure of self-regulation through data-driven ontology discovery
Ian W Eisenberg, Patrick G Bissett, A Zeynep Enkavi, James Li, David P MacKinnon, Lisa A Marsch, and Russell A Poldrack · 2019
Cited alongside, same era.
Efficient elicitation approaches to estimate collective crowd answers
John Joon Young Chung, Jean Y Song, Sindhu Kutty, Sungsoo Hong, Juho Kim, and Walter S Lasecki · 2019
Cited alongside, same era.
Socialiqa: Commonsense reasoning about social interactions
Maarten Sap, Hannah Rashkin, Derek Chen, Ronan LeBras, and Yejin Choi · 2019
Cited alongside, same era.
Generating plans that predict themselves
Jaime F Fisac, Chang Liu, Jessica B Hamrick, Shankar Sastry, and J Karl et al Hedrick · 2020
Cited alongside, same era.
Inferring the goals of communicating agents from actions and instructions
Lance Ying, Tan Zhi-Xuan, Vikash Mansinghka, and Joshua B Tenenbaum · 2023
Later among the works it cites.
Learning skillful medium-range global weather forecasting
Remi Lam, Alvaro Sanchez-Gonzalez, Matthew Willson, Peter Wirnsberger, Meire Fortunato, Ferran Alet, Suman Ravuri, Timo Ewalds, Zach Eaton-Rosen, Weihua Hu, et al · 2023
Later among the works it cites.
Fine-grained human feedback gives better rewards for language model training
Zeqiu Wu, Yushi Hu, Weijia Shi, Nouha Dziri, Alane Suhr, Prithviraj Ammanabrolu, Noah A Smith, Mari Ostendorf, and Hannaneh Hajishirzi · 2023
Later among the works it cites.
Tombench: Benchmarking theory of mind in large language models, 2024
Zhuang Chen, Jincenzi Wu, Jinfeng Zhou, Bosi Wen, Guanqun Bi, Gongyao Jiang, Yaru Cao, Mengting Hu, Yunghwei Lai, Zexuan Xiong, and Minlie Huang · 2024
Later among the works it cites.
Understanding social reasoning in language models with language models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Performance vs. competence in human–machine comparisons
Chaz Firestone · 2020
Cited alongside, same era.
The logic of universalization guides moral judgment
Sydney Levine, Max Kleiman-Weiner, Laura Schulz, Joshua Tenenbaum, and Fiery Cushman · 2020
Cited alongside, same era.
A case for soft loss functions
Alexandra Uma, Tommaso Fornaciari, Dirk Hovy, Silviu Paun, and Barbara et al Plank · 2020
Cited alongside, same era.
Systematic review and inventory of theory of mind measures for young children
Cindy Beaudoin, Élizabel Leblanc, Charlotte Gagner, and Miriam H Beauchamp · 2020
Cited alongside, same era.
The naive utility calculus as a unified, quantitative framework for action understanding
Julian Jara-Ettinger, Laura E Schulz, and Joshua B Tenenbaum · 2020
Cited alongside, same era.
Understanding human intelligence through human limitations
Thomas L Griffiths · 2020
Cited alongside, same era.
Resource-rational analysis: Understanding human cognition as the optimal use of limited computational resources
Falk Lieder and Thomas L Griffiths · 2020
Cited alongside, same era.
Aligning machine and human visual representations across abstraction levels
Lukas Muttenthaler, Klaus Greff, Frieda Born, Bernhard Spitzer, Simon Kornblith, Michael C Mozer, Klaus-Robert MÞller, Thomas Unterthiner, and Andrew K Lampinen · 2024
Later among the works it cites.
The turing test and our shifting conceptions of intelligence, 2024
Melanie Mitchell · 2024
Later among the works it cites.
Pragmatic instruction following and goal assistance via cooperative language-guided inverse planning
Tan Zhi-Xuan, Lance Ying, Vikash Mansinghka, and Joshua B Tenenbaum · 2024
Later among the works it cites.
Rehearsal: Simulating conflict to teach conflict resolution
Omar Shaikh, Valentino Emil Chai, Michele Gelfand, Diyi Yang, and Michael S Bernstein · 2024
Later among the works it cites.
Predicting results of social science experiments using large language models
Ashwini Ashokkumar, Luke Hewitt, Isaias Ghezae, and Robb Willer · 2024
Later among the works it cites.
Generative agent simulations of 1,000 people
Joon Sung Park, Carolyn Q Zou, Aaron Shaw, Benjamin Mako Hill, Carrie Cai, Meredith Ringel Morris, Robb Willer, Percy Liang, and Michael S Bernstein · 2024
Later among the works it cites.
Can ai serve as a substitute for human subjects in software engineering research?
Marco Gerosa, Bianca Trinkenreich, Igor Steinmacher, and Anita Sarma · 2024
Later among the works it cites.
Bringing comparative cognition to computers
Konstantinos Voudouris, Matthew Crosby, Marta Halina, and José Hernández-Orallo · 2024
Later among the works it cites.
Principles of animal cognition for LLM evaluations: A case study on transitive inference
Shrestha Rane, Nick Chater, and Mike Oaksford · 2024
Later among the works it cites.
Auxiliary task demands mask the capabilities of smaller language models
Jennifer Hu and Michael C Frank · 2024
Later among the works it cites.
Can language models handle recursively nested grammatical structures? A case study on comparing models and humans
Andrew Lampinen · 2024
Later among the works it cites.
When new experience leads to new knowledge: A computational framework for formalizing epistemically transformative experiences
Joan Danielle K Ongchoco, Isaac M Davis, Julian Jara-Ettinger, and LA Paul · 2024
Later among the works it cites.
The PRISM alignment dataset: What participatory, representative and individualised human feedback reveals about the subjective and multicultural alignment of large language models
Hannah Rose Kirk, Alexander Whitefield, Paul Röttger, Andrew Michael Bean, Katerina Margatina, Rafael Mosquera, Juan Manuel Ciro, Max Bartolo, Adina Williams, He He, Bertie Vidgen, and Scott A. Hale · 2024
Later among the works it cites.
A roadmap to pluralistic alignment
Taylor Sorensen, Jared Moore, Jillian Fisher, Mitchell Gordon, Niloofar Mireshghallah, Christopher Michael Rytting, Andre Ye, Liwei Jiang, Ximing Lu, Nouha Dziri, et al · 2024
Later among the works it cites.
A survey of confidence estimation and calibration in large language models
Jiahui Geng, Fengyu Cai, Yuxia Wang, Heinz Koeppl, Preslav Nakov, and Iryna Gurevych · 2024
Later among the works it cites.
Bayesian models of cognition: reverse engineering the mind
Thomas L Griffiths, Nick Chater, and Joshua B Tenenbaum · 2024
Later among the works it cites.
Evaluating large language models in theory of mind tasks
Michal Kosinski · 2024
Later among the works it cites.
MMToM-QA: Multimodal theory of mind question answering
Chuanyang Jin, Yutong Wu, Jing Cao, Jiannan Xiang, Yen-Ling Kuo, Zhiting Hu, Tomer Ullman, Antonio Torralba, Joshua B Tenenbaum, and Tianmin Shu · 2024
Later among the works it cites.
Pragmatic embodied spoken instruction following in human-robot collaboration with theory of mind
Lance Ying, Xinyi Li, Shivam Aarya, Yizirui Fang, Yifan Yin, Jason Xinyu Liu, Stefanie Tellex, Joshua B Tenenbaum, and Tianmin Shu · 2024
Later among the works it cites.
Larger and more instructable language models become less reliable
Lexin Zhou, Wout Schellaert, Fernando Martínez-Plumed, Yael Moros-Daval, Cèsar Ferri, and José Hernández-Orallo · 2024
Later among the works it cites.
Aurora: A foundation model of the atmosphere
Cristian Bodnar, Wessel P Bruinsma, Ana Lucic, Megan Stanley, Johannes Brandstetter, Patrick Garvan, Maik Riechert, Jonathan Weyn, Haiyu Dong, Anna Vaughan, et al · 2024
Later among the works it cites.
Large language models assume people are more rational than we really are
Ryan Liu, Jiayi Geng, Joshua C Peterson, Ilia Sucholutsky, and Thomas L Griffiths · 2024
Later among the works it cites.
Language models, like humans, show content effects on reasoning tasks
Andrew K Lampinen, Ishita Dasgupta, Stephanie C Y Chan, Antonia Creswell, Dharshan Kumaran, James L McClelland, and Felix Hill · 2024
Later among the works it cites.
How to evaluate the cognitive abilities of LLMs
Anna A Ivanova · 2025
Closest in time.
Measuring what matters: Construct validity in large language model benchmarks
Andrew Bean et al · 2025
Closest in time.
Re-evaluating theory of mind evaluation in large language models
Jennifer Hu, Felipe Sosa, and Tomer Ullman · 2025
Closest in time.
Wilka Carvalho and Andrew Lampinen · 2025
Closest in time.
On the informativeness of supervision signals
Ilia Sucholutsky, Ruairidh M Battleday, Katherine M Collins, Raja Marjieh, and Joshua et al Peterson · 2046
Closest in time.