Fetching the paper…
Reading the bibliography…
In order for AI systems to communicate effectively with people, they must understand how we make decisions.
The theory of decision making
Ward Edwards. 1954 · 1954
Earlier work this paper cites.
An economic theory of political action in a democracy
Anthony Downs. 1957 · 1957
Earlier work this paper cites.
Individual choice behavior: A theoretical analysis
R. Duncan Luce. 1959 · 1959
Earlier work this paper cites.
Judgment under Uncertainty: Heuristics and Biases: Biases in judgments reveal some heuristics of thinking under uncertainty
Amos Tversky and Daniel Kahneman. 1974 · 1974
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff. 1978 · 1978
Earlier work this paper cites.
Prospect theory: An analysis of decision under risk
Daniel Kahneman and Amos Tversky. 1979 · 1979
Earlier work this paper cites.
The Strategy of Conflict: with a new Preface by the Author
Thomas C Schelling. 1980 · 1980
Earlier work this paper cites.
Foundations of social theory
James S Coleman. 1994 · 1994
Earlier work this paper cites.
Affect, empathy, and regressive mispredictions of others’ preferences under risk
David Faro and Yuval Rottenstreich. 2006 · 2006
Earlier work this paper cites.
Action understanding as inverse planning
Chris L Baker, Rebecca Saxe, and Joshua B Tenenbaum. 2009 · 2009
Earlier work this paper cites.
Legibility and predictability of robot motion. In 2013 8th ACM/IEEE International Conference on Human-Robot Interaction (HRI) . IEEE, 301–308
Anca D Dragan, Kenton CT Lee, and Siddhartha S Srinivasa. 2013 · 2013
Earlier work this paper cites.
Democracy, bureaucracy and public choice: Economic approaches in political science
Patrick Dunleavy. 2014 · 2014
Earlier work this paper cites.
The child as econometrician: A rational model of preference understanding in children
Christopher G Lucas, Thomas L Griffiths, Fei Xu, Christine Fawcett, Alison Gopnik, Tamar Kushnir, Lori Markson, and Jane Hu. 2014 · 2014
Earlier work this paper cites.
Learning the preferences of ignorant, inconsistent agents
Owain Evans, Andreas Stuhlmueller, and Noah D. Goodman. 2015 · 2015
Earlier work this paper cites.
Planning for autonomous cars that leverage effects on human actions.. In Robotics: Science and systems
Dorsa Sadigh, Shankar Sastry, Sanjit A Seshia, and Anca D Dragan. 2016 · 2016
Earlier work this paper cites.
Rational quantitative attribution of beliefs, desires and percepts in human mentalizing
Chris L Baker, Julian Jara-Ettinger, Rebecca Saxe, and Joshua B Tenenbaum. 2017 · 2017
Earlier work this paper cites.
People learn other people’s preferences through inverse decision-making
Alan Jern, Christopher G Lucas, and Charles Kemp. 2017 · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Earlier work this paper cites.
Cognitive model priors for predicting human decisions. In Proceedings of the 36th International Conference on Machine Learning . 5133–5141
David D. Bourgin, Joshua C. Peterson, Daniel Reichman, Stuart J. Russell, and Thomas L. Griffiths. 2019 · 2019
Earlier work this paper cites.
The naive utility calculus as a unified, quantitative framework for action understanding
Julian Jara-Ettinger, Laura E Schulz, and Joshua B Tenenbaum. 2020 · 2020
Earlier work this paper cites.
Resource-rational analysis: Understanding human cognition as the optimal use of limited computational resources
Falk Lieder and Thomas L Griffiths. 2020 · 2020
Earlier work this paper cites.
Literal or pedagogic human? Analyzing human model misspecification in objective learning. In Proceedings of The 35th Uncertainty in Artificial Intelligence Conference
Smitha Milli and Anca D. Dragan. 2020 · 2020
Earlier work this paper cites.
Modeling the mistakes of boundedly rational agents within a Bayesian theory of mind
Arwa Alanqary, Gloria Z. Lin, Joie Le, Tan Zhi-Xuan, Vikash K. Mansinghka, and Joshua B. Tenenbaum. 2021 · 2021
Earlier work this paper cites.
A general language assistant as a laboratory for alignment
Amanda Askell, Yuntao Bai, Anna Chen, Dawn Drain, Deep Ganguli, Tom Henighan, Andy Jones, Nicholas Joseph, Ben Mann, Nova DasSarma, et al · 2021
Earlier work this paper cites.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al · 2021
Earlier work this paper cites.
Mark K. Ho and Thomas L. Griffiths. 2021 · 2021
Cited alongside, same era.
Using large-scale experiments and machine learning to discover theories of human decision-making
Joshua C. Peterson, David D. Bourgin, Mayank Agrawal, Daniel Reichman, and Thomas L. Griffiths. 2021 · 2021
Cited alongside, same era.
Using large language models to simulate multiple humans and replicate human subject studies. In International Conference on Machine Learning
Gati Aher, Rosa I. Arriaga, and Adam Tauman Kalai. 2022 · 2022
Cited alongside, same era.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Y Bai et al · 2022
Cited alongside, same era.
An ontology of decision models
Lisheng He, Wenjia Joyce Zhao, and Sudeep Bhatia. 2022 · 2022
Cited alongside, same era.
Direct preference optimization: Your language model is secretly a reward model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Stefano Ermon, Christopher D. Manning, and Chelsea Finn. 2023 · 2023
Later among the works it cites.
Melanie Sclar, Yejin Choi, Yulia Tsvetkov, and Alane Suhr. 2023 · 2023
Later among the works it cites.
Enhancing human persuasion with large language models
Minkyu Shin and Jin Kim. 2023 · 2023
Later among the works it cites.
Large language models fail on trivial alterations to theory-of-mind tasks
Tomer Ullman. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Training language models to follow instructions with human feedback
Long Ouyang et al · 2022
Cited alongside, same era.
Social simulacra: Creating populated prototypes for social computing systems. In Proceedings of the 35th Annual ACM Symposium on User Interface Software and Technology . 1–18
Joon Sung Park, Lindsay Popowski, Carrie Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2022 · 2022
Cited alongside, same era.
Neural theory-of-mind? on the limits of social intelligence in large LMs
Maarten Sap, Ronan LeBras, Daniel Fried, and Yejin Choi. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Cited alongside, same era.
Out of one, many: Using language models to simulate human samples
Lisa P Argyle, Ethan C Busby, Nancy Fulda, Joshua R Gubler, Christopher Rytting, and David Wingate. 2023 · 2023
Cited alongside, same era.
Turning large language models into cognitive models
Marcel Binz and Eric Schulz. 2023a · 2023
Cited alongside, same era.
Using cognitive psychology to understand GPT-3
Marcel Binz and Eric Schulz. 2023b · 2023
Cited alongside, same era.
Peiyi Wang, Lei Li, Liang Chen, Zefan Cai, Dawei Zhu, Binghuai Lin, Yunbo Cao, Qi Liu, Tianyu Liu, and Zhifang Sui. 2023 · 2023
Later among the works it cites.
Inferring the goals of communicating agents from actions and instructions. In Proceedings of the AAAI Symposium Series , Vol. 2. 26–33
Lance Ying, Tan Zhi-Xuan, Vikash Mansinghka, and Joshua B Tenenbaum. 2023 · 2023
Later among the works it cites.
Foundational challenges in assuring alignment and safety of large language models
Usman Anwar, Abulhair Saparov, Javier Rando, Daniel Paleka, Miles Turpin, Peter Hase, Ekdeep Singh Lubana, Erik Jenner, Stephen Casper, Oliver Sourbut, et al · 2024
Closest in time.
Measuring implicit bias in explicitly unbiased large language models
Xuechunzi Bai, Angelina Wang, Ilia Sucholutsky, and Thomas L Griffiths. 2024 · 2024
Closest in time.
CogBench: A large language model walks into a psychology lab
Julian Coda-Forno, Marcel Binz, Jane X Wang, and Eric Schulz. 2024 · 2024
Closest in time.
LLM-driven imitation of subrational behavior: Illusion or Reality?
Andrea Coletta, Kshama Dwarakanath, Penghang Liu, Svitlana Vyetrenko, and Tucker Balch. 2024 · 2024
Closest in time.
Understanding social reasoning in language models with language models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman. 2024 · 2024
Closest in time.
Teaching large language models to reason with reinforcement learning
Alex Havrilla, Yuqing Du, Sharath Chandra Raparthy, Christoforos Nalmpantis, Jane Dwivedi-Yu, Maksym Zhuravinskyi, Eric Hambro, Sainbayar Sukhbaatar, and Roberta Raileanu. 2024 · 2024
Closest in time.
Quantifying the Persona Effect in LLM Simulations
Tiancheng Hu and Nigel Collier. 2024 · 2024
Closest in time.
A survey on large language model hallucination via a creativity perspective
Xuhui Jiang, Yuxing Tian, Fengrui Hua, Chengjin Xu, Yuanzhuo Wang, and Jian Guo. 2024 · 2024
Closest in time.
Inna Wanyin Lin, Ashish Sharma, Christopher Michael Rytting, Adam S Miner, Jina Suh, and Tim Althoff. 2024 · 2024
Closest in time.
Ryan Liu, Jiayi Geng, Addison J Wu, Ilia Sucholutsky, Tania Lombrozo, and Thomas L Griffiths. 2024a · 2024
Closest in time.
(Ir) rationality and cognitive biases in large language models
Olivia Macmillan-Scott and Mirco Musolesi. 2024 · 2024
Closest in time.
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement. In International Conference on Learning Representations
Linlu Qiu, Liwei Jiang, Ximing Lu, Melanie Sclar, Valentina Pyatkin, Chandra Bhagavatula, Bailin Wang, Yoon Kim, Yejin Choi, Nouha Dziri, et al · 2024
Closest in time.
In-context impersonation reveals large language models’ strengths and biases
Leonard Salewski, Stephan Alaniz, Isabel Rio-Torto, Eric Schulz, and Zeynep Akata. 2024 · 2024
Closest in time.
Rehearsal: Simulating conflict to teach conflict resolution
Omar Shaikh, Valentino Chai, Michele J. Gelfand, Diyi Yang, and Michael S. Bernstein. 2024 · 2024
Closest in time.
Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models. In Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers) . 2257–2273
Natalie Shapira, Mosh Levy, Seyed Hossein Alavi, Xuhui Zhou, Yejin Choi, Yoav Goldberg, Maarten Sap, and Vered Shwartz. 2024 · 2024
Closest in time.
Seungjong Sun, Eungu Lee, Dongyan Nan, Xiangying Zhao, Wonbyung Lee, Bernard J. Jansen, and Jang Hyun Kim. 2024 · 2024
Closest in time.
Pragmatic instruction following and goal assistance via cooperative language-guided inverse planning
Tan Zhi-Xuan, Lance Ying, Vikash Mansinghka, and Joshua B Tenenbaum. 2024 · 2024
Closest in time.
LIMA: Less is more for alignment
Chunting Zhou, Pengfei Liu, Puxin Xu, Srinivasan Iyer, Jiao Sun, Yuning Mao, Xuezhe Ma, Avia Efrat, Ping Yu, Lili Yu, et al · 2024
Closest in time.
Xuhui Zhou, Zhe Su, Tiwalayo Eisape, Hyunwoo Kim, and Maarten Sap. 2024b · 2024
Closest in time.