Fetching the paper…
Reading the bibliography…
Do large language models (LLMs) display rational reasoning? LLMs have been shown to contain human biases due to the data they have been trained on; whether this is reflected in rational reasoning remains less clear.
Reasoning
Peter C. Wason · 1966
Earlier work this paper cites.
Subjective probability: A judgment of representativeness
Daniel Kahneman and Amos Tversky · 1972
Earlier work this paper cites.
Judgment under Uncertainty: Heuristics and Biases
Amos Tversky and Daniel Kahneman · 1974
Earlier work this paper cites.
Probabilistic reasoning in clinical medicine: Problems and opportunities
David M. Eddy · 1982
Earlier work this paper cites.
The psychology of preferences
Daniel Kahneman and Amos Tversky · 1982
Earlier work this paper cites.
Extensional versus intuitive reasoning: The conjunction fallacy in probability judgment
Amos Tversky and Daniel Kahneman · 1983
Earlier work this paper cites.
The bounded rationality of probabilistic mental models
Gerd Gigerenzer · 1993
Earlier work this paper cites.
Without Good Reason: The Rationality Debate in Philosophy and Cognitive Science
Edward Stein · 1996
Earlier work this paper cites.
Reasoning the fast and frugal way: models of bounded rationality
Gerd Gigerenzer and Daniel Goldstein · 1996
Earlier work this paper cites.
Monty Hall’s Three Doors: Construction and Deconstruction of a Choice Anomaly
Daniel Friedman · 1998
Earlier work this paper cites.
Rationality and Intelligence: A Brief Update
Stuart Russell · 2016
Earlier work this paper cites.
Machine behaviour
Iyad Rahwan, Manuel Cebrian, Nick Obradovich, Josh Bongard, Jean-François Bonnefon, Cynthia Breazeal, Jacob Crandall, Nicholas Christakis, Iain Couzin, Matthew Jackson, Nicholas Jennings, Ece Kamar, Isabel Kloumann, Hugo Larochelle, David Lazer, Richard McElreath, Alan Mislove, David Parkes, Alex Pentland, and Michael Wellman · 2019
Earlier work this paper cites.
Performance vs. competence in human–machine comparisons
Chaz Firestone · 2020
Earlier work this paper cites.
Language Models are Few-Shot Learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Earlier work this paper cites.
Tversky and Kahneman’s Cognitive Illusions: Who Can Solve Them, and Why?
Georg Bruckmaier, Stefan Krauss, Karin Binder, Sven Hilbert, and Martin Brunner · 2021
Earlier work this paper cites.
Large pre-trained language models contain human-like biases of what is right and wrong to do
Patrick Schramowski, Cigdem Turan, Nico Andersen, Constantin A. Rothkopf, and Kristian Kersting · 2022
Earlier work this paper cites.
Can language models learn from explanations in context?
Andrew Lampinen, Ishita Dasgupta, Stephanie Chan, Kory Mathewson, Mh Tessler, Antonia Creswell, James McClelland, Jane Wang, and Felix Hill · 2022
Cited alongside, same era.
LaMDA: Language Models for Dialog Applications
Romal Thoppilan, Daniel De Freitas, Jamie Hall, Noam Shazeer, Apoorv Kulshreshtha, Heng-Tze Cheng, Alicia Jin, Taylor Bos, Leslie Baker, Yu Du, YaGuang Li, Hongrae Lee, Huaixiu Steven Zheng, Amin Ghafouri, Marcelo Menegali, Yanping Huang, Maxim Krikun, Dmitry Lepikhin, James Qin, Dehao Chen, Yuanzhong Xu, Zhifeng Chen, Adam Roberts, Maarten Bosma, Vincent Zhao, Yanqi Zhou, Chung-Ching Chang, Igor Krivokon, Will Rusch, Marc Pickett, Pranesh Srinivasan, Laichee Man, Kathleen Meier-Hellstern, Meredith Ringel Morris, Tulsee Doshi, Renelito Delos Santos, Toju Duke, Johnny Soraker, Ben Zevenbergen, Vinodkumar Prabhakaran, Mark Diaz, Ben Hutchinson, Kristen Olson, Alejandra Molina, Erin Hoffman-John, Josh Lee, Lora Aroyo, Ravi Rajakumar, Alena Butryna, Matthew Lamm, Viktoriya Kuzmina, Joe Fenton, Aaron Cohen, Rachel Bernstein, Ray Kurzweil, Blaise Aguera-Arcas, Claire Cui, Marian Croak, Ed Chi, and Quoc Le · 2022
Cited alongside, same era.
Language Model Behavior: A Comprehensive Survey
Tyler A. Chang and Benjamin K. Bergen · 2023
Cited alongside, same era.
(Ir)rationality in AI: State of the Art, Research Challenges and Open Questions
Olivia Macmillan-Scott and Mirco Musolesi · 2023
Cited alongside, same era.
Language models show human-like content effects on reasoning tasks
Ishita Dasgupta, Andrew K. Lampinen, Stephanie C. Y. Chan, Hannah R. Sheahan, Antonia Creswell, Dharshan Kumaran, James L. McClelland, and Felix Hill · 2023
Later among the works it cites.
Does ChatGPT have Theory of Mind?
Bart Holterman and Kees van Deemter · 2023
Later among the works it cites.
Exploring the Intersection of Rationality, Reality, and Theory of Mind in AI Reasoning: An Analysis of GPT-4’s Responses to Paradoxes and ToM Tests
Lucas Freund · 2023
Later among the works it cites.
The Emergence of Economic Rationality of GPT
Yiting Chen, Tracy Xiao Liu, You Shan, and Songfa Zhong · 2023
Later among the works it cites.
Emergent analogical reasoning in large language models
Taylor Webb, Keith J. Holyoak, and Hongjing Lu · 2023
Later among the works it cites.
The Goldilocks of Pragmatic Understanding: Fine-Tuning Strategy Matters for Implicature Resolution by LLMs
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Can language models handle recursively nested grammatical structures? a case study on comparing models and humans
Andrew K Lampinen · 2023
Cited alongside, same era.
Human-like intuitive behavior and reasoning biases emerged in large language models but disappeared in ChatGPT
Thilo Hagendorff, Sarah Fabi, and Michal Kosinski · 2023
Cited alongside, same era.
Can AI language models replace human participants?
Danica Dillion, Niket Tandon, Yuling Gu, and Kurt Gray · 2023
Cited alongside, same era.
AI language models cannot replace human research participants
Jacqueline Harding, William D’Alessandro, N. G. Laskowski, and Robert Long · 2023
Cited alongside, same era.
Whose Opinions Do Language Models Reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee, Percy Liang, and Tatsunori Hashimoto · 2023
Cited alongside, same era.
In-Context Impersonation Reveals Large Language Models’ Strengths and Biases
Leonard Salewski, Stephan Alaniz, Isabel Rio-Torto, Eric Schulz, and Zeynep Akata · 2023
Cited alongside, same era.
Generative Agents: Interactive Simulacra of Human Behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein · 2023
Cited alongside, same era.
Large language models show human-like content biases in transmission chain experiments
Alberto Acerbi and Joseph M. Stubbersfield · 2023
Cited alongside, same era.
Laura Ruis, Akbir Khan, Stella Biderman, Sara Hooker, Tim Rocktäschel, and Edward Grefenstette · 2023
Later among the works it cites.
Sparks of Artificial General Intelligence: Early experiments with GPT-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, Harsha Nori, Hamid Palangi, Marco Tulio Ribeiro, and Yi Zhang · 2023
Later among the works it cites.
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
Itay Itzhak, Gabriel Stanovsky, Nir Rosenfeld, and Yonatan Belinkov · 2023
Later among the works it cites.
GPT-4 Technical Report
OpenAI · 2023
Later among the works it cites.
Model Card and Evaluations for Claude Models
Anthropic · 2023
Later among the works it cites.
XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
Paul Röttger, Hannah Rose Kirk, Bertie Vidgen, Giuseppe Attanasio, Federico Bianchi, and Dirk Hovy · 2023
Later among the works it cites.
Large language models in medicine
Arun James Thirunavukarasu, Darren Shu Jeng Ting, Kabilan Elangovan, Laura Gutierrez, Ting Fang Tan, and Daniel Shu Wei Ting · 2023
Later among the works it cites.
Inductive reasoning in humans and large language models
Simon Jerome Han, Keith J. Ransom, Andrew Perfors, and Charles Kemp · 2024
Closest in time.
Escalation Risks from Language Models in Military and Diplomatic Decision-Making
Juan-Pablo Rivera, Gabriel Mukobi, Anka Reuel, Max Lamparth, Chandler Smith, and Jacquelyn Schneider · 2024
Closest in time.
How AI Could Revolutionize Diplomacy
Andrew Moore · 2024
Closest in time.