Fetching the paper…
Reading the bibliography…
This paper examines the question of whether Large Language Models (LLMs) like ChatGPT possess minds, focusing specifically on whether they have a genuine folk psychology encompassing beliefs, desires, and intentions.
The Language of Thought
Jerry A. Fodor · 1975
Earlier work this paper cites.
Towards a causal theory of linguistic representation
Dennis W. Stampe · 1977
Earlier work this paper cites.
The Intentional Stance
Daniel Clement Dennett · 1981
Earlier work this paper cites.
Knowledge and the Flow of Information
Fred I. Dretske · 1981
Earlier work this paper cites.
Inquiries Into Truth and Interpretation
Donald Davidson · 1984
Earlier work this paper cites.
Language, Thought, and Other Biological Categories: New Foundations for Realism
Ruth Garrett Millikan · 1984
Earlier work this paper cites.
Inquiry
Robert C. Stalnaker · 1984
Earlier work this paper cites.
Psychosemantics: The Problem of Meaning in the Philosophy of Mind
Jerry A. Fodor · 1987
Earlier work this paper cites.
The status of content
Paul A. Boghossian · 1990
Earlier work this paper cites.
The symbol grounding problem
Stevan Harnad · 1990
Earlier work this paper cites.
Some revisionary proposals about belief and believing
Ruth Barcan Marcus · 1990
Earlier work this paper cites.
A statistical referential theory of content: Using information theory to account for misrepresentation
Marius Usher · 2001
Earlier work this paper cites.
Assessing the causal structure of function
Sergio E Chaigneau, Lawrence W Barsalou, and Steven A Sloman · 2004
Earlier work this paper cites.
Notes toward a structuralist theory of mental representation
Jonathan Opie and Gerard O’Brien · 2004
Earlier work this paper cites.
Consciousness and Mind
David M. Rosenthal · 2005
Earlier work this paper cites.
Consciousness and intentionality
George Graham, Terence E. Horgan, and John L. Tienson · 2007
Earlier work this paper cites.
Toward an informational teleosemantics
Karen Neander · 2013
Earlier work this paper cites.
Desire-fulfillment theory
Chris Heathwood · 2016
Earlier work this paper cites.
Take and took, gaggle and goose, book and read: Evaluating the utility of vector differences for lexical relation learning
Ekaterina Vylomova, Laura Rimell, Trevor Cohn, and Timothy Baldwin · 2016
Earlier work this paper cites.
A Mark of the Mental: A Defence of Informational Teleosemantics
Karen Neander · 2017
Earlier work this paper cites.
Understanding intermediate layers using linear classifier probes, 2018
Guillaume Alain and Yoshua Bengio · 2018
Earlier work this paper cites.
If you can’t change what you believe, you don’t believe it
Grace Helton · 2018
Earlier work this paper cites.
Jeff Da and Jungo Kasai · 2019
Earlier work this paper cites.
Do neural language representations learn physical commonsense?
Maxwell Forbes, Ari Holtzman, and Yejin Choi · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, Alexander H Miller, and Sebastian Riedel · 2019
Cited alongside, same era.
Human compatible: AI and the problem of control
Stuart Russell · 2019
Cited alongside, same era.
Climbing towards NLU: On meaning, form, and understanding in the age of data
Emily M. Bender and Alexander Koller · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Leace: Perfect linear concept erasure in closed form, 2023
Nora Belrose, David Schneider-Joseph, Shauli Ravfogel, Ryan Cotterell, Edward Raff, and Stella Biderman · 2023
Later among the works it cites.
Deep reinforcement learning from human preferences, 2023
Paul Christiano, Jan Leike, Tom B. Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2023
Later among the works it cites.
Palm-e: An embodied multimodal language model, 2023
Danny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Vuong, Tianhe Yu, Wenlong Huang, Yevgen Chebotar, Pierre Sermanet, Daniel Duckworth, Sergey Levine, Vincent Vanhoucke, Karol Hausman, Marc Toussaint, Klaus Greff, Andy Zeng, Igor Mordatch, and Pete Florence · 2023
Later among the works it cites.
Human-like systematic generalization through a meta-learning neural network
Brenden M Lake and Marco Baroni · 2023
Later among the works it cites.
Matthew Mandelkern and Tal Linzen · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
What bert is not: Lessons from a new suite of psycholinguistic diagnostics for language models, 2020
Allyson Ettinger · 2020
Cited alongside, same era.
Negated and misprimed probes for pretrained language models: Birds can talk, but cannot fly, 2020
Nora Kassner and Hinrich Schütze · 2020
Cited alongside, same era.
Null it out: Guarding protected attributes by iterative nullspace projection, 2020
Shauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg · 2020
Cited alongside, same era.
Can language models encode perceptual structure without grounding? a case study in color, 2021
Mostafa Abdou, Artur Kulmizev, Daniel Hershcovich, Stella Frank, Ellie Pavlick, and Anders Søgaard · 2021
Cited alongside, same era.
Causal Theories of Mental Content
Fred Adams and Ken Aizawa · 2021
Cited alongside, same era.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Cited alongside, same era.
Evaluating large language models trained on code, 2021
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba · 2021
Cited alongside, same era.
Later among the works it cites.
Locating and editing factual associations in gpt, 2023
Kevin Meng, David Bau, Alex Andonian, and Yonatan Belinkov · 2023
Later among the works it cites.
The vector grounding problem, 2023
Dimitri Coelho Mollo and Raphaël Millière · 2023
Later among the works it cites.
Testing causal models of word meaning in gpt-3 and-4
Sam Musker and Ellie Pavlick · 2023
Later among the works it cites.
Progress measures for grokking via mechanistic interpretability
Neel Nanda, Lawrence Chan, Tom Lieberum, Jess Smith, and Jacob Steinhardt · 2023
Later among the works it cites.
Mental Content
Peter Schulte · 2023
Later among the works it cites.
Role play with large language models
Murray Shanahan, Kyle McDonell, and Laria Reynolds · 2023
Later among the works it cites.
Attention is all you need, 2023
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2023
Later among the works it cites.
The clock and the pizza: Two stories in mechanistic explanation of neural networks, 2023
Ziqian Zhong, Ziming Liu, Max Tegmark, and Jacob Andreas · 2023
Later among the works it cites.
The reversal curse: Llms trained on "a is b" fail to learn "b is a", 2024
Lukas Berglund, Meg Tong, Max Kaufmann, Mikita Balesni, Asa Cooper Stickland, Tomasz Korbak, and Owain Evans · 2024
Closest in time.
Discovering latent knowledge in language models without supervision, 2024
Collin Burns, Haotian Ye, Dan Klein, and Jacob Steinhardt · 2024
Closest in time.
Language models represent space and time, 2024
Wes Gurnee and Max Tegmark · 2024
Closest in time.
Standards for belief representations in llms, 2024
Daniel A. Herrmann and Benjamin A. Levinstein · 2024
Closest in time.
Are language models more like libraries or like librarians? bibliotechnism, the novel reference problem, and the attitudes of llms, 2024
Harvey Lederman and Kyle Mahowald · 2024
Closest in time.
Still no lie detector for language models: Probing empirical and conceptual roadblocks
Benjamin A Levinstein and Daniel A Herrmann · 2024
Closest in time.
A turing test of whether ai chatbots are behaviorally similar to humans
Qiaozhu Mei, Yutong Xie, Walter Yuan, and Matthew O Jackson · 2024
Closest in time.
Evaluating the world model implicit in a generative model, 2024
Keyon Vafa, Justin Y. Chen, Jon Kleinberg, Sendhil Mullainathan, and Ashesh Rambachan · 2024
Closest in time.
From task structures to world models: what do llms know?
Ilker Yildirim and LA Paul · 2024
Closest in time.