Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have shown an impressive ability to perform tasks believed to require thought processes.
A spreading-activation theory of semantic processing
Allan M Collins and Elizabeth F Loftus · 1975
Earlier work this paper cites.
A spreading activation theory of memory
John R. Anderson · 1983
Earlier work this paper cites.
Mental Models: Towards a Cognitive Science of Language, Inference, and Consciousness
P.N. Johnson-Laird · 1983
Earlier work this paper cites.
Theories of priming: I. associative distance and lag
Timothy P McNamara · 1992
Earlier work this paper cites.
Understanding similarity: A joint project for psychology, case-based reasoning, and law
Ulrike Hahn and Nick Chater · 1998
Earlier work this paper cites.
Models of working memory: Mechanisms of active maintenance and executive control
Akira Miyake and Priti Shah · 1999
Earlier work this paper cites.
Baddeley a. working memory: looking back and looking forward. nat rev neurosci 4: 829-839
Alan Baddeley · 2003
Earlier work this paper cites.
Introduction to experimental pragmatics
Dan Sperber and Ira Noveck · 2004
Earlier work this paper cites.
The Cambridge Handbook of Thinking and Reasoning
Keith Holyoak and Robert Morrison · 2005
Earlier work this paper cites.
Bayesian Rationality: The probabilistic approach to human reasoning
Mike Oaksford and Nick Chater · 2007
Earlier work this paper cites.
Remembering words in context as predicted by an associative read-out model
Markus J. Hofmann, Lars Kuchinke, Chris Biemann, Sascha Tamm, and Arthur M. Jacobs · 2011
Earlier work this paper cites.
Subtracting “ought” from “is”: Descriptivism versus normativism in the study of human thinking
Shira Elqayam and Jonathan St. B. T. Evans · 2011
Cited alongside, same era.
The ‘whys’ and ‘whens’ of individual differences in thinking biases
Wim De Neys and Jean-François Bonnefon · 2013
Cited alongside, same era.
What makes us think? a three-stage dual-process model of analytic engagement
Gordon Pennycook, Jonathan Fugelsang, and Derek Koehler · 2015
Cited alongside, same era.
The semantic distance task: Quantifying semantic distance with semantic network path length
Yoed N Kenett, Effi Levi, David Anaki, and Miriam Faust · 2017
Cited alongside, same era.
Investigating gender bias in language models using causal mediation analysis
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov, Sharon Qian, Daniel Nevo, Yaron Singer, and Stuart Shieber · 2020
Cited alongside, same era.
interpreting gpt: the logit lens, 2020
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Later among the works it cites.
Measuring and narrowing the compositionality gap in language models, 2023
Ofir Press, Muru Zhang, Sewon Min, Ludwig Schmidt, Noah A. Smith, and Mike Lewis · 2023
Later among the works it cites.
Mansi Sakarvadia, Aswathy Ajith, Arham Khan, Daniel Grzenda, Nathaniel Hudson, André Bauer, Kyle Chard, and Ian Foster · 2023
Later among the works it cites.
Dissecting recall of factual associations in auto-regressive language models
Mor Geva, Jasmijn Bastings, Katja Filippova, and Amir Globerson · 2023
Later among the works it cites.
Faithful explanations of black-box nlp models using llm-generated counterfactuals
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
nostalgebraist · 2020
Cited alongside, same era.
Causal abstractions of neural networks
Atticus Geiger, Hanson Lu, Thomas Icard, and Christopher Potts · 2021
Cited alongside, same era.
A mathematical framework for transformer circuits
Nelson Elhage, Neel Nanda, Catherine Olsson, Tom Henighan, Nicholas Joseph, Ben Mann, Amanda Askell, Yuntao Bai, Anna Chen, Tom Conerly, et al · 2021
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Cited alongside, same era.
Transformer feed-forward layers build predictions by promoting concepts in the vocabulary space
Mor Geva, Avi Caciularu, Kevin Wang, and Yoav Goldberg · 2022
Cited alongside, same era.
Sparks of artificial general intelligence: Early experiments with gpt-4. arxiv
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al · 2023
Cited alongside, same era.
What makes us think? a three-stage dual-process model of analytic engagement
Gordon Pennycook, Jonathan A Fugelsang, and Derek J Koehler
Cited in the paper.
Yair Ori Gat, Nitay Calderon, Amir Feder, Alexander Chapanin, Amit Sharma, and Roi Reichart · 2023
Later among the works it cites.
Mistral 7b, 2023
Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, Lélio Renard Lavaud, Marie-Anne Lachaux, Pierre Stock, Teven Le Scao, Thibaut Lavril, Thomas Wang, Timothée Lacroix, and William El Sayed · 2023
Later among the works it cites.
Do large language models latently perform multi-hop reasoning?, 2024
Sohee Yang, Elena Gribovskaya, Nora Kassner, Mor Geva, and Sebastian Riedel · 2024
Closest in time.
Understanding and patching compositional reasoning in llms, 2024
Zhaoyi Li, Gangwei Jiang, Hong Xie, Linqi Song, Defu Lian, and Ying Wei · 2024
Closest in time.
Interpretability at scale: Identifying causal mechanisms in alpaca
Zhengxuan Wu, Atticus Geiger, Thomas Icard, Christopher Potts, and Noah Goodman · 2024
Closest in time.
Selfie: Self-interpretation of large language model embeddings, 2024
Haozhe Chen, Carl Vondrick, and Chengzhi Mao · 2024
Closest in time.
Llama 3 model card
AI@Meta · 2024
Closest in time.