Fetching the paper…
Reading the bibliography…
Large language models are powerful systems that excel at many tasks, ranging from translation to mathematical reasoning.
Subjective probability: A judgment of representativeness
Daniel Kahneman and Amos Tversky · 1972
Earlier work this paper cites.
On the limited memory bfgs method for large scale optimization
Dong C Liu and Jorge Nocedal · 1989
Earlier work this paper cites.
Unified theories of cognition and the role of soar
Allen Newell · 1992
Earlier work this paper cites.
Decisions from experience and the effect of rare events in risky choice
Ralph Hertwig, Greg Barron, Elke U Weber, and Ido Erev · 2004
Earlier work this paper cites.
Bayesian model selection for group studies—revisited
Lionel Rigoux, Klaas Enno Stephan, Karl J Friston, and Jean Daunizeau · 2014
Earlier work this paper cites.
Humans use directed and random exploration to solve the explore–exploit dilemma
Robert C Wilson, Andra Geana, John M White, Elliot A Ludvig, and Jonathan D Cohen · 2014
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
Sebastian Bach, Alexander Binder, Grégoire Montavon, Frederick Klauschen, Klaus-Robert Müller, and Wojciech Samek · 2015
Earlier work this paper cites.
From anomalies to forecasts: Toward a descriptive model of decisions under risk, under ambiguity, and from experience
Ido Erev, Eyal Ert, Ori Plonsky, Doron Cohen, and Oded Cohen · 2017
Earlier work this paper cites.
Deconstructing the human algorithms for exploration
Samuel J Gershman · 2018
Earlier work this paper cites.
When and how can social scientists add value to data scientists? a choice prediction competition for human decision making
Ori Plonsky, Reut Apel, Ido Erev, Eyal Ert, and Moshe Tennenholtz · 2018
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al · 2019
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf · 2019
Earlier work this paper cites.
The algorithmic architecture of exploration in the human brain
Eric Schulz and Samuel J Gershman · 2019
Earlier work this paper cites.
Task representations in neural networks trained to perform many cognitive tasks
Guangyu Robert Yang, Madhura R Joglekar, H Francis Song, William T Newsome, and Xiao-Jing Wang · 2019
Cited alongside, same era.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Cited alongside, same era.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al · 2021
Cited alongside, same era.
Transformer interpretability beyond attention visualization
Hila Chefer, Shir Gur, and Lior Wolf · 2021
Cited alongside, same era.
The dynamics of explore–exploit decisions reveal a signal-to-noise mechanism for random exploration
Emergent abilities of large language models
Jason Wei, Yi Tay, Rishi Bommasani, Colin Raffel, Barret Zoph, Sebastian Borgeaud, Dani Yogatama, Maarten Bosma, Denny Zhou, Donald Metzler, Ed H. Chi, Tatsunori Hashimoto, Oriol Vinyals, Percy Liang, Jeff Dean, and William Fedus · 2022
Later among the works it cites.
Scaling laws for language encoding models in fmri
Richard Antonello, Aditya Vaidya, and Alexander G Huth · 2023
Closest in time.
Using cognitive psychology to understand gpt-3
Marcel Binz and Eric Schulz · 2023
Closest in time.
Meta-learned models of cognition
Marcel Binz, Ishita Dasgupta, Akshay Jagadish, Matthew Botvinick, Jane X Wang, and Eric Schulz · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Samuel F Feng, Siyu Wang, Sylvia Zarnescu, and Robert C Wilson · 2021
Cited alongside, same era.
Using large-scale experiments and machine learning to discover theories of human decision-making
Joshua C Peterson, David D Bourgin, Mayank Agrawal, Daniel Reichman, and Thomas L Griffiths · 2021
Cited alongside, same era.
The neural architecture of language: Integrative modeling converges on predictive processing
Martin Schrimpf, Idan Asher Blank, Greta Tuckute, Carina Kauf, Eghbal A Hosseini, Nancy Kanwisher, Joshua B Tenenbaum, and Evelina Fedorenko · 2021
Cited alongside, same era.
Language models show human-like content effects on reasoning
Ishita Dasgupta, Andrew K Lampinen, Stephanie CY Chan, Antonia Creswell, Dharshan Kumaran, James L McClelland, and Felix Hill · 2022
Cited alongside, same era.
Machine intuition: Uncovering human-like intuitive decision-making in gpt-3.5
Thilo Hagendorff, Sarah Fabi, and Michal Kosinski · 2022
Cited alongside, same era.
Reconstructing the cascade of language processing in the brain using the internal computations of a transformer-based language model
Sreejan Kumar, Theodore R Sumers, Takateru Yamakoshi, Ariel Goldstein, Uri Hasson, Kenneth A Norman, Thomas L Griffiths, Robert D Hawkins, and Samuel A Nastase · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
A neural model of task compositionality with natural language instructions
Reidar Riveland and Alexandre Pouget · 2022
Cited alongside, same era.
Closest in time.
Inducing anxiety in large language models increases exploration and bias
Julian Coda-Forno, Kristin Witte, Akshay K Jagadish, Marcel Binz, Zeynep Akata, and Eric Schulz · 2023
Closest in time.
Gpts are gpts: An early look at the labor market impact potential of large language models
Tyna Eloundou, Sam Manning, Pamela Mishkin, and Daniel Rock · 2023
Closest in time.
Experiential values are underweighted in decisions involving symbolic options
Basile Garcia, Maël Lebreton, Sacha Bourgeois-Gironde, and Stefano Palminteri · 2023
Closest in time.
Chatgpt for good? on opportunities and challenges of large language models for education
Enkelejda Kasneci, Kathrin Seßler, Stefan Küchemann, Maria Bannert, Daryna Dementieva, Frank Fischer, Urs Gasser, Georg Groh, Stephan Günnemann, Eyke Hüllermeier, et al · 2023
Closest in time.
Ethics of large language models in medicine and medical research
Hanzhou Li, John T Moon, Saptarshi Purkayastha, Leo Anthony Celi, Hari Trivedi, and Judy W Gichoya · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Driving and suppressing the human language network using large language models
Greta Tuckute, Aalok Sathe, Shashank Srikant, Maya Taliaferro, Mingye Wang, Martin Schrimpf, Kendrick Kay, and Evelina Fedorenko · 2023
Closest in time.