Fetching the paper…
Reading the bibliography…
Large language models (LLMs) show increasingly advanced emergent capabilities and are being incorporated across various societal domains.
“Language Models are Few-Shot Learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“Research methods in the behavioral sciences.”
Leon Festinger and Daniel Katz · 1953
Earlier work this paper cites.
“The Nature and History of Experimental Control”
Edwin Boring · 1954
Earlier work this paper cites.
“Statistical Methods for the Behavioral Sciences”
Allen Edwards · 1954
Earlier work this paper cites.
“Aspects of the Theory of Syntax”
Noam Chomsky · 1965
Earlier work this paper cites.
“Collected Papers”
Gilbert Ryle · 1971
Earlier work this paper cites.
“More is different: Broken symmetry and the nature of the hierarchical structure of science”
Philip Anderson · 1972
Earlier work this paper cites.
“The Interpretation of Cultures: Selected Essays”
Clifford Geertz · 1973
Earlier work this paper cites.
“Judgment under Uncertainty: Heuristics and Biases”
Amos Tversky and Daniel Kahneman · 1974
Earlier work this paper cites.
“Minds, brains, and programs”
John Searle · 1980
Earlier work this paper cites.
“The Framing of Decisions and the Psychology of Choice”
Amos Tversky and Daniel Kahneman · 1981
Earlier work this paper cites.
“Beliefs about beliefs: representation and constraining function of wrong beliefs in young children’s understanding of deception”
H. Wimmer and J Perner · 1983
Earlier work this paper cites.
“Pragmatic reasoning schemas”
Patricia Cheng and Keith Holyoak · 1985
Earlier work this paper cites.
“Three-year-olds’ difficulty with false belief: The case for a conceptual deficit”
Josef Perner, Susan. Leekam and Heinz Wimmer · 1987
Earlier work this paper cites.
“Spacing effects and their implications for theory and practice”
Frank. Dempster · 1989
Earlier work this paper cites.
“Sentence comprehension: A parallel distributed processing approach”
Jay McClelland, Mark St. and Roman Taraban · 1989
Earlier work this paper cites.
“Distributed representations, simple recurrent networks, and grammatical structure”
Jeffrey Elman · 1991
Earlier work this paper cites.
“Empiricism and the Philosophy of Mind”
Wilfrid Sellars · 1997
Earlier work this paper cites.
“Thematic Roles Assigned along the Garden Path Linger”
Kiel Christianson, Andrew Hollingworth, John. Halliwell and Fernanda Ferreira · 2001
Earlier work this paper cites.
“Alike performance during nonverbal episodic learning from diversely imprinted neural networks”
Georg Grön, David Schul, Volker Bretschneider, AP Wunderlich and Matthias Riepe · 2003
Earlier work this paper cites.
“Cognitive psychology and self-reports: models and methods”
Jared Jobe · 2003
Earlier work this paper cites.
“COMPARING LEARNING METHODS FOR CLASSIFICATION”
Yuhong Yang · 2006
Earlier work this paper cites.
“Curriculum learning”
Yoshua Bengio, Jérôme Louradour, Ronan Collobert and Jason Weston · 2009
Earlier work this paper cites.
“Reinforcement learning in the brain”
Yael Niv · 2009
Earlier work this paper cites.
“A survey of cross-validation procedures for model selection”
Sylvain Arlot and Alain Celisse · 2010
Earlier work this paper cites.
“Working memory”
Alan Baddeley · 2010
Earlier work this paper cites.
“Heuristic decision making”
Gerd Gigerenzer and Wolfgang Gaissmaier · 2011
Earlier work this paper cites.
“Ecological Rationality: Intelligence in the World”
Peter Todd and Gerd Gigerenzer · 2012
Earlier work this paper cites.
“Model selection and psychological theory: a discussion of the differences between the Akaike information criterion (AIC) and the Bayesian information criterion (BIC)”
Scott. Vrieze · 2012
Earlier work this paper cites.
“A large annotated corpus for learning natural language inference”
Samuel. Bowman, Gabor Angeli, Christopher Potts and Christopher. Manning · 2015
Earlier work this paper cites.
“The benefits of interleaved and blocked study: different tasks benefit from different schedules of study”
Paulo. Carvalho and Robert. Goldstone · 2015
Earlier work this paper cites.
“Estimating the reproducibility of psychological science”
Open Science Collaboration · 2015
Earlier work this paper cites.
“ImageNet Large Scale Visual Recognition Challenge”
Olga Russakovsky et al · 2015
Earlier work this paper cites.
“Could a Neuroscientist Understand a Microprocessor?”
Eric Jonas and Konrad Kording · 2017
Earlier work this paper cites.
“The Stroop Color and Word Test”
Federica Scarpina and Sofia Tagini · 2017
Earlier work this paper cites.
“Working Memory From the Psychological and Neurosciences Perspectives: A Review”
Wen Chai, Aini Abd and Jafri Abdullah · 2018
Earlier work this paper cites.
“Explainable Artificial Intelligence (XAI): Concepts, Taxonomies, Opportunities and Challenges toward Responsible AI”
Alejandro Barredo Arrieta et al · 2019
Earlier work this paper cites.
“Exploring Neural Networks with Activation Atlases”
Shan Carter, Zan Armstrong, Ludwig Schubert, Ian Johnson and Chris Olah · 2019
Earlier work this paper cites.
“Psycholinguistics Meets Continual Learning: Measuring Catastrophic Forgetting in Visual Question Answering”
Claudio Greco, Barbara Plank, Raquel Fernández and Raffaella Bernardi · 2019
Earlier work this paper cites.
“Right for the Wrong Reasons: Diagnosing Syntactic Heuristics in Natural Language Inference”
Tom McCoy, Ellie Pavlick and Tal Linzen · 2019
Earlier work this paper cites.
“Using Priming to Uncover the Organization of Syntactic Representations in Neural Language Models”
Grusha Prasad, Marten Van and Tal Linzen · 2019
Earlier work this paper cites.
“Machine behaviour”
Iyad Rahwan, Manuel Cebrian, Nick Obradovich, Josh Bongard, Jean-François Bonnefon and Cynthia Breazeal · 2019
Earlier work this paper cites.
“Apply rich psychological terms in AI with care”
Henry Shevlin and Marta Halina · 2019
Earlier work this paper cites.
“HellaSwag: Can a Machine Really Finish Your Sentence?”
Rowan Zellers, Ari Holtzman, Yonatan Bisk, Ali Farhadi and Yejin Choi · 2019
Earlier work this paper cites.
“Abstraction and Reasoning Challenge”
François Chollet, Katherine Tong, Walter Reade and Julia Elliott · 2020
Earlier work this paper cites.
“Can GPT-3 Pass a Writer’s Turing Test?”
Katherine Elkins and Jon Chun · 2020
Earlier work this paper cites.
“Performance vs. competence in human–machine comparisons”
Chaz Firestone · 2020
Earlier work this paper cites.
“GPT-3: Its Nature, Scope, Limits, and Consequences”
Luciano Floridi and Massimo Chiriatti · 2020
Earlier work this paper cites.
“Shortcut learning in deep neural networks”
Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis, Richard Zemel, Wieland Brendel, Matthias Bethge and Felix. Wichmann · 2020
Earlier work this paper cites.
“Transparency and reproducibility in artificial intelligence”
Benjamin Haibe-Kains et al · 2020
Earlier work this paper cites.
“What shapes feature representations? Exploring datasets, architectures, and training”
Katherine Hermann and Andrew Lampinen · 2020
Earlier work this paper cites.
“Scaling Laws for Neural Language Models”
Jared Kaplan et al · 2020
Earlier work this paper cites.
“Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks”
Patrick Lewis et al · 2020
Earlier work this paper cites.
“On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?”
Emily Bender, Timnit Gebru, Angelina McMillan-Major and Shmargaret Shmitchell · 2021
Earlier work this paper cites.
“On the opportunities and risks of foundation models”
Rishi Bommasani et al · 2021
Earlier work this paper cites.
“Evaluating Large Language Models Trained on Code”
Mark Chen et al · 2021
Earlier work this paper cites.
“Training Verifiers to Solve Math Word Problems”
Karl Cobbe et al · 2021
Earlier work this paper cites.
“Measuring Mathematical Problem Solving With the MATH Dataset”
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song and Jacob Steinhardt · 2021
Earlier work this paper cites.
“Surface Form Competition: Why the Highest Probability Answer Isn’t Always Right”
Ari Holtzman, Peter West, Vered Shwartz, Yejin Choi and Luke Zettlemoyer · 2021
Earlier work this paper cites.
“Syntactic Structure from Deep Learning”
Tal Linzen and Marco Baroni · 2021
Cited alongside, same era.
“Using large-scale experiments and machine learning to discover theories of human decision-making”
Joshua. Peterson, David. Bourgin, Mayank Agrawal, Daniel Reichman and Thomas. Griffiths · 2021
Cited alongside, same era.
“Causal ImageNet: How to discover spurious features in Deep Learning?”
Sahil Singla and Soheil Feizi · 2021
Cited alongside, same era.
“Calibrate Before Use: Improving Few-Shot Performance of Language Models”
Tony. Zhao, Eric Wallace, Shi Feng, Dan Klein and Sameer Singh · 2021
Cited alongside, same era.
“What learning algorithm is in-context learning? Investigations with linear models”
Ekin Akyürek, Dale Schuurmans, Jacob Andreas, Tengyu Ma and Denny Zhou · 2022
Cited alongside, same era.
“Improving Language Models by Retrieving from Trillions of Tokens”
“Task Contamination: Language Models May Not Be Few-Shot Anymore”
Changmao Li and Jeffrey Flanigan · 2023
Closest in time.
“Embers of Autoregression: Understanding Large Language Models Through the Problem They are Trained to Solve”
R McCoy, Shunyu Yao, Dan Friedman, Matthew Hardy and Thomas Griffiths · 2023
Closest in time.
“Inverse Scaling: When Bigger Isn’t Better”
Ian. McKenzie et al · 2023
Closest in time.
“Augmented Language Models: a Survey”
Grégoire Mialon, Roberto Dessì, Maria Lomeli, Christoforos Nalmpantis, Ram Pasunuru and Roberta Raileanu · 2023
Closest in time.
“Boosting Theory-of-Mind Performance in Large Language Models via Prompting”
Shima Moghaddam and Christopher. Honey · 2023
Closest in time.
“DERA: Enhancing Large Language Model Completions with Dialog-Enabled Resolving Agents”
Varun Nair, Elliot Schumacher, Geoffrey Tso and Anitha Kannan · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai, Eliza Rutherford and Katie Millican · 2022
Cited alongside, same era.
“Data Distributional Properties Drive Emergent In-Context Learning in Transformers”
Stephanie Chan, Adam Santoro, Andrew Lampinen, Jane Wang, Aaditya Singh, Pierre Richemond, James McClelland and Felix Hill · 2022
Cited alongside, same era.
“Language models show human-like content effects on reasoning”
Ishita Dasgupta, Andrew. Lampinen, Stephanie.. Chan, Antonia Creswell, Dharshan Kumaran, James. McClelland and Felix Hill · 2022
Cited alongside, same era.
“Towards artificial general intelligence via a multimodal foundation model”
Nanyi Fei et al · 2022
Cited alongside, same era.
“Towards Reasoning in Large Language Models: A Survey”
Jie Huang and Kevin-Chuan Chang · 2022
Cited alongside, same era.
“Capturing failures of large language models via human cognitive biases”
Erik Jones and Jacob Steinhardt · 2022
Cited alongside, same era.
“Language Models (Mostly) Know What They Know”
Saurav Kadavath et al · 2022
Cited alongside, same era.
Closest in time.
“GPT-4 Technical Report”, 2023
OpenAI · 2023
Closest in time.
“GPT-4V(ision) System Card”, 2023
OpenAI · 2023
Closest in time.
“Transformers learn in-context by gradient descent”
Johannes von Oswald, Eyvind Niklasson, Ettore Randazzo, João Sacramento, Alexander Mordvintsev, Andrey Zhmoginov and Max Vladymyrov · 2023
Closest in time.
“Robust speech recognition via large-scale weak supervision”
Alec Radford, Jong Kim, Tao Xu, Greg Brockman, Christine McLeavey and Ilya Sutskever · 2023
Closest in time.
“The Goldilocks of Pragmatic Understanding: Fine-Tuning Strategy Matters for Implicature Resolution by LLMs”
Laura Ruis, Akbir Khan, Stella Biderman, Sara Hooker, Tim Rocktäschel and Edward Grefenstette · 2023
Closest in time.
“In-Context Impersonation Reveals Large Language Models’ Strengths and Biases”
Leonard Salewski, Stephan Alaniz, Isabel Rio-Torto, Eric Schulz and Zeynep Akata · 2023
Closest in time.
“Are Emergent Abilities of Large Language Models a Mirage?”
Rylan Schaeffer, Brando Miranda and Sanmi Koyejo · 2023
Closest in time.
“Toolformer: Language Models Can Teach Themselves to Use Tools”
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda and Thomas Scialom · 2023
Closest in time.
“Visual cognition in multimodal large language models”
Luca Schulze, Elif Akata, Matthias Bethge and Eric Schulz · 2023
Closest in time.
“Long-form analogies generated by chatGPT lack human-like psycholinguistic properties”
S.. Seals and Valerie. Shalin · 2023
Closest in time.
“Role play with large language models”
Murray Shanahan, Kyle McDonell and Laria Reynolds · 2023
Closest in time.
“A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis”
Alessandro Stolfo, Yonatan Belinkov and Mrinmaya Sachan · 2023
Closest in time.
“Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks”
Tomer Ullman · 2023
Closest in time.
“Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora”
Alex Warstadt et al · 2023
Closest in time.
“Emergent analogical reasoning in large language models”
Taylor Webb, Keith Holyoak and Hongjing Lu · 2023
Closest in time.
“Using Computational Models to Test Syntactic Learnability”
Ethan Wilcox, Richard Futrell and Roger Levy · 2023
Closest in time.
“Transmission Versus Truth, Imitation Versus Innovation: What Children Can Do That Large Language and Language-and-Vision Models Cannot (Yet)”
Eunice Yiu, Eliza Kosoy and Alison Gopnik · 2023
Closest in time.
“Explainability for Large Language Models: A Survey”
Haiyan Zhao, Hanjie Chen, F. Yang, Ninghao Liu, Huiqi Deng, Hengyi Cai, Shuaiqiang Wang, Dawei Yin and Mengnan Du · 2023
Closest in time.
“Mindstorms in Natural Language-Based Societies of Mind”
Mingchen Zhuge et al · 2023
Closest in time.
“Many-shot jailbreaking”, 2024, pp. 1–34
Cem Anil et al · 2024
Closest in time.
“The Claude 3 Model Family: Opus, Sonnet, Haiku”, 2024, pp. 1–42
Anthropic · 2024
Closest in time.
“Language Model Behavior: A Comprehensive Survey”
Tyler Chang and Benjamin Bergen · 2024
Closest in time.
“CogBench: a large language model walks into a psychology lab”
Julian Coda-Forno, Marcel Binz, Jane Wang and Eric Schulz · 2024
Closest in time.
“Experimentology: An Open Science Approach to Experimental Psychology Methods”
Michael. Frank, Mika Braginsky, Julie Cachia, Nicholas Coles, Tom. Hardwicke, Robert. Hawkins, Maya. Mathur and Rondeline Williams · 2024
Closest in time.
“Scaling and evaluating sparse autoencoders”
Leo Gao, Tomé la Tour, Henk Tillman, Gabriel Goh, Rajan Troll, Alec Radford, Ilya Sutskever, Jan Leike and Jeffrey Wu · 2024
Closest in time.
“Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context”
Gemini Team et al · 2024
Closest in time.
“Large Language Model based Multi-Agents: A Survey of Progress and Challenges”
Taicheng Guo, Xiuying Chen, Yaqi Wang, Ruidi Chang, Shichao Pei, Nitesh. Chawla, Olaf Wiest and Xiangliang Zhang · 2024
Closest in time.
“Deception abilities emerged in large language models”
Thilo Hagendorff · 2024
Closest in time.
“Mapping the Ethics of Generative AI: A Comprehensive Scoping Review”
Thilo Hagendorff · 2024
Closest in time.
“Relative Value Biases in Large Language Models”
William Hayes, Nicolas Yax and Stefano Palminteri · 2024
Closest in time.
“Elements of World Knowledge (EWOK): A cognition-inspired framework for evaluating basic world knowledge in language models”
Anna Ivanova et al · 2024
Closest in time.
“Human-like Category Learning by Injecting Ecological Priors from Large Language Models into Neural Networks”
Akshay Jagadish, Julian Coda-Forno, Mirko Thalmann, Eric Schulz and Marcel Binz · 2024
Closest in time.
“Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test”
Aditi Khandelwal, Utkarsh Agarwal, Kumar Tanmay and Monojit Choudhury · 2024
Closest in time.
“How do Large Language Models Navigate Conflicts between Honesty and Helpfulness?”
Ryan Liu, Theodore Sumers, Ishita Dasgupta and Thomas Griffiths · 2024
Closest in time.
“DrEureka: Language Model Guided Sim-To-Real Transfer”
Yecheng Ma, William Liang, Hung-Ju Wang, Sam Wang, Yuke Zhu, Linxi Fan, Osbert Bastani and Dinesh Jayaraman · 2024
Closest in time.
“(Ir)rationality and cognitive biases in large language models”
Olivia Macmillan-Scott and Mirco Musolesi · 2024
Closest in time.
“Dissociating language and thought in large language models”
Kyle Mahowald, Anna Ivanova, Idan Blank, Nancy Kanwisher, Joshua Tenenbaum and Evelina Fedorenko · 2024
Closest in time.
“DataPerf: Benchmarks for Data-Centric AI Development”
Mark Mazumder et al · 2024
Closest in time.
“Can You Learn Semantics Through Next-Word Prediction? The Case of Entailment”
William Merrill, Zhaofeng Wu, Norihito Naka, Yoon Kim and Tal Linzen · 2024
Closest in time.
“Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention”
Tsendsuren Munkhdalai, Manaal Faruqui and Siddharth Gopal · 2024
Closest in time.
“Network Formation and Dynamics Among Multi-LLMs”
Marios Papachristou and Yuan Yuan · 2024
Closest in time.
“The Machine Psychology of Cooperation: Can GPT models operationalise prompts for altruism, cooperation, competitiveness and selfishness in economic games?”
Steve Phelps and Yvan. Russell · 2024
Closest in time.
“The Effect of Sampling Temperature on Problem Solving in Large Language Models”
Matthew Renze and Erhan Guven · 2024
Closest in time.
“Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models”
Paul Röttger, Valentin Hofmann, Valentina Pyatkin, Musashi Hinck, Hannah Kirk, Hinrich Schütze and Dirk Hovy · 2024
Closest in time.
“In-context learning agents are asymmetric belief updaters”
Johannes Schubert, Akshay Jagadish, Marcel Binz and Eric Schulz · 2024
Closest in time.
“Testing theory of mind in large language models and humans”
James.. Strachan et al · 2024
Closest in time.
“LLMs achieve adult human performance on higher-order theory of mind tasks”
Winnie Street et al · 2024
Closest in time.
“Predictions from language models for multiple-choice tasks are not robust under variation of scoring methods”
Polina Tsvilodub, Hening Wang, Sharon Grosch and Michael Franke · 2024
Closest in time.
“Studying and improving reasoning in humans and machines”
Nicolas Yax, Hernan Anlló and Stefano Palminteri · 2024
Closest in time.
“Vision-Language Models for Vision Tasks: A Survey”
Jingyi Zhang, Jiaxing Huang, Sheng Jin and Shijian Lu · 2024
Closest in time.
“Natural Language Embedded Programs for Hybrid Language Symbolic Reasoning”
Tianhua Zhang et al · 2024
Closest in time.
“Improved Few-Shot Jailbreaking Can Circumvent Aligned Language Models and Their Defenses”
Xiaosen Zheng, Tianyu Pang, Chao Du, Qian Liu, Jing Jiang and Min Lin · 2024
Closest in time.