Fetching the paper…
Reading the bibliography…
This chapter critically examines the potential contributions of modern language models to theoretical linguistics.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
A solution to Plato’s problem: The latent semantic analysis theory of acquisition, induction, and representation of knowledge
Thomas K. Landauer and Susan T. Dumais · 1939
Earlier work this paper cites.
Philosophy of cognitive science in the age of deep learning
Raphaël Millière · 1939
Earlier work this paper cites.
The nature and measurement of meaning
Charles E. Osgood · 1939
Earlier work this paper cites.
Linguistic evidence and grammatical theory
Carson T. Schütze · 1939
Earlier work this paper cites.
A mathematical theory of communication
C. E. Shannon · 1948
Earlier work this paper cites.
Prediction and Entropy of Printed English
C. E. Shannon · 1951
Earlier work this paper cites.
Distributional Structure
Zellig S. Harris · 1954
Earlier work this paper cites.
Translation
Warren Weaver · 1955
Earlier work this paper cites.
Syntactic Structures
Noam Chomsky · 1957
Earlier work this paper cites.
A synopsis of linguistic theory, 1930-1955
John R. Firth · 1957
Earlier work this paper cites.
On certain formal properties of grammars
Noam Chomsky · 1959
Earlier work this paper cites.
The Structure of a Semantic Theory
Jerrold J. Katz and Jerry A. Fodor · 1963
Earlier work this paper cites.
Aspects of the Theory of Syntax
Noam Chomsky · 1965
Earlier work this paper cites.
Language identification in the limit
E Mark Gold · 1967
Earlier work this paper cites.
Constraints on variables in syntax
John Robert Ross · 1967
Earlier work this paper cites.
Procedures as a Representation for Data in a Computer Program for Understanding Natural Language
Terry Winograd · 1971
Earlier work this paper cites.
Progress in natural language understanding: An application to lunar geology
W. A. Woods · 1973
Earlier work this paper cites.
A vector space model for automatic indexing
G. Salton, A. Wong, and C. S. Yang · 1975
Earlier work this paper cites.
Montague Grammar, Mental Representations, and Reality
Barbara Hall Partee · 1981
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Jerry A. Fodor and Zenon W. Pylyshyn · 1988
Earlier work this paper cites.
On language and connectionism: Analysis of a parallel distributed processing model of language acquisition
Steven Pinker and Alan Prince · 1988
Earlier work this paper cites.
Indexing by latent semantic analysis
Scott Deerwester, Susan T. Dumais, George W. Furnas, Thomas K. Landauer, and Richard Harshman · 1990
Earlier work this paper cites.
Distributed representations, simple recurrent networks, and grammatical structure
Jeffrey L. Elman · 1991
Earlier work this paper cites.
The Minimalist Program
Noam Chomsky · 1995
Earlier work this paper cites.
Rethinking Innateness: A Connectionist Perspective on Development
Jeffrey L. Elman, Elizabeth A. Bates, Mark H. Johnson, and Annette Karmiloff-Smith · 1996
Earlier work this paper cites.
Rethinking Eliminative Connectionism
Gary F. Marcus · 1998
Earlier work this paper cites.
Toward a connectionist model of recursion in human linguistic performance
Morten H. Christiansen and Nick Chater · 1999
Earlier work this paper cites.
Riau Indonesian as a Pivotless Language
David Gil · 1999
Earlier work this paper cites.
A Neural Probabilistic Language Model
Yoshua Bengio, Réjean Ducharme, and Pascal Vincent · 2000
Earlier work this paper cites.
Large language models are better than theoretical linguists at theoretical linguistics
Ben Ambridge and Liam Blything · 2002
Earlier work this paper cites.
The Faculty of Language: What Is It, Who Has It, and How Did It Evolve?
Marc D. Hauser, Noam Chomsky, and W. Tecumseh Fitch · 2002
Earlier work this paper cites.
Language input and child syntax
Janellen Huttenlocher, Marina Vasilyeva, Elina Cymerman, and Susan Levine · 2002
Earlier work this paper cites.
Empirical assessment of stimulus poverty arguments
Geoffrey K. Pullum and Barbara C. Scholz · 2002
Earlier work this paper cites.
The faculty of language: What’s special about it?
Steven Pinker and Ray Jackendoff · 2004
Earlier work this paper cites.
Universal Grammar, statistics or both?
Charles D. Yang · 2004
Earlier work this paper cites.
Simpler Syntax
Peter W. Culicover and Ray Jackendoff · 2005
Earlier work this paper cites.
Expectation-based syntactic comprehension
Roger Levy · 2007
Earlier work this paper cites.
Representational similarity analysis - connecting the branches of systems neuroscience
Nikolaus Kriegeskorte, Marieke Mur, and Peter Bandettini · 2008
Earlier work this paper cites.
Constructing a Language
Michael Tomasello · 2009
Earlier work this paper cites.
Linguistic Nativism and the Poverty of the Stimulus
Alexander Clark and Shalom Lappin · 2010
Earlier work this paper cites.
Poverty of the Stimulus Revisited
Robert C. Berwick, Paul Pietroski, Beracah Yankama, and Noam Chomsky · 2011
Earlier work this paper cites.
Montague Meets Markov: Deep Semantics with Probabilistic Logical Form
Islam Beltagy, Cuong Chau, Gemma Boleda, Dan Garrette, Katrin Erk, and Raymond Mooney · 2013
Earlier work this paper cites.
Intensionality was only alleged: On adjective-noun composition in distributional semantics
Gemma Boleda, Marco Baroni, The Nghia Pham, and Louise McNally · 2013
Earlier work this paper cites.
Morgan’s Canon, meet Hume’s Dictum: Avoiding anthropofabulation in cross-species comparisons
Cameron Buckner · 2013
Earlier work this paper cites.
Towards a semantics for distributional representations
Katrin Erk · 2013
Earlier work this paper cites.
Efficient Estimation of Word Representations in Vector Space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Frege in Space: A Program for Composition Distributional Semantics
Marco Baroni, Raffaella Bernardi, and Roberto Zamparelli · 2014
Earlier work this paper cites.
How the Tiger Bush Got its Stripes: ‘How Possibly’ vs. ‘How Actually’ Model Explanations
Alisa Bokulich · 2014
Earlier work this paper cites.
Temporal Analysis of Language through Neural Language Models
Yoon Kim, Yi-I Chiu, Kentaro Hanaki, Darshan Hegde, and Slav Petrov · 2014
Earlier work this paper cites.
Empiricism and Language Learnability
Nick Chater, Alexander Clark, John A. Goldsmith, and Amy Perfors · 2015
Earlier work this paper cites.
Learnability
Alexander Clark · 2015
Earlier work this paper cites.
Explanation in Linguistics
Paul Egré · 2015
Earlier work this paper cites.
Structures, Not Strings: Linguistics as Part of the Cognitive Sciences
Martin B. H. Everaert, Marinus A. C. Huybregts, Noam Chomsky, Robert C. Berwick, and Johan J. Bolhuis · 2015
Earlier work this paper cites.
Model-Based Cognitive Neuroscience: A Conceptual Introduction
Birte U. Forstmann and Eric-Jan Wagenmakers · 2015
Earlier work this paper cites.
Building a shared world: Mapping distributional to model-theoretic semantic spaces
Aurélie Herbelot and Eva Maria Vecchi · 2015
Earlier work this paper cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Earlier work this paper cites.
Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg · 2016
Earlier work this paper cites.
Creating Language: Integrating Evolution, Acquisition, and Processing
Morten H. Christiansen and Nick Chater · 2016
Earlier work this paper cites.
Diachronic Word Embeddings Reveal Statistical Laws of Semantic Change
William L. Hamilton, Jure Leskovec, and Dan Jurafsky · 2016
Earlier work this paper cites.
Exploring socioeconomic differences in syntactic development through the lens of real-time processing
Yi Ting Huang, Kathryn Leech, and Meredith L. Rowe · 2016
Earlier work this paper cites.
Assessing the Ability of LSTMs to Learn Syntax-Sensitive Dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg · 2016
Earlier work this paper cites.
Impossible Languages
Andrea Moro · 2016
Earlier work this paper cites.
Does String-Based Neural MT Learn Source Syntax?
Xing Shi, Inkit Padhi, and Kevin Knight · 2016
Earlier work this paper cites.
Probing Classifiers: Promises, Shortcomings, and Advances
Yonatan Belinkov · 2017
Earlier work this paper cites.
Formal distributional semantics: Introduction to the special issue
Gemma Boleda and Aurélie Herbelot · 2017
Earlier work this paper cites.
Conceptual Versus Referential Affordance in Concept Composition
Louise McNally and Gemma Boleda · 2017
Earlier work this paper cites.
On Chomsky and the Two Cultures of Statistical Learning
Peter Norvig · 2017
Earlier work this paper cites.
Attention is All you Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Understanding intermediate layers using linear classifier probes, November 2018
Guillaume Alain and Yoshua Bengio · 2018
Earlier work this paper cites.
RNN Simulations of Grammaticality Judgments on Long-distance Dependencies
Shammur Absar Chowdhury and Roberto Zamparelli · 2018
Earlier work this paper cites.
What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, German Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Under the Hood: Using Diagnostic Classifiers to Investigate and Improve how Language Models Track Agreement Information
Mario Giulianelli, Jack Harding, Florian Mohnert, Dieuwke Hupkes, and Willem Zuidema · 2018
Earlier work this paper cites.
Colorless Green Recurrent Networks Dream Hierarchically
Kristina Gulordava, Piotr Bojanowski, Edouard Grave, Tal Linzen, and Marco Baroni · 2018
Earlier work this paper cites.
Visualisation and ’Diagnostic Classifiers’ Reveal How Recurrent and Recursive Neural Networks Process Hierarchical Structure
Dieuwke Hupkes, Sara Veldhoen, and Willem Zuidema · 2018
Earlier work this paper cites.
Generalization without Systematicity: On the Compositional Skills of Sequence-to-Sequence Recurrent Networks
Brenden Lake and Marco Baroni · 2018
Earlier work this paper cites.
Distributional Models of Word Meaning
Alessandro Lenci · 2018
Earlier work this paper cites.
Targeted Syntactic Evaluation of Language Models
Rebecca Marvin and Tal Linzen · 2018
Earlier work this paper cites.
An Analysis of Encoder Representations in Transformer-Based Machine Translation
Alessandro Raganato and Jörg Tiedemann · 2018
Earlier work this paper cites.
Can LSTM Learn to Capture Agreement? The Case of Basque
Shauli Ravfogel, Yoav Goldberg, and Francis Tyers · 2018
Cited alongside, same era.
Representation in Cognitive Science
Nicholas Shea · 2018
Cited alongside, same era.
Acceptability judgments and grammaticality, prospects and challenges
Jon Sprouse · 2018
Cited alongside, same era.
How could models possibly provide how-possibly explanations?
Philippe Verreault-Julien · 2018
Cited alongside, same era.
What do RNN Language Models Learn about Filler–Gap Dependencies?
Ethan Gotlieb Wilcox, Roger Levy, Takashi Morita, and Richard Futrell · 2018
Cited alongside, same era.
What Do North American Babies Hear? A large-scale cross-corpus analysis
Elika Bergelson, Marisa Casillas, Melanie Soderstrom, Amanda Seidl, Anne S. Warlaumont, and Andrei Amatuni · 2019
Cited alongside, same era.
Characterizing Intrinsic Compositionality in Transformers with Tree Projections, November 2022
Shikhar Murty, Pratyusha Sharma, Jacob Andreas, and Christopher D. Manning · 2022
Later among the works it cites.
Progress measures for grokking via mechanistic interpretability
Neel Nanda, Lawrence Chan, Tom Lieberum, Jess Smith, and Jacob Steinhardt · 2022
Later among the works it cites.
Scientific Representation
James Nguyen and Roman Frigg · 2022
Later among the works it cites.
Does BERT Rediscover a Classical NLP Pipeline?
Jingcheng Niu, Wenjie Lu, and Gerald Penn · 2022
Later among the works it cites.
Making Transformers Solve Compositional Tasks
Santiago Ontanon, Joshua Ainslie, Zachary Fisher, and Vaclav Cvicek · 2022
Later among the works it cites.
Semantic Structure in Deep Learning
Ellie Pavlick · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep learning: A philosophical introduction
Cameron Buckner · 2019
Cited alongside, same era.
Correlating Neural and Symbolic Representations of Language
Grzegorz Chrupała and Afra Alishahi · 2019
Cited alongside, same era.
Deep Neural Networks as Scientific Models
Radoslaw M. Cichy and Daniel Kaiser · 2019
Cited alongside, same era.
What Does BERT Look at? An Analysis of BERT’s Attention
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D. Manning · 2019
Cited alongside, same era.
Child-Directed Speech Is Infrequent in a Forager-Farmer Population: A Time Allocation Study
Alejandrina Cristia, Emmanuel Dupoux, Michael Gurven, and Jonathan Stieglitz · 2019
Cited alongside, same era.
Neural language models as psycholinguistic subjects: Representations of syntactic state
Richard Futrell, Ethan Wilcox, Takashi Morita, Peng Qian, Miguel Ballesteros, and Roger Levy · 2019
Cited alongside, same era.
Meaning without reference in large language models, August 2022
Steven Piantadosi and Felix Hill · 2022
Later among the works it cites.
Improving Compositional Generalization with Latent Structure and Data Augmentation
Linlu Qiu, Peter Shaw, Panupong Pasupat, Pawel Nowak, Tal Linzen, Fei Sha, and Kristina Toutanova · 2022
Later among the works it cites.
The Importance of Understanding Deep Learning
Tim Räz and Claus Beisbart · 2022
Later among the works it cites.
Large-Scale Evidence for Logarithmic Effects of Word Predictability on Reading Time, November 2022
Cory Shain, Clara Meister, Tiago Pimentel, Ryan Cotterell, and Roger Philip Levy · 2022
Later among the works it cites.
Understanding models understanding language
Anders Søgaard · 2022
Later among the works it cites.
Understanding from Machine Learning Models
Emily Sullivan · 2022
Later among the works it cites.
Interpretability in the Wild: A Circuit for Indirect Object Identification in GPT-2 small
Kevin Ro Wang, Alexandre Variengien, Arthur Conmy, Buck Shlegeris, and Jacob Steinhardt · 2022
Later among the works it cites.
What Artificial Neural Networks Can Tell Us about Human Language Acquisition
Alex Warstadt and Samuel R. Bowman · 2022
Later among the works it cites.
LexSym: Compositionality as Lexical Symmetry
Ekin Akyurek and Jacob Andreas · 2023
Later among the works it cites.
Physics of Language Models: Part 1, Context-Free Grammar, May 2023
Zeyuan Allen-Zhu and Yuanzhi Li · 2023
Later among the works it cites.
Large Linguistic Models: Analyzing theoretical linguistic abilities of LLMs, May 2023
Gašper Beguš, Maksymilian Dąbkowski, and Ryan Rhodes · 2023
Later among the works it cites.
Simplicity Bias in Transformers and their Ability to Learn Sparse Boolean Functions, July 2023
Satwik Bhattamishra, Arkil Patel, Varun Kanade, and Phil Blunsom · 2023
Later among the works it cites.
What are large language models supposed to model?
Idan A. Blank · 2023
Later among the works it cites.
Not all layers are equally as important: Every Layer Counts BERT
Lucas Georges Gabriel Charpentier and David Samuel · 2023
Later among the works it cites.
Sudden Drops in the Loss: Syntax Acquisition, Phase Transitions, and Simplicity Bias in MLMs, September 2023
Angelica Chen, Ravid Schwartz-Ziv, Kyunghyun Cho, Matthew L. Leavitt, and Naomi Saphra · 2023
Later among the works it cites.
Noam Chomsky: The False Promise of ChatGPT
Noam Chomsky, Ian Roberts, and Jeffrey Watumull · 2023
Later among the works it cites.
Large Language Models Demonstrate the Potential of Statistical Learning in Language
Pablo Contreras Kallens, Ross Deans Kristensen-McLachlan, and Morten H. Christiansen · 2023
Later among the works it cites.
Large language models and (non-)linguistic recursion, June 2023
Maksymilian Dąbkowski and Gašper Beguš · 2023
Later among the works it cites.
Systematic testing of three Language Models reveals low language accuracy, absence of response stability, and a yes-response bias
Vittoria Dentella, Fritz Günther, and Evelina Leivada · 2023
Later among the works it cites.
The neuroconnectionist research programme
Adrien Doerig, Rowan P. Sommers, Katja Seeliger, Blake Richards, Jenann Ismael, Grace W. Lindsay, Konrad P. Kording, Talia Konkle, Marcel A. J. van Gerven, Nikolaus Kriegeskorte, and Tim C. Kietzmann · 2023
Later among the works it cites.
Compositionality in Computational Linguistics
Lucia Donatelli and Alexander Koller · 2023
Later among the works it cites.
TinyStories: How Small Can Language Models Be and Still Speak Coherent English?, May 2023
Ronen Eldan and Yuanzhi Li · 2023
Later among the works it cites.
Language acquisition: Do children and language models follow similar learning stages?, June 2023
Linnea Evanson, Yair Lakretz, and Jean-Rémi King · 2023
Later among the works it cites.
Verb Conjugation in Transformers Is Determined by Linear Encodings of Subject Number, October 2023
Sophie Hao and Tal Linzen · 2023
Later among the works it cites.
Operationalising Representation in Natural Language Processing, June 2023
Jacqueline Harding · 2023
Later among the works it cites.
Prompt-based methods may underestimate large language models’ linguistic generalizations, May 2023
Jennifer Hu and Roger Levy · 2023
Later among the works it cites.
Why large language models are poor theories of human linguistic cognition. A reply to Piantadosi (2023)., April 2023
Roni Katzir · 2023
Later among the works it cites.
Human-like systematic generalization through a meta-learning neural network
Brenden M. Lake and Marco Baroni · 2023
Later among the works it cites.
Can language models handle recursively nested grammatical structures? A case study on comparing models and humans, February 2023
Andrew Kyle Lampinen · 2023
Later among the works it cites.
Modeling rapid language learning by distilling Bayesian priors into artificial neural networks, May 2023
R. Thomas McCoy and Thomas L. Griffiths · 2023
Later among the works it cites.
The Vector Grounding Problem, April 2023
Dimitri Coelho Mollo and Raphaël Millière · 2023
Later among the works it cites.
Large languages, impossible languages and human brains
Andrea Moro, Matteo Greco, and Stefano F. Cappa · 2023
Later among the works it cites.
How to Plant Trees in Language Models: Data and Architectural Effects on the Emergence of Syntactic Inductive Biases, May 2023
Aaron Mueller and Tal Linzen · 2023
Later among the works it cites.
Grokking of Hierarchical Structure in Vanilla Transformers, May 2023
Shikhar Murty, Pratyusha Sharma, Jacob Andreas, and Christopher D. Manning · 2023
Later among the works it cites.
Transformer-Based LM Surprisal Predicts Human Reading Times Best with About Two Billion Training Tokens, April 2023
Byung-Doh Oh and William Schuler · 2023
Later among the works it cites.
GPT-4 Technical Report, March 2023
OpenAI · 2023
Later among the works it cites.
Pretrain on just structure: Understanding linguistic inductive biases using transfer learning, April 2023
Isabel Papadimitriou and Dan Jurafsky · 2023
Later among the works it cites.
Modeling syntactic acquisition
Lisa S. Pearl · 2023
Later among the works it cites.
Modern language models refute Chomsky’s approach to language, March 2023
Steven Piantadosi · 2023
Later among the works it cites.
Characterizing English Preposing in PP constructions, August 2023
Christopher Potts · 2023
Later among the works it cites.
The Best Game in Town: The Re-Emergence of the Language of Thought Hypothesis Across the Cognitive Sciences
Jake Quilty-Dunn, Nicolas Porot, and Eric Mandelbaum · 2023
Later among the works it cites.
Can AI-Generated Text be Reliably Detected?, June 2023
Vinu Sankar Sadasivan, Aounon Kumar, Sriram Balasubramanian, Wenxiao Wang, and Soheil Feizi · 2023
Later among the works it cites.
Trained on 100 million words and still in shape: BERT meets British National Corpus
David Samuel, Andrey Kutuzov, Lilja Øvrelid, and Erik Velldal · 2023
Later among the works it cites.
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R. Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, Agnieszka Kluska, Aitor Lewkowycz, Akshat Agarwal, Alethea Power, Alex Ray, Alex Warstadt, Alexander W. Kocurek, Ali Safaya, Ali Tazarv, Alice Xiang, Alicia Parrish, Allen Nie, Aman Hussain, Amanda Askell, Amanda Dsouza, Ambrose Slone, Ameet Rahane, Anantharaman S. Iyer, Anders Johan Andreassen, Andrea Madotto, Andrea Santilli, Andreas Stuhlmüller, Andrew M. Dai, Andrew La, Andrew Lampinen, Andy Zou, Angela Jiang, Angelica Chen, Anh Vuong, Animesh Gupta, Anna Gottardi, Antonio Norelli, Anu Venkatesh, Arash Gholamidavoodi, Arfa Tabassum, Arul Menezes, Arun Kirubarajan, Asher Mullokandov, Ashish Sabharwal, Austin Herrick, Avia Efrat, Aykut Erdem, Ayla Karakaş, B. Ryan Roberts, Bao Sheng Loe, Barret Zoph, Bartłomiej Bojanowski, Batuhan Özyurt, Behnam Hedayatnia, Behnam Neyshabur, Benjamin Inden, Benno Stein, Berk Ekmekci, Bill Yuchen Lin, Blake Howald, Bryan Orinion, Cameron Diao, Cameron Dour, Catherine Stinson, Cedrick Argueta, Cesar Ferri, Chandan Singh, Charles Rathkopf, Chenlin Meng, Chitta Baral, Chiyu Wu, Chris Callison-Burch, Christopher Waites, Christian Voigt, Christopher D. Manning, Christopher Potts, Cindy Ramirez, Clara E. Rivera, Clemencia Siro, Colin Raffel, Courtney Ashcraft, Cristina Garbacea, Damien Sileo, Dan Garrette, Dan Hendrycks, Dan Kilman, Dan Roth, C. Daniel Freeman, Daniel Khashabi, Daniel Levy, Daniel Moseguí González, Danielle Perszyk, Danny Hernandez, Danqi Chen, Daphne Ippolito, Dar Gilboa, David Dohan, David Drakard, David Jurgens, Debajyoti Datta, Deep Ganguli, Denis Emelin, Denis Kleyko, Deniz Yuret, Derek Chen, Derek Tam, Dieuwke Hupkes, Diganta Misra, Dilyar Buzan, Dimitri Coelho Mollo, Diyi Yang, Dong-Ho Lee, Dylan Schrader, Ekaterina Shutova, Ekin Dogus Cubuk, Elad Segal, Eleanor Hagerman, Elizabeth Barnes, Elizabeth Donoway, Ellie Pavlick, Emanuele Rodolà, Emma Lam, Eric Chu, Eric Tang, Erkut Erdem, Ernie Chang, Ethan A. Chi, Ethan Dyer, Ethan Jerzak, Ethan Kim, Eunice Engefu Manyasi, Evgenii Zheltonozhskii, Fanyue Xia, Fatemeh Siar, Fernando Martínez-Plumed, Francesca Happé, Francois Chollet, Frieda Rong, Gaurav Mishra, Genta Indra Winata, Gerard de Melo, Germán Kruszewski, Giambattista Parascandolo, Giorgio Mariani, Gloria Xinyue Wang, Gonzalo Jaimovitch-Lopez, Gregor Betz, Guy Gur-Ari, Hana Galijasevic, Hannah Kim, Hannah Rashkin, Hannaneh Hajishirzi, Harsh Mehta, Hayden Bogar, Henry Francis Anthony Shevlin, Hinrich Schuetze, Hiromu Yakura, Hongming Zhang, Hugh Mee Wong, Ian Ng, Isaac Noble, Jaap Jumelet, Jack Geissinger, Jackson Kernion, Jacob Hilton, Jaehoon Lee, Jaime Fernández Fisac, James B. Simon, James Koppel, James Zheng, James Zou, Jan Kocon, Jana Thompson, Janelle Wingfield, Jared Kaplan, Jarema Radom, Jascha Sohl-Dickstein, Jason Phang, Jason Wei, Jason Yosinski, Jekaterina Novikova, Jelle Bosscher, Jennifer Marsh, Jeremy Kim, Jeroen Taal, Jesse Engel, Jesujoba Alabi, Jiacheng Xu, Jiaming Song, Jillian Tang, Joan Waweru, John Burden, John Miller, John U. Balis, Jonathan Batchelder, Jonathan Berant, Jörg Frohberg, Jos Rozen, Jose Hernandez-Orallo, Joseph Boudeman, Joseph Guerr, Joseph Jones, Joshua B. Tenenbaum, Joshua S. Rule, Joyce Chua, Kamil Kanclerz, Karen Livescu, Karl Krauth, Karthik Gopalakrishnan, Katerina Ignatyeva, Katja Markert, Kaustubh Dhole, Kevin Gimpel, Kevin Omondi, Kory Wallace Mathewson, Kristen Chiafullo, Ksenia Shkaruta, Kumar Shridhar, Kyle McDonell, Kyle Richardson, Laria Reynolds, Leo Gao, Li Zhang, Liam Dugan, Lianhui Qin, Lidia Contreras-Ochando, Louis-Philippe Morency, Luca Moschella, Lucas Lam, Lucy Noble, Ludwig Schmidt, Luheng He, Luis Oliveros-Colón, Luke Metz, Lütfi Kerem Senel, Maarten Bosma, Maarten Sap, Maartje Ter Hoeve, Maheen Farooqi, Manaal Faruqui, Mantas Mazeika, Marco Baturan, Marco Marelli, Marco Maru, Maria Jose Ramirez-Quintana, Marie Tolkiehn, Mario Giulianelli, Martha Lewis, Martin Potthast, Matthew L. Leavitt, Matthias Hagen, Mátyás Schubert, Medina Orduna Baitemirova, Melody Arnaud, Melvin McElrath, Michael Andrew Yee, Michael Cohen, Michael Gu, Michael Ivanitskiy, Michael Starritt, Michael Strube, Michał Swędrowski, Michele Bevilacqua, Michihiro Yasunaga, Mihir Kale, Mike Cain, Mimee Xu, Mirac Suzgun, Mitch Walker, Mo Tiwari, Mohit Bansal, Moin Aminnaseri, Mor Geva, Mozhdeh Gheini, Mukund Varma T, Nanyun Peng, Nathan Andrew Chi, Nayeon Lee, Neta Gur-Ari Krakover, Nicholas Cameron, Nicholas Roberts, Nick Doiron, Nicole Martinez, Nikita Nangia, Niklas Deckers, Niklas Muennighoff, Nitish Shirish Keskar, Niveditha S. Iyer, Noah Constant, Noah Fiedel, Nuan Wen, Oliver Zhang, Omar Agha, Omar Elbaghdadi, Omer Levy, Owain Evans, Pablo Antonio Moreno Casares, Parth Doshi, Pascale Fung, Paul Pu Liang, Paul Vicol, Pegah Alipoormolabashi, Peiyuan Liao, Percy Liang, Peter W. Chang, Peter Eckersley, Phu Mon Htut, Pinyu Hwang, Piotr Miłkowski, Piyush Patil, Pouya Pezeshkpour, Priti Oli, Qiaozhu Mei, Qing Lyu, Qinlang Chen, Rabin Banjade, Rachel Etta Rudolph, Raefer Gabriel, Rahel Habacker, Ramon Risco, Raphaël Millière, Rhythm Garg, Richard Barnes, Rif A. Saurous, Riku Arakawa, Robbe Raymaekers, Robert Frank, Rohan Sikand, Roman Novak, Roman Sitelew, Ronan Le Bras, Rosanne Liu, Rowan Jacobs, Rui Zhang, Russ Salakhutdinov, Ryan Andrew Chi, Seungjae Ryan Lee, Ryan Stovall, Ryan Teehan, Rylan Yang, Sahib Singh, Saif M. Mohammad, Sajant Anand, Sam Dillavou, Sam Shleifer, Sam Wiseman, Samuel Gruetter, Samuel R. Bowman, Samuel Stern Schoenholz, Sanghyun Han, Sanjeev Kwatra, Sarah A. Rous, Sarik Ghazarian, Sayan Ghosh, Sean Casey, Sebastian Bischoff, Sebastian Gehrmann, Sebastian Schuster, Sepideh Sadeghi, Shadi Hamdan, Sharon Zhou, Shashank Srivastava, Sherry Shi, Shikhar Singh, Shima Asaadi, Shixiang Shane Gu, Shubh Pachchigar, Shubham Toshniwal, Shyam Upadhyay, Shyamolima Shammie Debnath, Siamak Shakeri, Simon Thormeyer, Simone Melzi, Siva Reddy, Sneha Priscilla Makini, Soo-Hwan Lee, Spencer Torene, Sriharsha Hatwar, Stanislas Dehaene, Stefan Divic, Stefano Ermon, Stella Biderman, Stephanie Lin, Stephen Prasad, Steven Piantadosi, Stuart Shieber, Summer Misherghi, Svetlana Kiritchenko, Swaroop Mishra, Tal Linzen, Tal Schuster, Tao Li, Tao Yu, Tariq Ali, Tatsunori Hashimoto, Te-Lin Wu, Théo Desbordes, Theodore Rothschild, Thomas Phan, Tianle Wang, Tiberius Nkinyili, Timo Schick, Timofei Kornev, Titus Tunduny, Tobias Gerstenberg, Trenton Chang, Trishala Neeraj, Tushar Khot, Tyler Shultz, Uri Shaham, Vedant Misra, Vera Demberg, Victoria Nyamai, Vikas Raunak, Vinay Venkatesh Ramasesh, Vinay Uday Prabhu, Vishakh Padmakumar, Vivek Srikumar, William Fedus, William Saunders, William Zhang, Wout Vossen, Xiang Ren, Xiaoyu Tong, Xinran Zhao, Xinyi Wu, Xudong Shen, Yadollah Yaghoobzadeh, Yair Lakretz, Yangqiu Song, Yasaman Bahri, Yejin Choi, Yichi Yang, Yiding Hao, Yifu Chen, Yonatan Belinkov, Yu Hou, Yufang Hou, Yuntao Bai, Zachary Seid, Zhuoye Zhao, Zijian Wang, Zijie J. Wang, Zirui Wang, and Ziyi Wu · 2023
Later among the works it cites.
Finding Structure in One Child’s Linguistic Experience
Wentao Wang, Wai Keen Vong, Najoung Kim, and Brenden M. Lake · 2023
Later among the works it cites.
Call for Papers – The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus, January 2023
Alex Warstadt, Leshem Choshen, Aaron Mueller, Adina Williams, Ethan Gotlieb Wilcox, and Chengxu Zhuang · 2023
Later among the works it cites.
Are Deep Neural Networks Adequate Behavioral Models of Human Visual Perception?
Felix A. Wichmann and Robert Geirhos · 2023
Later among the works it cites.
Using Computational Models to Test Syntactic Learnability
Ethan Gotlieb Wilcox, Richard Futrell, and Roger Levy · 2023
Later among the works it cites.
Causal interventions expose implicit situation models for commonsense language understanding, June 2023
Takateru Yamakoshi, James L. McClelland, Adele E. Goldberg, and Robert D. Hawkins · 2023
Later among the works it cites.
How poor is the stimulus? Evaluating hierarchical generalization in neural networks trained on child-directed speech, June 2023
Aditya Yedetore, Tal Linzen, Robert Frank, and R. Thomas McCoy · 2023
Later among the works it cites.
Do Transformers Parse while Predicting the Masked Word?, March 2023
Haoyu Zhao, Abhishek Panigrahi, Rong Ge, and Sanjeev Arora · 2023
Later among the works it cites.
The Clock and the Pizza: Two Stories in Mechanistic Explanation of Neural Networks, June 2023
Ziqian Zhong, Ziming Liu, Max Tegmark, and Jacob Andreas · 2023
Later among the works it cites.
Basic syntax from speech: Spontaneous concatenation in unsupervised deep neural networks, July 2024
Gašper Beguš, Thomas Lu, and Zili Wang · 2024
Closest in time.
Cognitive Plausibility in Natural Language Processing
Lisa Beinborn and Nora Hollenstein · 2024
Closest in time.
The difficulty and importance of estimating the lower and upper bounds of infant speech exposure, June 2024
Joseph Coffey, Okko Räsänen, Camila Scaff, and Alejandrina Cristia · 2024
Closest in time.
What Can Language Models Tell Us About Human Cognition?
Louise Connell and Dermot Lynott · 2024
Closest in time.
Language in Vivo vs. in Silico: Size Matters but Larger Language Models Still Do Not Comprehend Language on a Par with Humans, April 2024
Vittoria Dentella, Fritz Guenther, and Evelina Leivada · 2024
Closest in time.
Auxiliary task demands mask the capabilities of smaller language models, April 2024
Jennifer Hu and Michael C. Frank · 2024
Closest in time.
Language models align with human judgments on key grammatical constructions, January 2024
Jennifer Hu, Kyle Mahowald, Gary Lupyan, Anna Ivanova, and Roger Levy · 2024
Closest in time.
Mission: Impossible Language Models, January 2024
Julie Kallini, Isabel Papadimitriou, Richard Futrell, Kyle Mahowald, and Christopher Potts · 2024
Closest in time.
Evaluating the Language Abilities of Large Language Models vs. Humans: Three Caveats
Evelina Leivada, Vittoria Dentella, and Fritz Günther · 2024
Closest in time.
Testing learning hypotheses using neural networks by manipulating learning data, July 2024
Cara Su-Yi Leong and Tal Linzen · 2024
Closest in time.
Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models, March 2024
Samuel Marks, Can Rager, Eric J. Michaud, Yonatan Belinkov, David Bau, and Aaron Mueller · 2024
Closest in time.
A Philosophical Introduction to Language Models – Part I: Continuity With Classic Debates, January 2024
Raphaël Millière and Cameron Buckner · 2024
Closest in time.
Anthropocentric bias and the possibility of artificial cognition, July 2024
Raphaël Millière and Charles Rathkopf · 2024
Closest in time.
Language Models Learn Rare Phenomena from Less Rare Phenomena: The Case of the Missing AANNs, March 2024
Kanishka Misra and Kyle Mahowald · 2024
Closest in time.
The Philosophy of Theoretical Linguistics: A Contemporary Outlook
Ryan M. Nefdt · 2024
Closest in time.
Filtered Corpus Training (FiCT) Shows that Language Models can Generalize from Indirect Evidence, May 2024
Abhinav Patil, Jaap Jumelet, Yu Ying Chiu, Andy Lapastora, Peter Shen, Lexie Wang, Clevis Willrich, and Shane Steinert-Threlkeld · 2024
Closest in time.
A systematic investigation of learnability from single child linguistic input
Yulu Qin, Wentao Wang, and Brenden Lake · 2024
Closest in time.
Language in brains, minds, and machines
Greta Tuckute, Nancy Kanwisher, and Evelina Fedorenko · 2024
Closest in time.
Insights from the first BabyLM Challenge: Training sample-efficient language models on a developmentally plausible corpus
Alex Warstadt, Aaron Mueller, Leshem Choshen, Ethan Gotlieb Wilcox, Chengxu Zhuang, Adina Williams, Ryan Cotterell, and Tal Linzen · 2024
Closest in time.
Bigger is not always better: The importance of human-scale language modeling for psycholinguistics, July 2024
Ethan Gotlieb Wilcox, Michael Hu, Aaron Mueller, Tal Linzen, Alex Warstadt, Leshem Choshen, Chengxu Zhuang, Ryan Cotterell, and Adina Williams · 2024
Closest in time.
A framework for rigorous evaluation of human performance in human and machine learning comparison studies
Hannah P. Cowley, Mandy Natter, Karla Gray-Roncal, Rebecca E. Rhodes, Erik C. Johnson, Nathan Drenkow, Timothy M. Shead, Frances S. Chance, Brock Wester, and William Gray-Roncal · 2045
Closest in time.