Fetching the paper…
Reading the bibliography…
Language models learn rare syntactic phenomena, but the extent to which this is attributable to generalization vs.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Syntactic Structures
N. Chomsky. 1957 · 1957
Earlier work this paper cites.
Aspects of the Theory of Syntax
N. Chomsky. 1965 · 1965
Earlier work this paper cites.
Knowledge of language: Its nature, origin, and use
N. Chomsky. 1986 · 1986
Earlier work this paper cites.
Category-based Induction
Daniel N Osherson, Edward E Smith, Ormond Wilkie, Alejandro Lopez, and Eldar Shafir. 1990 · 1990
Earlier work this paper cites.
Regular morphology and the lexicon
Joan Bybee. 1995 · 1995
Earlier work this paper cites.
Constructions: A Construction Grammar Approach to Argument Structure
Adele E Goldberg. 1995 · 1995
Earlier work this paper cites.
The CHILDES project: Tools for analyzing talk
B. MacWhinney. 2000 · 2000
Earlier work this paper cites.
Dialogue act modeling for automatic tagging and recognition of conversational speech
Andreas Stolcke, Klaus Ries, Noah Coccaro, Elizabeth Shriberg, Rebecca Bates, Daniel Jurafsky, Paul Taylor, Rachel Martin, Carol Van Ess-Dykema, and Marie Meteer. 2000 · 2000
Earlier work this paper cites.
Empirical assessment of stimulus poverty arguments
Geoffrey K Pullum and Barbara C Scholz. 2002 · 2002
Earlier work this paper cites.
Constructions at Work: The Nature of Generalization in Language
Adele E Goldberg. 2005 · 2005
Earlier work this paper cites.
On the syntax of quantity in English
Richard S Kayne. 2007 · 2007
Earlier work this paper cites.
Two types of modified cardinals
Stephanie Solt. 2007 · 2007
Earlier work this paper cites.
Word learning as Bayesian inference
Fei Xu and Joshua B Tenenbaum. 2007 · 2007
Earlier work this paper cites.
Corpus linguistics in morphology: Morphological productivity
R Harald Baayen. 2009 · 2009
Earlier work this paper cites.
KenLM: Faster and smaller language model queries
Kenneth Heafield. 2011 · 2011
Earlier work this paper cites.
Stubborn Distributivity, Multiparticipant Nouns and the Count/Mass Distinction
Roger Schwarzschild. 2011 · 2011
Earlier work this paper cites.
The partial productivity of constructions as induction
Laura Suttle and Adele E Goldberg. 2011 · 2011
Earlier work this paper cites.
Large-scale syntactic language modeling with treelets
Adam Pauls and Dan Klein. 2012 · 2012
Earlier work this paper cites.
“A pleasant three days in Philadelphia”: Arguments for a pseudopartitive analysis
Caitlin Keenan. 2013 · 2013
Earlier work this paper cites.
The AMARA corpus: Building parallel language resources for the educational domain
Ahmed Abdelali, Francisco Guzman, Hassan Sajjad, and Stephan Vogel. 2014 · 2014
Earlier work this paper cites.
Productivity and reuse in language: A theory of linguistic computation and storage
Timothy J O’Donnell. 2015 · 2015
Earlier work this paper cites.
The Goldilocks Principle: Reading Children’s Books with Explicit Memory Representations
Felix Hill, Antoine Bordes, Sumit Chopra, and Jason Weston. 2016 · 2016
Earlier work this paper cites.
Assessing the ability of LSTMs to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
Earlier work this paper cites.
OpenSubtitles2016: Extracting large parallel corpora from movie and TV subtitles
Pierre Lison and Jörg Tiedemann. 2016 · 2016
Cited alongside, same era.
Grammaticality, acceptability, and probability: A probabilistic view of linguistic knowledge
Jey Han Lau, Alexander Clark, and Shalom Lappin. 2017 · 2017
Cited alongside, same era.
Theory, data, and the epistemology of syntax
Geoffrey K Pullum. 2017 · 2017
Cited alongside, same era.
What do RNN language models learn about filler–gap dependencies?
Ethan Wilcox, Roger Levy, Takashi Morita, and Richard Futrell. 2018 · 2018
Cited alongside, same era.
An amazing four doctoral dissertations
Mary Dalrymple and Tracy Holloway King. 2019 · 2019
Cited alongside, same era.
Neural language models as psycholinguistic subjects: Representations of syntactic state
Richard Futrell, Ethan Wilcox, Takashi Morita, Peng Qian, Miguel Ballesteros, and Roger Levy. 2019 · 2019
Neural reality of argument structure constructions
Bai Li, Zining Zhu, Guillaume Thomas, Frank Rudzicz, and Yang Xu. 2022 · 2022
Later among the works it cites.
minicons: Enabling flexible behavioral and representational analyses of transformer language models
Kanishka Misra. 2022 · 2022
Later among the works it cites.
Poverty of the stimulus without tears
Lisa Pearl. 2022 · 2022
Later among the works it cites.
CxLM: A construction and context-aware language model
Yu-Hsiang Tseng, Cing-Fang Shih, Pin-Er Chen, Hsin-Yu Chou, Mao-Chang Ku, and Shu-Kai Hsieh. 2022 · 2022
Later among the works it cites.
Artificial Neural Networks as Models of Human Language Acquisition
Alex Warstadt. 2022 · 2022
Later among the works it cites.
The better your syntax, the better your semantics? probing pretrained language models for the English comparative correlative
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Explain me this: Creativity, competition, and the partial productivity of constructions
Adele E Goldberg. 2019 · 2019
Cited alongside, same era.
It’s all in the name: Mitigating gender bias with name-based counterfactual data substitution
Rowan Hall Maudslay, Hila Gonen, Ryan Cotterell, and Simone Teufel. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
A standardized Project Gutenberg corpus for statistical analysis of natural language and quantitative linguistics
Martin Gerlach and Francesc Font-Clos. 2020 · 2020
Cited alongside, same era.
spaCy: Industrial-strength natural language processing in python
Matthew Honnibal, Ines Montani, Sofie Van Landeghem, and Adriane Boyd. 2020 · 2020
Cited alongside, same era.
How can we accelerate progress towards human-like linguistic generalization?
Tal Linzen. 2020 · 2020
Cited alongside, same era.
Leonie Weissweiler, Valentin Hofmann, Abdullatif Köksal, and Hinrich Schütze. 2022 · 2022
Later among the works it cites.
Noam Chomsky: The False Promise of ChatGPT
Noam Chomsky, Ian Roberts, and Jeffrey Watumull. 2023 · 2023
Later among the works it cites.
TinyStories: How Small Can Language Models Be and Still Speak Coherent English?
Ronen Eldan and Yuanzhi Li. 2023 · 2023
Later among the works it cites.
Language models can learn exceptions to syntactic rules
Cara Su-Yi Leong and Tal Linzen. 2023 · 2023
Later among the works it cites.
A discerning several thousand judgments: GPT-3 rates the article + adjective + numeral + noun construction
Kyle Mahowald. 2023 · 2023
Later among the works it cites.
How much do language models copy from their training data? evaluating linguistic novelty in text generation using RAVEN
R. Thomas McCoy, Paul Smolensky, Tal Linzen, Jianfeng Gao, and Asli Celikyilmaz. 2023 · 2023
Later among the works it cites.
Abstraction via exemplars? A representational case study on lexical category inference in BERT
Kanishka Misra and Najoung Kim. 2023 · 2023
Later among the works it cites.
Characterizing English Preposing in PP constructions
Christopher Potts. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Using collostructional analysis to evaluate BERT’s representation of linguistic constructions
Tim Veenboer and Jelke Bloem. 2023 · 2023
Later among the works it cites.
Findings of the BabyLM challenge: Sample-efficient pretraining on developmentally plausible corpora
Alex Warstadt, Aaron Mueller, Leshem Choshen, Ethan Wilcox, Chengxu Zhuang, Juan Ciro, Rafael Mosquera, Bhargavi Paranjabe, Adina Williams, Tal Linzen, and Ryan Cotterell. 2023 · 2023
Later among the works it cites.
OPT: Open Pre-trained Transformer Language Models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2023
Later among the works it cites.
Mission: Impossible language models
Julie Kallini, Isabel Papadimitriou, Richard Futrell, Kyle Mahowald, and Christopher Potts. 2024 · 2024
Closest in time.
Testing learning hypotheses using neural networks by manipulating learning data
Cara Su-Yi Leong and Tal Linzen. 2024 · 2024
Closest in time.
Dissociating language and thought in large language models
Kyle Mahowald, Anna A Ivanova, Idan A Blank, Nancy Kanwisher, Joshua B Tenenbaum, and Evelina Fedorenko. 2024 · 2024
Closest in time.
Filtered Corpus Training (FiCT) Shows that Language Models can Generalize from Indirect Evidence
Abhinav Patil, Jaap Jumelet, Yu Ying Chiu, Andy Lapastora, Peter Shen, Lexie Wang, Clevis Willrich, and Shane Steinert-Threlkeld. 2024 · 2024
Closest in time.
Hybrid human-LLM corpus construction and LLM evaluation for rare linguistic phenomena
Leonie Weissweiler, Abdullatif Köksal, and Hinrich Schütze. 2024 · 2024
Closest in time.
Language modelling as a multi-task problem
Lucas Weber, Jaap Jumelet, Elia Bruni, and Dieuwke Hupkes. 2021 · 2060
Closest in time.