Fetching the paper…
Reading the bibliography…
While state-of-the-art neural network models continue to achieve lower perplexity scores on language modeling benchmarks, it remains unknown whether optimizing for broad-coverage predictive performance leads to human-like syntactic knowledge.
Assessing BERT’s syntactic abilities
Yoav Goldberg. 2019 · 1901
Earlier work this paper cites.
Syntactic structures
Noam Chomsky. 1957 · 1957
Earlier work this paper cites.
Finitary models of language users
George A. Miller and Noam Chomsky. 1963 · 1963
Earlier work this paper cites.
Constraints on Variables in Syntax
John Robert Ross. 1967 · 1967
Earlier work this paper cites.
The cognitive basis for linguistic structures
Tom Bever. 1970 · 1970
Earlier work this paper cites.
The Pseudo-Cleft Construction in English
Francis Roger Higgins. 1973 · 1973
Earlier work this paper cites.
Polarity Sensitivity as Inherent Scope Relations
William Ladusaw. 1979 · 1979
Earlier work this paper cites.
Definite NP anaphora and c-command domains
Tanya Reinhart. 1981 · 1981
Earlier work this paper cites.
Making and correcting errors during sentence comprehension: Eye movements in the analysis of structurally ambiguous sentences
Lyn Frazier and Keith Rayner. 1982 · 1982
Earlier work this paper cites.
How can grammars help parsers?
Stephen Crain and Janet Dean Fodor. 1985 · 1985
Earlier work this paper cites.
The independence of syntactic processing
Fernanda Ferreira and Charles Clifton, Jr. 1986 · 1986
Earlier work this paper cites.
Parsing wh-constructions: Evidence for on-line gap location
Laurie A Stowe. 1986 · 1986
Earlier work this paper cites.
BLLIP 1987-89 WSJ Corpus Release 1 LDC2000T43
Eugene Charniak, Don Blaheta, Niyu Ge, Keith Hall, John Hale, and Mark Johnson. 2000 · 1987
Earlier work this paper cites.
Lexical guidance in human parsing: Locus and processing characteristics
Don C. Mitchell. 1987 · 1987
Earlier work this paper cites.
Broken agreement
Kathryn Bock and Carol A. Miller. 1991 · 1991
Earlier work this paper cites.
Building a large annotated corpus of English: The Penn Treebank
Mitchell P. Marcus, Mary Ann Marcinkiewicz, and Beatrice Santorini. 1993 · 1993
Earlier work this paper cites.
Semantic influences on parsing: Use of thematic role information in syntactic ambiguity resolution
John C. Trueswell, Michael K. Tanenhaus, and Susan M. Garnsey. 1994 · 1994
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Plausibility and recovery from garden paths: An eye-tracking study
Martin J. Pickering and Matthew J. Traxler. 1998 · 1998
Earlier work this paper cites.
Structural change and reanalysis difficulty in language comprehension
Patrick Sturt, Martin J. Pickering, and Matthew W. Crocker. 1999 · 1999
Earlier work this paper cites.
A probabilistic Earley parser as a psycholinguistic model
John Hale. 2001 · 2001
Earlier work this paper cites.
NLTK: The natural language toolkit
Steven Bird and Edward Loper. 2004 · 2004
Cited alongside, same era.
Improved inference for unlexicalized parsing
Slav Petrov and Dan Klein. 2007 · 2007
Cited alongside, same era.
The parser doesn’t ignore intransitivity, after all
Adrian Staub. 2007 · 2007
Cited alongside, same era.
Processing polarity: How the ungrammatical intrudes on the grammatical
Shravan Vasishth, Sven Brüssow, Richard L Lewis, and Heiner Drenhaus. 2008 · 2008
Cited alongside, same era.
Insensitivity of the human sentence-processing system to hierarchical structure
Stefan L Frank and Rens Bod. 2011 · 2011
Cited alongside, same era.
Negative and positive polarity items: Variation, licensing, and compositionality
Anastasia Giannakidou. 2011 · 2011
Cited alongside, same era.
Under the hood: Using diagnostic classifiers to investigate and improve how language models track agreement information
Mario Giulianelli, Jack Harding, Florian Mohnert, Dieuwke Hupkes, and Willem Zuidema. 2018 · 2018
Later among the works it cites.
Predictive power of word surprisal for reading times is a linear function of language model quality
Adam Goodkind and Klinton Bicknell. 2018 · 2018
Later among the works it cites.
Colorless green recurrent networks dream hierarchically
Kristina Gulordava, Piotr Bojanowski, Edouard Grave, Tal Linzen, and Marco Baroni. 2018 · 2018
Later among the works it cites.
Constituency parsing with a self-attentive encoder
Nikita Kitaev and Dan Klein. 2018 · 2018
Later among the works it cites.
Targeted syntactic evaluation of language models
Rebecca Marvin and Tal Linzen. 2018 · 2018
Later among the works it cites.
Modeling garden path effects without explicit hierarchical syntax
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Syntax: A generative introduction , volume 18
Andrew Carnie. 2012 · 2012
Cited alongside, same era.
Sequential vs. hierarchical syntactic models of human incremental sentence processing
Victoria Fossum and Roger P. Levy. 2012 · 2012
Cited alongside, same era.
The effect of word predictability on reading time is logarithmic
Nathaniel J. Smith and Roger P. Levy. 2013 · 2013
Cited alongside, same era.
Exploring the limits of language modeling
Rafal Jozefowicz, Oriol Vinyals, Mike Schuster, Noam Shazeer, and Yonghui Wu. 2016 · 2016
Cited alongside, same era.
Assessing the ability of LSTMs to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Marten van Schijndel and Tal Linzen. 2018 · 2018
Later among the works it cites.
What do RNN language models learn about filler–gap dependencies?
Ethan Wilcox, Roger P. Levy, Takashi Morita, and Richard Futrell. 2018 · 2018
Later among the works it cites.
Transformer-XL: Attentive language models beyond a fixed-length context
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime G. Carbonell, Quoc V. Le, and Ruslan Salakhutdinov. 2019 · 2019
Later among the works it cites.
Neural language models as psycholinguistic subjects: Representations of syntactic state
Richard Futrell, Ethan Wilcox, Takashi Morita, Peng Qian, Miguel Ballesteros, and Roger Levy. 2019 · 2019
Later among the works it cites.
A structural probe for finding syntax in word representations
John Hewitt and Christopher D Manning. 2019 · 2019
Later among the works it cites.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
Tom McCoy, Ellie Pavlick, and Tal Linzen. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Quantity doesn’t buy quality syntax with neural language models
Marten van Schijndel, Aaron Mueller, and Tal Linzen. 2019 · 2019
Later among the works it cites.
Ordered neurons: Integrating tree structures into recurrent neural networks
Yikang Shen, Shawn Tan, Alessandro Sordoni, and Aaron Courville. 2019 · 2019
Later among the works it cites.
Hierarchical representation in neural language models: Suppression and recovery of expectations
Ethan Wilcox, Roger P. Levy, and Richard Futrell. 2019a · 2019
Later among the works it cites.
Structural supervision improves learning of non-local grammatical dependencies
Ethan Wilcox, Peng Qian, Richard Futrell, Miguel Ballestros, and Roger P. Levy. 2019b · 2019
Later among the works it cites.
XLNet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Ruslan Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
What don’t RNN language models learn about filler-gap dependencies?
Rui P. Chaves. 2020 · 2020
Closest in time.
A closer look at the performance of neural language models on reflexive anaphor licensing
Jennifer Hu, Sherry Yong Chen, and Roger P. Levy. 2020 · 2020
Closest in time.
BLiMP: A Benchmark of Linguistic Minimal Pairs for English
Alex Warstadt, Alicia Parrish, Haokun Liu, Anhad Mohananey, Wei Peng, Sheng-Fu Wang, and Samuel R Bowman. 2020 · 2020
Closest in time.
Evaluating neural networks as models of human online language processing
Ethan Wilcox, Jon Gauthier, Jennifer Hu, Peng Qian, and Roger P. Levy. 2020 · 2020
Closest in time.