Fetching the paper…
Reading the bibliography…
In computational psycholinguistics, various language models have been evaluated against human reading behavior (e.g., eye movement) to build human-like computational models.
Effects of contextual constraint on eye movements in reading: A further examination
Keith Rayner and Arnold D Well. 1996 · 1996
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Locality and Parsing Complexity
Lars Konieczny. 2000 · 2000
Earlier work this paper cites.
A Probabilistic Earley Parser as a Psycholinguistic Model
John Hale. 2001 · 2001
Earlier work this paper cites.
Scaling Laws for Neural Language Models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2001
Earlier work this paper cites.
Entropy rate constancy in text
Dmitriy Genzel and Eugene Charniak. 2002 · 2002
Earlier work this paper cites.
The dundee corpus
Alan Kennedy, Robin Hill, and Joël Pynte. 2003 · 2003
Earlier work this paper cites.
Probabilistic models of word order and syntactic discontinuity
Roger Levy. 2005 · 2005
Earlier work this paper cites.
MeCab: Yet Another Part-of-speech and Morphological Analyzer
Taku Kudo. 2006 · 2006
Earlier work this paper cites.
Speakers optimize information density through syntactic reduction
T Jaeger and Roger Levy. 2007 · 2007
Earlier work this paper cites.
Data from eye-tracking corpora as evidence for theories of syntactic processing complexity
Vera Demberg and Frank Keller. 2008 · 2008
Earlier work this paper cites.
Expectation-based syntactic comprehension
Roger Levy. 2008 · 2008
Earlier work this paper cites.
Deriving lexical and syntactic expectation-based measures for psycholinguistic modeling via incremental top-down parsing
Brian Roark, Asaf Bachrach, Carlos Cardenas, and Christophe Pallier. 2009 · 2009
Earlier work this paper cites.
Why are some word orders more common than others? a uniform information density account
Luke Maurits, Dan Navarro, and Amy Perfors. 2010 · 2010
Earlier work this paper cites.
On Achieving and Evaluating Language-Independence in NLP
Emily M Bender. 2011 · 2011
Cited alongside, same era.
Insensitivity of the Human Sentence-Processing System to Hierarchical Structure
Stefan L Frank and Rens Bod. 2011 · 2011
Cited alongside, same era.
Extracting training data from large language models
Nicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom B. Brown, Dawn Song, Úlfar Erlingsson, Alina Oprea, and Colin Raffel. 2020 · 2012
Cited alongside, same era.
Sequential vs. Hierarchical Syntactic Models of Human Incremental Sentence Processing
Victoria Fossum and Roger Levy. 2012 · 2012
Cited alongside, same era.
The effect of word predictability on reading time is logarithmic
Nathaniel J. Smith and Roger Levy. 2013 · 2013
Cited alongside, same era.
Attention Is All You Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Between Reading Time and Clause Boundaries in Japanese - Wrap-up Effect in a Head-final Language
Masayuki Asahara. 2018 · 2018
Later among the works it cites.
Predictive power of word surprisal for reading times is a linear function of language model quality
Adam Goodkind and Klinton Bicknell. 2018 · 2018
Later among the works it cites.
Finding Syntax in Human Encephalography with Beam Search
John Hale, Chris Dyer, Adhiguna Kuncoro, and Jonathan R. Brennan. 2018 · 2018
Later among the works it cites.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing
Taku Kudo and John Richardson. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fitting linear mixed-effects models using lme4
Douglas Bates, Martin Mächler, Ben Bolker, and Steve Walker. 2015 · 2015
Cited alongside, same era.
The erp response to the amount of information conveyed by words in sentences
Stefan L. Frank, Leun J. Otten, Giulia Galli, and Gabriella Vigliocco. 2015 · 2015
Cited alongside, same era.
Reading-Time Annotations for “Balanced Corpus of Contemporary Written Japanese”
Masayuki Asahara, Hajime Ono, and Edson T Miyamoto. 2016 · 2016
Cited alongside, same era.
Testing the processing hypothesis of word order variation using a probabilistic language model
Jelke Bloem. 2016 · 2016
Cited alongside, same era.
Abstract linguistic structure correlates with temporal activity during naturalistic comprehension
Jonathan R Brennan, Edward P Stabler, Sarah E Van Wagenen, Wen-Ming Luh, and John T Hale. 2016 · 2016
Cited alongside, same era.
Neural Machine Translation of Rare Words with Subword Units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Between Reading Time and Information Structure
Masayuki Asahara. 2017 · 2017
Cited alongside, same era.
Alec Radrof, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2018 · 2018
Later among the works it cites.
Comparing Gated and Simple Recurrent Neural Network Architectures as Models of Human Sentence Processing
C Aurnhammer and S L Frank. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
Language Models are Few-Shot Learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Later among the works it cites.
If beam search is the answer, what was the question?
Clara Meister, Ryan Cotterell, and Tim Vieira. 2020 · 2020
Later among the works it cites.
On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior
Ethan Gotlieb Wilcox, Jon Gauthier, Jennifer Hu, Peng Qian, and Roger Levy. 2020 · 2020
Later among the works it cites.
Human Sentence Processing: Recurrence or Attention?
Danny Merkx and Stefan L. Frank. 2020 · 2021
Closest in time.
A cognitive regularizer for language modeling
Jason Wei, Clara Meister, and Ryan Cotterell. 2021 · 2021
Closest in time.