Fetching the paper…
Reading the bibliography…
We introduce a new test of how well language models capture meaning in children's books.
Goldilocks and the three bears
Hassall, John · 1904
Earlier work this paper cites.
Interaction with context during human sentence processing
Altmann, Gerry and Steedman, Mark · 1988
Earlier work this paper cites.
A cache-based natural language model for speech recognition
Kuhn, Roland and De Mori, Renato · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, Ronald J · 1992
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Yoshua, Simard, Patrice, and Frasconi, Paolo · 1994
Earlier work this paper cites.
Word frequency distributions and lexical semantics
Baayen, R Harald and Lieber, Rochelle · 1996
Earlier work this paper cites.
Unconstrained on-line handwriting recognition with recurrent neural networks
Graves, Alex, Liwicki, Marcus, Bunke, Horst, Schmidhuber, Jürgen, and Fernández, Santiago · 2008
Earlier work this paper cites.
Large scale image annotation: learning to rank with joint word-image embeddings
Weston, Jason, Bengio, Samy, and Usunier, Nicolas · 2010
Earlier work this paper cites.
The neurobiology of semantic memory
Binder, Jeffrey R and Desai, Rutvik H · 2011
Earlier work this paper cites.
The microsoft research sentence completion challenge
Zweig, Geoffrey and Burges, Christopher JC · 2011
Cited alongside, same era.
Improving neural networks by preventing co-adaptation of feature detectors
Hinton, Geoffrey E, Srivastava, Nitish, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan R · 2012
Cited alongside, same era.
Improving word representations via global context and multiple word prototypes
Huang, Eric H, Socher, Richard, Manning, Christopher D, and Ng, Andrew Y · 2012
Cited alongside, same era.
Context dependent recurrent neural network language model
Mikolov, Tomas and Zweig, Geoffrey · 2012
Cited alongside, same era.
Scalable modified Kneser-Ney language model estimation
Heafield, Kenneth, Pouzyrevsky, Ivan, Clark, Jonathan H., and Koehn, Philipp · 2013
Cited alongside, same era.
Transition-based dependency parsing with stack long short-term memory
Dyer, Chris, Ballesteros, Miguel, Ling, Wang, Matthews, Austin, and Smith, Noah A · 2015
Closest in time.
Learning to transduce with unbounded memory
Grefenstette, Edward, Hermann, Karl Moritz, Suleyman, Mustafa, and Blunsom, Phil · 2015
Closest in time.
Teaching machines to read and comprehend
Hermann, Karl Moritz, Kočiský, Tomáš, Grefenstette, Edward, Espeholt, Lasse, Kay, Will, Suleyman, Mustafa, and Blunsom, Phil · 2015
Closest in time.
Inferring algorithmic patterns with stack-augmented recurrent nets
Joulin, Armand and Mikolov, Tomas · 2015
Closest in time.
Ask me anything: Dynamic memory networks for natural language processing
Kumar, Ankit, Irsoy, Ozan, Su, Jonathan, Bradbury, James, English, Robert, Pierce, Brian, Ondruska, Peter, Gulrajani, Ishaan, and Socher, Richard · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Richardson, Matthew, Burges, Christopher JC, and Renshaw, Erin · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Wan, Li, Zeiler, Matthew, Zhang, Sixin, Cun, Yann L, and Fergus, Rob · 2013
Cited alongside, same era.
The stanford corenlp natural language processing toolkit
Manning, Christopher D, Surdeanu, Mihai, Bauer, John, Finkel, Jenny, Bethard, Steven J, and McClosky, David · 2014
Cited alongside, same era.
Large-scale simple question answering with memory networks
Bordes, Antoine, Usunier, Nicolas, Chopra, Sumit, and Weston, Jason · 2015
Cited alongside, same era.
Towards ai-complete question answering: a set of prerequisite toy tasks
Weston, Jason, Bordes, Antoine, Chopra, Sumit, and Mikolov, Tomas
Cited in the paper.
Memory networks
Weston, Jason, Chopra, Sumit, and Bordes, Antoine
Cited in the paper.
Effective approaches to attention-based neural machine translation
Luong, Minh-Thang, Pham, Hieu, and Manning, Christopher D · 2015
Closest in time.
A neural attention model for abstractive sentence summarization
Rush, Alexander M, Chopra, Sumit, and Weston, Jason · 2015
Closest in time.
End-to-end memory networks
Sukhbaatar, Sainbayar, Szlam, Arthur, Weston, Jason, and Fergus, Rob · 2015
Closest in time.
Show, attend and tell: Neural image caption generation with visual attention
Xu, Kelvin, Ba, Jimmy, Kiros, Ryan, Courville, Aaron, Salakhutdinov, Ruslan, Zemel, Richard, and Bengio, Yoshua · 2015
Closest in time.