Fetching the paper…
Reading the bibliography…
Text documents are structured on multiple levels of detail: individual words are related by syntax, but larger units of text are related by discourse structure.
Building a large annotated corpus of English: The Penn Treebank
Mitchell P. Marcus, Mary Ann Marcinkiewicz, and Beatrice Santorini · 1993
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Yoshua Bengio, Patrice Simard, and Paolo Frasconi · 1994
Earlier work this paper cites.
Bootstrap methods and their application , volume 1
Anthony Christopher Davison and David Victor Hinkley · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Speech and language processing
Dan Jurafsky and James H Martin · 2000
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Janvin · 2003
Earlier work this paper cites.
Modeling local coherence: An entity-based approach
Regina Barzilay and Mirella Lapata · 2008
Earlier work this paper cites.
Introduction to Information Retrieval
Christopher D Manning, Prabhakar Raghavan, and Hinrich Schütze · 2008
Earlier work this paper cites.
BLLIP North American News Text, Complete
David McClosky, Eugene Charniak, and Mark Johnson · 2008
Earlier work this paper cites.
Statistical Machine Translation
Philipp Koehn · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Cited alongside, same era.
Recurrent neural network based language model
Tomas Mikolov, Martin Karafiát, Lukas Burget, Jan Cernockỳ, and Sanjeev Khudanpur · 2010
Cited alongside, same era.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer · 2011
Cited alongside, same era.
Context dependent recurrent neural network language model
Tomas Mikolov and Geoffrey Zweig · 2012
Cited alongside, same era.
On the difficulty of training recurrent neural networks
Razvan Pascanu, Tomas Mikolov, and Yoshua Bengio · 2012
Cited alongside, same era.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
A model of coherence based on distributed sentence representation
Jiwei Li and Eduard Hovy · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc VV Le · 2014
Later among the works it cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Closest in time.
A survey on the application of recurrent neural networks to statistical language modeling
Wim De Mulder, Steven Bethard, and Marie-Francine Moens · 2015
Closest in time.
A Hierarchical Neural Autoencoder for Paragraphs and Documents
Jiwei Li, Thang Luong, and Dan Jurafsky · 2015
Closest in time.
Hierarchical Recurrent Neural Network for Document Modeling
Rui Lin, Shujie Liu, Muyun Yang, Mu Li, Ming Zhou, and Sheng Li · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Modelling, Visualising and Summarising Documents with a Single Convolutional Neural Network
Misha Denil, Alban Demiraj, Nal Kalchbrenner, Phil Blunsom, and Nando de Freitas · 2014
Cited alongside, same era.
Distributed Representations of Sentences and Documents
Quoc Le and Tomas Mikolov · 2014
Cited alongside, same era.
Closest in time.
A neural network approach to context-sensitive generation of conversational responses
Alessandro Sordoni, Michel Galley, Michael Auli, Chris Brockett, Yangfeng Ji, Margaret Mitchell, Jian-Yun Nie, Jianfeng Gao, and Bill Dolan · 2015
Closest in time.
Document Modeling with Gated Recurrent Neural Network for Sentiment Classification
Duyu Tang, Bing Qin, and Ting Liu · 2015
Closest in time.
Larger-Context Language Modelling
Tian Wang and Kyunghyun Cho · 2015
Closest in time.