Fetching the paper…
Reading the bibliography…
Recurrent Neural Networks (RNNs), and specifically a variant with Long Short-Term Memory (LSTM), are enjoying renewed interest as a result of successful applications in a wide range of machine learning problems that involve sequential data.
Learning internal representations by error propagation
Rumelhart, David E, Hinton, Geoffrey E, and Williams, Ronald J · 1985
Earlier work this paper cites.
Generalization of backpropagation with application to a recurrent gas market model
Werbos, Paul J · 1988
Earlier work this paper cites.
A dynamic language model for speech recognition
Jelinek, Frederick, Merialdo, Bernard, Roukos, Salim, and Strauss, Martin · 1991
Earlier work this paper cites.
Building a large annotated corpus of english: The penn treebank
Marcus, Mitchell P, Marcinkiewicz, Mary Ann, and Santorini, Beatrice · 1993
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Yoshua, Simard, Patrice, and Frasconi, Paolo · 1994
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
An empirical study of smoothing techniques for language modeling
Chen, Stanley F and Goodman, Joshua · 1999
Earlier work this paper cites.
Spoken language processing: A guide to theory, algorithm, and system development
Huang, Xuedong, Acero, Alex, Hon, Hsiao-Wuen, and Foreword By-Reddy, Raj · 2001
Earlier work this paper cites.
Visualizing data using t-sne
Van der Maaten, Laurens and Hinton, Geoffrey · 2008
Earlier work this paper cites.
Recurrent neural network based language model
Mikolov, Tomas, Karafiát, Martin, Burget, Lukas, Cernockỳ, Jan, and Khudanpur, Sanjeev · 2010
Earlier work this paper cites.
Generating text with recurrent neural networks
Sutskever, Ilya, Martens, James, and Hinton, Geoffrey E · 2011
Earlier work this paper cites.
Diagnosing error in object detectors
Hoiem, Derek, Chodpathumwan, Yodsawalai, and Dai, Qieyun · 2012
Earlier work this paper cites.
The human knowledge compression contest
Hutter, Marcus · 2012
Cited alongside, same era.
Statistical language models based on neural networks
Mikolov, Tomáš · 2012
Cited alongside, same era.
On the difficulty of training recurrent neural networks
Pascanu, Razvan, Mikolov, Tomas, and Bengio, Yoshua · 2012
Cited alongside, same era.
Generating sequences with recurrent neural networks
Graves, Alex · 2013
Cited alongside, same era.
Speech recognition with deep recurrent neural networks
Graves, Alex, Mohamed, A-R, and Hinton, Geoffrey · 2013
Cited alongside, same era.
Scalable modified Kneser-Ney language model estimation
Heafield, Kenneth, Pouzyrevsky, Ivan, Clark, Jonathan H., and Koehn, Philipp · 2013
Deep learning in neural networks: An overview
Schmidhuber, J · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc VV · 2014
Later among the works it cites.
Weston, Jason, Chopra, Sumit, and Bordes, Antoine · 2014
Later among the works it cites.
Semi-supervised sequence learning
Dai, Andrew M and Le, Quoc V · 2015
Closest in time.
Rmsprop and equilibrated adaptive learning rates for non-convex optimization
Dauphin, Yann N, de Vries, Harm, Chung, Junyoung, and Bengio, Yoshua · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Training and analysing deep recurrent neural networks
Hermans, Michiel and Schrauwen, Benjamin · 2013
Cited alongside, same era.
How to construct deep recurrent neural networks
Pascanu, Razvan, Gülçehre, Çaglar, Cho, Kyunghyun, and Bengio, Yoshua · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2014
Cited alongside, same era.
On the properties of neural machine translation: Encoder-decoder approaches
Cho, Kyunghyun, van Merriënboer, Bart, Bahdanau, Dzmitry, and Bengio, Yoshua · 2014
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Chung, Junyoung, Gulcehre, Caglar, Cho, KyungHyun, and Bengio, Yoshua · 2014
Cited alongside, same era.
Graves, Alex, Wayne, Greg, and Danihelka, Ivo · 2014
Cited alongside, same era.
Long-term recurrent convolutional networks for visual recognition and description
Donahue, Jeff, Hendricks, Lisa Anne, Guadarrama, Sergio, Rohrbach, Marcus, Venugopalan, Subhashini, Saenko, Kate, and Darrell, Trevor · 2015
Closest in time.
Greff, Klaus, Srivastava, Rupesh Kumar, Koutník, Jan, Steunebrink, Bas R., and Schmidhuber, Jürgen · 2015
Closest in time.
Inferring algorithmic patterns with stack-augmented recurrent nets
Joulin, Armand and Mikolov, Tomas · 2015
Closest in time.
An empirical exploration of recurrent network architectures
Jozefowicz, Rafal, Zaremba, Wojciech, and Sutskever, Ilya · 2015
Closest in time.
Deep visual-semantic alignments for generating image descriptions
Karpathy, Andrej and Fei-Fei, Li · 2015
Closest in time.
Show and tell: A neural image caption generator
Vinyals, Oriol, Toshev, Alexander, Bengio, Samy, and Erhan, Dumitru · 2015
Closest in time.