Fetching the paper…
Reading the bibliography…
In this work, we propose a novel recurrent neural network (RNN) architecture.
Untersuchungen zu dynamischen neuronalen Netzen. Diploma thesis, Institut für Informatik, Lehrstuhl Prof. Brauer, Technische Universität München, 1991
Hochreiter, Sepp · 1991
Earlier work this paper cites.
Learning complex, extended sequences using the principle of history compression
Schmidhuber, Jürgen · 1992
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Yoshua, Simard, Patrice, and Frasconi, Paolo · 1994
Earlier work this paper cites.
Hierarchical recurrent neural networks for long-term dependencies
El Hihi, Salah and Bengio, Yoshua · 1995
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
The vanishing gradient problem during learning recurrent neural nets and problem solutions
Hochreiter, Sepp · 1998
Earlier work this paper cites.
Learning to forget: Continual prediction with LSTM
Gers, Felix A., Schmidhuber, Jürgen, and Cummins, Fred A · 2000
Earlier work this paper cites.
A neural probabilistic language model
Bengio, Yoshua, Ducharme, Réjean, and Vincent, Pascal · 2001
Earlier work this paper cites.
Generating text with recurrent neural networks
Sutskever, Ilya, Martens, James, and Hinton, Geoffrey E · 2011
Earlier work this paper cites.
Theano: new features and speed improvements
Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Bergstra, James, Goodfellow, Ian J., Bergeron, Arnaud, Bouchard, Nicolas, and Bengio, Yoshua · 2012
Cited alongside, same era.
Neural networks for machine learning
Hinton, Geoffrey · 2012
Cited alongside, same era.
The human knowledge compression contest
Hutter, Marcus · 2012
Cited alongside, same era.
Statistical Language Models based on Neural Networks
Mikolov, Tomas · 2012
Cited alongside, same era.
Subword language modeling with neural networks
Mikolov, Tomas, Sutskever, Ilya, Deoras, Anoop, Le, Hai-Son, Kombrink, Stefan, and Cernocky, J · 2012
Cited alongside, same era.
Pylearn2: a machine learning research library
Goodfellow, Ian J., Warde-Farley, David, Lamblin, Pascal, Dumoulin, Vincent, Mirza, Mehdi, Pascanu, Razvan, Bergstra, James, Bastien, Frédéric, and Bengio, Yoshua · 2013
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2014
Later among the works it cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, Kyunghyun, Van Merriënboer, Bart, Gulcehre, Caglar, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2014
Later among the works it cites.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2014
Later among the works it cites.
A clockwork rnn
Koutník, Jan, Greff, Klaus, Gomez, Faustino, and Schmidhuber, Jürgen · 2014
Later among the works it cites.
Deep networks with internal selective attention through feedback connections
Stollenga, Marijn F, Masci, Jonathan, Gomez, Faustino, and Schmidhuber, Jürgen · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Generating sequences with recurrent neural networks
Graves, Alex · 2013
Cited alongside, same era.
Training and analysing deep recurrent neural networks
Hermans, Michiel and Schrauwen, Benjamin · 2013
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc VV · 2014
Later among the works it cites.
Zaremba, Wojciech and Sutskever, Ilya · 2014
Later among the works it cites.
Recurrent neural network regularization
Zaremba, Wojciech, Sutskever, Ilya, and Vinyals, Oriol · 2014
Later among the works it cites.