Fetching the paper…
Reading the bibliography…
In this paper we compare different types of recurrent units in recurrent neural networks (RNNs).
Untersuchungen zu dynamischen neuronalen Netzen. Diploma thesis, Institut für Informatik, Lehrstuhl Prof. Brauer, Technische Universität München, 1991
S. Hochreiter · 1991
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Y. Bengio, P. Simard, and P. Frasconi · 1994
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Theano: a CPU and GPU math expression compiler
J. Bergstra, O. Breuleux, F. Bastien, P. Lamblin, R. Pascanu, G. Desjardins, J. Turian, D. Warde-Farley, and Y. Bengio · 2010
Earlier work this paper cites.
Practical variational inference for neural networks
A. Graves · 2011
Earlier work this paper cites.
Learning recurrent neural networks with Hessian-free optimization
J. Martens and I. Sutskever · 2011
Earlier work this paper cites.
Theano: new features and speed improvements
F. Bastien, P. Lamblin, R. Pascanu, J. Bergstra, I. J. Goodfellow, A. Bergeron, N. Bouchard, and Y. Bengio · 2012
Earlier work this paper cites.
Random search for hyper-parameter optimization
J. Bergstra and Y. Bengio · 2012
Cited alongside, same era.
Modeling temporal dependencies in high-dimensional sequences: Application to polyphonic music generation and transcription
N. Boulanger-Lewandowski, Y. Bengio, and P. Vincent · 2012
Cited alongside, same era.
Supervised Sequence Labelling with Recurrent Neural Networks
A. Graves · 2012
Cited alongside, same era.
Neural networks for machine learning
G. Hinton · 2012
Cited alongside, same era.
Advances in optimizing recurrent networks
Y. Bengio, N. Boulanger-Lewandowski, and R. Pascanu · 2013
Cited alongside, same era.
Pylearn2: a machine learning research library
I. J. Goodfellow, D. Warde-Farley, P. Lamblin, V. Dumoulin, M. Mirza, R. Pascanu, J. Bergstra, F. Bastien, and Y. Bengio · 2013
Speech recognition with deep recurrent neural networks
A. Graves, A.-r. Mohamed, and G. Hinton · 2013
Later among the works it cites.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2013
Later among the works it cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Closest in time.
On the properties of neural machine translation: Encoder-decoder approaches
K. Cho, B. van Merrienboer, D. Bahdanau, and Y. Bengio · 2014
Closest in time.
Learned-norm pooling for deep feedforward and recurrent neural networks
C. Gulcehre, K. Cho, R. Pascanu, and Y. Bengio · 2014
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Generating sequences with recurrent neural networks
A. Graves · 2013
Cited alongside, same era.
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Closest in time.