Fetching the paper…
Reading the bibliography…
In this paper, we study novel neural network structures to better model long term dependency in sequential data.
Approximation by superpositions of a sigmoidal function
Cybenko, G · 1989
Earlier work this paper cites.
Finding structure in time
Elman, J. L · 1990
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Werbos, P. J · 1990
Earlier work this paper cites.
Approximation capabilities of multilayer feedforward networks
Hornik, K · 1991
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Y., Simard, P., and Frasconi, P · 1994
Earlier work this paper cites.
On the computational power of neural nets
Siegelmann, H. T. and Sontag, E. D · 1995
Earlier work this paper cites.
Hierarchical recurrent neural networks for long-term dependencies
Hihi, Salah and Bengio, Yoshua · 1996
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
A neural probabilistic language model
Bengio, Y., Ducharme, R., Vincent, P., and Janvin, C · 2003
Earlier work this paper cites.
Extensions of recurrent neural network language model
Mikolov, T., Kombrink, S., Burget, L., Černockỳ, J.H., and Khudanpur, S · 2011
Earlier work this paper cites.
Generating text with recurrent neural networks
Sutskever, I., Martens, J., and Hinton, G · 2011
Earlier work this paper cites.
Neural networks for handwriting recognition, Book Chapter, Computational intelligence paradigms in advanced pattern classification
Liwicki, M., Graves, A., and Bunke, H · 2012
Cited alongside, same era.
Statistical Language Models based on Neural Networks
Mikolov, T · 2012
Cited alongside, same era.
Lstm neural networks for language modeling
Sundermeyer, M., Schlüter, R., and Ne, H · 2012
Cited alongside, same era.
Generating sequences with recurrent neural networks
Graves, A · 2013
Cited alongside, same era.
Speech recognition with deep recurrent neural
Graves, A., Mohamed, A., and Hinton, G · 2013
Cited alongside, same era.
Regularization and nonlinearities for neural language models: when are they needed?
How to construct deep recurrent neural networks
Pascanu, R., Gulcehre, C., Cho, K., and Bengio, Y · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q · 2014
Later among the works it cites.
Going deeper with convolutions
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., and Rabinovich, A · 2014
Later among the works it cites.
Recurrent neural network regularization
Zaremba, W., Sutskever, I., and Vinyals, O.l · 2014
Later among the works it cites.
Gated feedback recurrent neural networks
Chung, J., Gulcehre, C., Cho, K., and Bengio, Y · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pachitariu, M. and Sahani, M · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y · 2014
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Cited alongside, same era.
A clockwork rnn
Koutnik, J., Greff, K., Gomez, F., and Schmidhuber, J · 2014
Cited alongside, same era.
Lee, C. Y., Xie, S., Gallagher, P., Zhang, Z., and Tu, Z · 2014
Cited alongside, same era.
Learning longer memory in recurrent neural networks
Mikolov, T., Joulin, A., Chopra, S., Mathieu, M., and Ranzato, M · 2014
Cited alongside, same era.
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Later among the works it cites.
Character-aware neural language models
Kim, Y., Jernite, Y., Sontag, D., and Rush, A. M · 2015
Later among the works it cites.
Highway networks
Srivastava, R. K., Greff, K., and Schmidhuber, J · 2015
Later among the works it cites.
End-to-end memory networks
Sukhbaatar, S., Szlam, A., Weston, J., and Fergus, R · 2015
Later among the works it cites.
The fixed-size ordinally-forgetting encoding method for neural network language models
Zhang, S., Jiang, H., Xu, M., Hou, J., and Dai, L · 2015
Later among the works it cites.