Fetching the paper…
Reading the bibliography…
Recurrent neural networks (RNNs) have been used extensively and with increasing success to model various types of sequential data.
Finding structure in time
Jeffrey L. Elman · 1990
Earlier work this paper cites.
Neural sequence chunkers
Jürgen Schmidhuber · 1991
Earlier work this paper cites.
Learning complex, extended sequences using the principle of history compression
Jürgen Schmidhuber · 1992
Earlier work this paper cites.
Induction of multiscale temporal structure
Michael C Mozer · 1993
Earlier work this paper cites.
A neural network based, speaker independent, large vocabulary, continuous speech recognition system: the WERNICKE project
Tony Robinson, Luís B. Almeida, Jean-Marc Boite, Hervé Bourlard, Frank Fallside, Mike Hochberg, Dan J. Kershaw, Phil Kohn, Yochai Konig, Nelson Morgan, João Paulo Neto, Steve Renals, Marco Saerens, and Chuck Wooters · 1993
Earlier work this paper cites.
Factorial hidden markov models
Zoubin Ghahramani and Michael I. Jordan · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
The hierarchical hidden markov model: Analysis and applications
Shai Fine, Yoram Singer, and Naftali Tishby · 1998
Earlier work this paper cites.
Variational learning for switching state-space models
Zoubin Ghahramani and Geoffrey E. Hinton · 2000
Cited alongside, same era.
Linear-time inference in hierarchical hmms
Kevin P. Murphy and Mark A. Paskin · 2001
Cited alongside, same era.
The infinite factorial hidden markov model
Jurgen Van Gael, Yee Whye Teh, and Zoubin Ghahramani · 2008
Cited alongside, same era.
Statistical language models based on neural networks
Tomáš Mikolov · 2012
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart van Merrienboer, Dzmitry Bahdanau, and Yoshua Bengio · 2014
Towards end-to-end speech recognition with recurrent neural networks
Alex Graves and Navdeep Jaitly · 2014
Later among the works it cites.
A clockwork RNN
Jan Koutník, Klaus Greff, Faustino J. Gomez, and Jürgen Schmidhuber · 2014
Later among the works it cites.
Learning longer memory in recurrent neural networks
Tomas Mikolov, Armand Joulin, Sumit Chopra, Michaël Mathieu, and Marc’Aurelio Ranzato · 2014
Later among the works it cites.
Alternative structures for character-level rnns
Piotr Bojanowski, Armand Joulin, and Tomas Mikolov · 2015
Later among the works it cites.
Hierarchical multiscale recurrent neural networks
Junyoung Chung, Sungjin Ahn, and Yoshua Bengio · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Çaglar Gülçehre, KyungHyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Alex Graves · 2016
Closest in time.
Exploring the limits of language modeling
Rafal Józefowicz, Oriol Vinyals, Mike Schuster, Noam Shazeer, and Yonghui Wu · 2016
Closest in time.
Architectural complexity measures of recurrent neural networks
Saizheng Zhang, Yuhuai Wu, Tong Che, Zhouhan Lin, Roland Memisevic, Ruslan Salakhutdinov, and Yoshua Bengio · 2016
Closest in time.