Fetching the paper…
Reading the bibliography…
Leveraging advances in variational inference, we propose to enhance recurrent neural networks with latent variables, resulting in Stochastic Recurrent Networks (STORNs).
Turing computability with neural nets
Siegelmann, Hava T and Sontag, Eduardo D · 1991
Earlier work this paper cites.
Probability and random processes , volume 2
Grimmett, Geoffrey and Stirzaker, David · 1992
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
On the approximation capability of recurrent neural networks
Hammer, Barbara · 2000
Earlier work this paper cites.
Exponential family harmoniums with an application to information retrieval
Welling, Max, Rosen-Zvi, Michal, and Hinton, Geoffrey E · 2004
Earlier work this paper cites.
Style translation for human motion
Hsu, Eugene, Pulli, Kari, and Popović, Jovan · 2005
Earlier work this paper cites.
Modeling human motion using binary latent variables
Taylor, Graham W, Hinton, Geoffrey E, and Roweis, Sam T · 2006
Earlier work this paper cites.
Unconstrained online handwriting recognition with recurrent neural networks
Graves, Alex, Fernández, Santiago, Liwicki, Marcus, Bunke, Horst, and Schmidhuber, Jurgen · 2008
Earlier work this paper cites.
The recurrent temporal restricted boltzmann machine
Sutskever, I., Hinton, G., and Taylor, G · 2008
Earlier work this paper cites.
Learning recurrent neural networks with hessian-free optimization
Martens, J. and Sutskever, I · 2011
Cited alongside, same era.
Advances in optimizing recurrent networks
Bengio, Y., Boulanger-Lewandowski, N., and Pascanu, R · 2012
Cited alongside, same era.
Random search for hyper-parameter optimization
Bergstra, James and Bengio, Yoshua · 2012
Cited alongside, same era.
Boulanger-Lewandowski, N., Bengio, Y., and Vincent, P · 2012
Cited alongside, same era.
Adadelta: An adaptive learning rate method
Zeiler, Matthew D · 2012
Cited alongside, same era.
Speech recognition with deep recurrent neural networks
Graves, Alex, Mohamed, Abdel-rahman, and Hinton, Geoffrey · 2013
Later among the works it cites.
Auto-encoding variational bayes
Kingma, Diederik P and Welling, Max · 2013
Later among the works it cites.
How to construct deep recurrent neural networks
Pascanu, Razvan, Gulcehre, Caglar, Cho, Kyunghyun, and Bengio, Yoshua · 2013
Later among the works it cites.
On the importance of initialization and momentum in deep learning
Sutskever, Ilya, Martens, James, Dahl, George, and Hinton, Geoffrey · 2013
Later among the works it cites.
A new learning algorithm for stochastic feedforward neural nets
Tang, Yichuan and Salakhutdinov, Ruslan · 2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Training neural networks with implicit variance
Bayer, Justin, Osendorfer, Christian, Urban, Sebastian, et al · 2013
Cited alongside, same era.
High-dimensional sequence transduction
Boulanger-Lewandowski, Nicolas, Bengio, Yoshua, and Vincent, Pascal · 2013
Cited alongside, same era.
Generating sequences with recurrent neural networks
Graves, Alex · 2013
Cited alongside, same era.
On fast dropout and its applicability to recurrent networks
Bayer, Justin, Osendorfer, Christian, Korhammer, Daniela, Chen, Nutan, Urban, Sebastian, and van der Smagt, Patrick
Cited in the paper.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, Kyunghyun, van Merrienboer, Bart, Gulcehre, Caglar, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2014
Closest in time.
Stochastic back-propagation and variational inference in deep latent gaussian models
Rezende, Danilo Jimenez, Mohamed, Shakir, and Wierstra, Daan · 2014
Closest in time.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc V · 2014
Closest in time.