Fetching the paper…
Reading the bibliography…
We consider the general problem of modeling temporal data with long-range dependencies, wherein new observations are fully or partially predictable based on temporally-distant, past observations.
A new approach to linear filtering and prediction problems
R. E. Kalman · 1960
Earlier work this paper cites.
A tutorial on hidden markov models and selected applications in speech recognition
L. R. Rabiner · 1989
Earlier work this paper cites.
Dyna, an integrated architecture for learning, planning, and reacting
R. S. Sutton · 1991
Earlier work this paper cites.
Gradient calculations for dynamic recurrent neural networks: A survey
B. A. Pearlmutter · 1995
Earlier work this paper cites.
Parameter estimation for linear dynamical systems
Z. Ghahramani and G. E. Hinton · 1996
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Predictability, complexity, and learning
W. Bialek, I. Nemenman, and N. Tishby · 2001
Earlier work this paper cites.
Stochastic gradient estimation
M. C. Fu · 2005
Earlier work this paper cites.
Time series prediction with variational bayesian nonlinear state-space models
M. Tornio, A. Honkela, and J. Karhunen · 2007
Earlier work this paper cites.
Pilco: A model-based and data-efficient approach to policy search
M. Deisenroth and C. E. Rasmussen · 2011
Earlier work this paper cites.
Bayesian filtering and smoothing , volume 3
S. Särkkä · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Learning stochastic recurrent networks
J. Bayer and C. Osendorfer · 2014
Earlier work this paper cites.
A. Graves, G. Wayne, and I. Danihelka · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Cited alongside, same era.
Auto-encoding variational Bayes
D. P. Kingma and M. Welling · 2014
Cited alongside, same era.
Learning neural network policies with guided policy search under unknown dynamics
S. Levine and P. Abbeel · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
D. J. Rezende, S. Mohamed, and D. Wierstra · 2014
Cited alongside, same era.
J. Weston, S. Chopra, and A. Bordes · 2014
Cited alongside, same era.
Data-efficient learning of feedback policies from image pixels using deep dynamical models
Neural programmer-interpreters
S. Reed and N. de Freitas · 2015
Later among the works it cites.
Variational inference with normalizing flows
D. J. Rezende and S. Mohamed · 2015
Later among the works it cites.
End-to-end memory networks
S. Sukhbaatar, J. Weston, R. Fergus, et al · 2015
Later among the works it cites.
Pointer networks
O. Vinyals, M. Fortunato, and N. Jaitly · 2015
Later among the works it cites.
Embed to control: A locally linear latent dynamics model for control from raw images
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller · 2015
Later among the works it cites.
Attend, infer, repeat: Fast scene understanding with generative models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J.-A. M. Assael, N. Wahlström, T. B. Schön, and M. P. Deisenroth · 2015
Cited alongside, same era.
Learning to transduce with unbounded memory
E. Grefenstette, K. M. Hermann, M. Suleyman, and P. Blunsom · 2015
Cited alongside, same era.
Draw: A recurrent neural network for image generation
K. Gregor, I. Danihelka, A. Graves, D. Jimenez Rezende, and D. Wierstra · 2015
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Cited alongside, same era.
Teaching machines to read and comprehend
K. M. Hermann, T. Kocisky, E. Grefenstette, L. Espeholt, W. Kay, M. Suleyman, and P. Blunsom · 2015
Cited alongside, same era.
Inferring algorithmic patterns with stack-augmented recurrent nets
A. Joulin and T. Mikolov · 2015
Cited alongside, same era.
R. G. Krishnan, U. Shalit, and D. Sontag · 2015
Cited alongside, same era.
S. A. Eslami, N. Heess, T. Weber, Y. Tassa, D. Szepesvari, G. E. Hinton, et al · 2016
Later among the works it cites.
Sequential neural models with stochastic layers
M. Fraccaro, S. K. Sønderby, U. Paquet, and O. Winther · 2016
Later among the works it cites.
Hybrid computing using a neural network with dynamic external memory
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska-Barwińska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou, et al · 2016
Later among the works it cites.
Text understanding with the attention sum reader network
R. Kadlec, M. Schmid, O. Bajgar, and J. Kleindienst · 2016
Later among the works it cites.
Learning to generate with memory
C. Li, J. Zhu, and B. Zhang · 2016
Later among the works it cites.
Hierarchical variational models
R. Ranganath, D. Tran, and D. M. Blei · 2016
Later among the works it cites.
Programming with a differentiable forth interpreter
S. Riedel, M. Bošnjak, and T. Rocktäschel · 2016
Later among the works it cites.
One-shot learning with memory-augmented neural networks
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap · 2016
Later among the works it cites.