Fetching the paper…
Reading the bibliography…
Data-efficient reinforcement learning (RL) in continuous state-action spaces using very high-dimensional observations remains a key challenge in developing fully autonomous systems.
Learning internal representations by error propagation
D. E. Rumelhart, G. E. Hinton, and R. J. Williams · 1986
Earlier work this paper cites.
An on-line algorithm for dynamic reinforcement learning and planning in reactive environments
J. Schmidhuber · 1990
Earlier work this paper cites.
Learning tasks from a single demonstration
C. G. Atkeson and S. Schaal · 1997
Earlier work this paper cites.
Exploiting model uncertainty estimates for safe dynamic control learning
J. G. Schneider · 1997
Earlier work this paper cites.
Learning from demonstration
S. Schaal · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Autonomous helicopter control using reinforcement learning policy search methods
J. A. Bagnell and J. G. Schneider · 2001
Earlier work this paper cites.
A generalized iterative LQG method for locally-optimal feedback control of constrained nonlinear stochastic systems
E. Todorov and W. Li · 2005
Earlier work this paper cites.
Reducing the dimensionality of data with neural networks
G. Hinton and R. Salakhutdinov · 2006
Earlier work this paper cites.
Greedy layer-wise training of deep networks
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle · 2007
Cited alongside, same era.
Extracting and composing robust features with denoising autoencoders
P. Vincent, L. Hugo, Y. Bengio, and P.-A. Manzagol · 2008
Cited alongside, same era.
Robot trajectory optimization using approximate inference
M. Toussaint · 2009
Cited alongside, same era.
Learning deep architectures for AI
Y. Bengio · 2009
Cited alongside, same era.
Learning multiple layers of features from tiny images
A. Krizhevsky · 2009
Cited alongside, same era.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. E. Hinton · 2010
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A. M. Saxe, J. L. McClelland, and S. Ganguli · 2013
Later among the works it cites.
Probabilistic differential dynamic programming
Y. Pan and E. Theodorou · 2014
Later among the works it cites.
Deep learning in neural networks: An overview
J. Schmidhuber · 2014
Later among the works it cites.
Stochastic backpropagation and variational inference in deep latent Gaussian models
D. Jimenez Rezende, S. Mohamed, and D. Wierstra · 2014
Later among the works it cites.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio · 2010
Cited alongside, same era.
PILCO: A model-based and data-efficient approach to policy search
M. P. Deisenroth and C. E. Rasmussen · 2011
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
From pixels to torques: Policy learning with deep dynamical models
N. Wahlström, T. B. Schön, and M. P. Deisenroth
Cited in the paper.
Learning deep dynamical models from image pixels
N. Wahlström, T. B. Schön, and M. P. Deisenroth
Cited in the paper.
Gaussian processes for data-efficient learning in robotics and control
M. P. Deisenroth, D. Fox, and C. E. Rasmussen · 2015
Closest in time.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2015
Closest in time.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Closest in time.
Embed to control: A locally linear latent dynamics model for control from raw images
M. Watter, J. T. Springenberg, J. Boedecker, and M. A. Riedmiller · 2015
Closest in time.