Fetching the paper…
Reading the bibliography…
We introduce Embed to Control (E2C), a method for model learning and control of non-linear dynamical systems from raw pixel images.
Differential dynamic programming
D. Jacobson and D. Mayne · 1970
Earlier work this paper cites.
Optimal Control and Estimation
R. F. Stengel · 1994
Earlier work this paper cites.
An approach to fuzzy control of nonlinear systems; stability and design issues
H. Wang, K. Tanaka, and M. Griffin · 1996
Earlier work this paper cites.
Bayesian Forecasting and Dynamic Models (Springer Series in Statistics)
M. West and J. Harrison · 1997
Earlier work this paper cites.
Introduction to Reinforcement Learning
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
M. I. Jordan, Z. Ghahramani, T. S. Jaakkola, and L. K. Saul · 1999
Earlier work this paper cites.
Iterative Linear Quadratic Regulator Design for Nonlinear Biological Movement Systems
W. Li and E. Todorov · 2004
Earlier work this paper cites.
A generalized iterative LQG method for locally-optimal feedback control of constrained nonlinear stochastic systems
E. Todorov and W. Li · 2005
Earlier work this paper cites.
Receding horizon differential dynamic programming
Y. Tassa, T. Erez, and W. D. Smart · 2008
Earlier work this paper cites.
Robot Trajectory Optimization using Approximate Inference
M. Toussaint · 2009
Earlier work this paper cites.
Learning nonlinear dynamic models
J. Langford, R. Salakhutdinov, and T. Zhang · 2009
Earlier work this paper cites.
Deconvolutional networks
M. D. Zeiler, D. Krishnan, G. W. Taylor, and R. Fergus · 2010
Earlier work this paper cites.
Deep auto-encoder neural networks in reinforcement learning
S. Lange and M. Riedmiller · 2010
Earlier work this paper cites.
Dynamical binary latent variable models for 3d human pose tracking
G. W. Taylor, L. Sigal, D. J. Fleet, and G. E. Hinton · 2010
Cited alongside, same era.
Reinforcement learning on slow features of high-dimensional input streams
R. Legenstein, N. Wilbert, and L. Wiskott · 2010
Cited alongside, same era.
Transforming auto-encoders
G. Hinton, A. Krizhevsky, and S. Wang · 2011
Cited alongside, same era.
Unsupervised learning of visual invariance with temporal coherence
W. Zou, A. Ng, and K. Yu · 2011
Cited alongside, same era.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Cited alongside, same era.
Variational policy search via trajectory optimization
S. Levine and V. Koltun · 2013
Cited alongside, same era.
State representation learning in robotics: Using prior knowledge about physical interaction
R. Jonschkowski and O. Brock · 2014
Later among the works it cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A. M. Saxe, J. L. McClelland, and S. Ganguli · 2014
Later among the works it cites.
Learning to generate chairs with convolutional neural networks
A. Dosovitskiy, J. T. Springenberg, and T. Brox · 2015
Closest in time.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2015
Closest in time.
Autonomous learning of state representations for control
W. Böhmer, J. T. Springenberg, J. Boedecker, M. Riedmiller, and K. Obermayer · 2015
Closest in time.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning to relate images
R. Memisevic · 2013
Cited alongside, same era.
Probabilistic differential dynamic programming
Y. Pan and E. Theodorou · 2014
Cited alongside, same era.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
D. J. Rezende, S. Mohamed, and D. Wierstra · 2014
Cited alongside, same era.
Learning stochastic recurrent networks
J. Bayer and C. Osendorfer · 2014
Cited alongside, same era.
Latent Kullback Leibler control for continuous-state systems using probabilistic graphical models
T. Matsubara, V. Gómez, and H. J. Kappen · 2014
Cited alongside, same era.
Closest in time.
Learning of non-parametric control policies with high-dimensional state features
H. van Hoof, J. Peters, and G. Neumann · 2015
Closest in time.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2015
Closest in time.
From pixels to torques: Policy learning with deep dynamical models
N. Wahlström, T. B. Schön, and M. P. Deisenroth · 2015
Closest in time.
DRAW: A recurrent neural network for image generation
K. Gregor, I. Danihelka, A. Graves, D. Rezende, and D. Wierstra · 2015
Closest in time.
Nice: Non-linear independent components estimation
L. Dinh, D. Krueger, and Y. Bengio · 2015
Closest in time.
Transformation properties of learned visual representations
T. Cohen and M. Welling · 2015
Closest in time.
Deep convolutional inverse graphics network
T. D. Kulkarni, W. Whitney, P. Kohli, and J. B. Tenenbaum · 2015
Closest in time.