Evolutionary principles in self-referential learning
J. Schmidhuber · 1987
Earlier work this paper cites.
Visual space task specification, planning and control
M. Jagersand and R. Nelson · 1995
Earlier work this paper cites.
Learning to learn
S. Thrun and L. Pratt · 1998
Earlier work this paper cites.
Image-based simultaneous control of robot and target object motions by direct-image-interpretation method
K. Deguchi and I. Takahashi · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng and S. Russell · 2000
Earlier work this paper cites.
Learning to learn using gradient descent
S. Hochreiter, A. Younger, and P. Conwell · 2001
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
Robot motor skill coordination with em-based reinforcement learning
P. Kormushev, S. Calinon, and D. G. Caldwell · 2010
Earlier work this paper cites.
Relative entropy inverse reinforcement learning
A. Boularias, J. Kober, and J. Peters · 2011
Earlier work this paper cites.
Autonomous reinforcement learning on raw visual input data in a real world application
S. Lange, M. Riedmiller, and A. Voigtlander · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Learning objective functions for manipulation
M. Kalakrishnan, P. Pastor, L. Righetti, and S. Schaal · 2013
Earlier work this paper cites.
The Cross-Entropy Method: A Unified Approach to Combinatorial Optimization, Monte-Carlo Simulation and Machine Learning
R. Rubinstein and D. Kroese · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller · 2015
Earlier work this paper cites.
Pouring skills with planning and learning modeled from human demonstrations
A. Yamaguchi, C. G. Atkeson, and T. Ogasawara · 2015
Earlier work this paper cites.
Human-level concept learning through probabilistic program induction
B. M. Lake, R. Salakhutdinov, and J. B. Tenenbaum · 2015
Earlier work this paper cites.