Fetching the paper…
Reading the bibliography…
Trial-and-error based reinforcement learning (RL) has seen rapid advancements in recent times, especially with the advent of deep neural networks.
Dynamic Programming
R. E. Bellman · 1957
Earlier work this paper cites.
The Mathematical Theory of Optimal Processes
L. S. Pontryagin, E. F. Mishchenko, V. G. Boltyanskii, and R. V. Gamkrelidze · 1962
Earlier work this paper cites.
Partitioned Variable Metric Updates for Large Structured Optimization Problems
A. Griewank and P. L. Toint · 1982
Earlier work this paper cites.
A Multiple Shooting Algorithm for Direct Solution of Optimal Control Problems
H. G. Bock and K. J. Plitt · 1984
Earlier work this paper cites.
Optimization and Non-Smooth Analysis
F. H. Clarke · 1990
Earlier work this paper cites.
Exploiting Model Uncertainty Estimates for Safe Dynamic Control Learning
J. G. Schneider · 1997
Earlier work this paper cites.
Introduction to Gaussian Processes
D. J. C. MacKay · 1998
Earlier work this paper cites.
Reinforcement Learning: An Introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Constrained Model Predictive Control: Stability and Optimality
D. Q. Mayne, J. B. Rawlings, C. V. Rao, and P. O. Scokaert · 2000
Earlier work this paper cites.
PEGASUS: A Policy Search Method for Large MDPs and POMDPs
A. Y. Ng and M. I. Jordan · 2000
Earlier work this paper cites.
A Family of Algorithms for Approximate Bayesian Inference
T. P. Minka · 2001
Earlier work this paper cites.
Expectation Propagation for Approximate Bayesian Inference
T. P. Minka · 2001
Earlier work this paper cites.
Tractable Approximations for Probabilistic Models: The Adaptive TAP Mean Field Approach
M. Opper and O. Winther · 2001
Earlier work this paper cites.
Gaussian Process Priors with Uncertain Inputs-Application to Multiple-Step Ahead Time Series Forecasting
A. Girard, C. E. Rasmussen, J. Quinonero-Candela, and R. Murray-Smith · 2003
Earlier work this paper cites.
Optimal Control Systems
D. S. D. S. Naidu and R. C. Naidu, Subbaram/Dorf · 2003
Earlier work this paper cites.
A Unifying View of Sparse Approximate Gaussian Process Regression
J. Quiñonero-Candela and C. E. Rasmussen · 2005
Earlier work this paper cites.
A Generalized Iterative LQG Method for Locally-Optimal Feedback Control of Constrained Nonlinear Stochastic Systems
E. Todorov and Weiwei Li · 2005
Earlier work this paper cites.
Numerical Optimization
J. Nocedal and S. J. Wright · 2006
Cited alongside, same era.
Gaussian Processes for Machine Learning
C. E. Rasmussen and C. K. I. Williams · 2006
Cited alongside, same era.
Explicit Stochastic Predictive Control of Combustion Plants based on Gaussian Process Models
A. Grancharova, J. Kocijan, and T. A. Johansen · 2008
Cited alongside, same era.
Variational Bayesian Learning of Nonlinear Hidden State-Space Models for Model Predictive Control
T. Raiko and M. Tornio · 2009
Cited alongside, same era.
Efficient Computation of Optimal Actions
E. Todorov · 2009
Cited alongside, same era.
Robot Trajectory Optimization using Approximate Inference
M. Toussaint · 2009
Cited alongside, same era.
Probabilistic Differential Dynamic Programming
Y. Pan and E. Theodorou · 2014
Later among the works it cites.
Efficient Reinforcement Learning for Robots using Informative Simulated Priors
M. Cutler and J. P. How · 2015
Later among the works it cites.
Gaussian Processes for Data-Efficient Learning in Robotics and Control
M. P. Deisenroth, D. Fox, and C. E. Rasmussen · 2015
Later among the works it cites.
Human-Level Control through Deep Reinforcement Learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Later among the works it cites.
Learning-based Nonlinear Model Predictive Control to Improve Vision-based Mobile Robot Path Tracking
C. Ostafew, A. Schoellig, T. Barfoot, and J. Collier · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Efficient Reinforcement Learning using Gaussian Processes
M. P. Deisenroth · 2010
Cited alongside, same era.
PILCO: A Model-Based and Data-Efficient Approach to Policy Search
M. P. Deisenroth and C. E. Rasmussen · 2011
Cited alongside, same era.
Stability and Suboptimality Using Stabilizing Constraints
L. Grüne and J. Pannek · 2011
Cited alongside, same era.
Optimal Reinforcement Learning for Gaussian Systems
P. Hennig · 2011
Cited alongside, same era.
On Stochastic Optimal Control and Reinforcement Learning by Approximate Inference
K. Rawlik, M. Toussaint, and S. Vijayakumar · 2012
Cited alongside, same era.
Geometric Optimal Control: Theory, Methods and Examples
H. Schättler and U. Ledzewicz · 2012
Cited alongside, same era.
Y. Pan, E. Theodorou, and M. Kontitsis · 2015
Later among the works it cites.
Gaussian Processes for Data-Efficient Learning in Robotics and Control
M. P. Deisenroth, D. Fox, and C. E. Rasmussen · 2015
Later among the works it cites.
Reinforcement Learning with Unsupervised Auxiliary Tasks
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu · 2016
Later among the works it cites.
Gaussian Process-Based Predictive Control for Periodic Error Correction
E. D. Klenske, M. N. Zeilinger, B. Schölkopf, and P. Hennig · 2016
Later among the works it cites.
Modelling and Control of Dynamic Systems Using Gaussian Process Models
J. Kocijan · 2016
Later among the works it cites.
Robust Constrained Learning-based NMPC enabling reliable mobile robot path tracking
C. Ostafew, A. Schoellig, T. Barfoot · 2016
Later among the works it cites.
Mastering the Game of Go with Deep Neural Networks and Tree Search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis · 2016
Later among the works it cites.
Collective Robot Reinforcement Learning with Distributed Asynchronous Guided Policy Search
A. Yahya, A. Li, M. Kalakrishnan, Y. Chebotar, and S. Levine · 2016
Later among the works it cites.
GP-ILQG: Data-driven Robust Optimal Control for Uncertain Nonlinear Dynamical Systems
G. Lee, S. S. Srinivasa, and M. T. Mason · 2017
Closest in time.
Data-driven Demand Response Modeling and Control of Buildings with Gaussian Processes
T. X. Nghiem and C. N. Jones · 2017
Closest in time.
Survey of Model-Based Reinforcement Learning: Applications on Robotics
A. S. Polydoros and L. Nalpantidis · 2017
Closest in time.