Fetching the paper…
Reading the bibliography…
Model-free Reinforcement Learning (RL) offers an attractive approach to learn control policies for high-dimensional systems, but its relatively poor sample complexity often forces training in simulated environments.
E. A. Coddington and N. Levinson, Theory of Ordinary Differential Equations . McGraw-Hill Inc., 1955
1955
Earlier work this paper cites.
M. G. Crandall and P.-l. Lions, “Viscosity solutions of Hamilton-Jacobi equations,” Trans. American Mathematical Society , 1983
1983
Earlier work this paper cites.
L. Ljung, System identification . Springer, 1998
1998
Earlier work this paper cites.
I. M. Mitchell and C. J. Tomlin, “Overapproximating Reachable Sets by Hamilton-Jacobi Projections,” J. Scientific Computing , 2003
2003
Earlier work this paper cites.
S. M. LaValle, Planning algorithms . Cambridge university press, 2006
2006
Earlier work this paper cites.
I. M. Mitchell, “The Flexible, Extensible and Efficient Toolbox of Level Set Methods,” J. Scientific Computing , 2008
2008
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” International Conference on Machine Learning (ICML) , 2009
2009
Earlier work this paper cites.
J. H. Gillula, H. Huang, M. P. Vitus, and C. J. Tomlin, “Design of guaranteed safe maneuvers using reachable sets: Autonomous quadrotor aerobatics in theory and practice,” IEEE International Conference on Robotics and Automation (ICRA) , 2010
2010
Earlier work this paper cites.
M. Deisenroth and C. E. Rasmussen, “PILCO: A model-based and data-efficient approach to policy search,” International Conference on Machine Learning (ICML) , 2011
2011
Earlier work this paper cites.
A. Karpathy and M. Van De Panne, “Curriculum learning for motor skills,” Canadian Conference on Artificial Intelligence , 2012
2012
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing atari with deep reinforcement learning,” Neural Information Processing Systems (NIPS) , 2013
2013
Earlier work this paper cites.
A. Baranes and P.-Y. Oudeyer, “Active learning of inverse models with intrinsically motivated goal exploration in robots,” Robotics and Autonomous Systems , 2013
2013
Earlier work this paper cites.
D. J. Webb and J. van den Berg, “Kinodynamic RRT*: Asymptotically optimal motion planning for robots with linear dynamics,” IEEE International Conference on Robotics and Automation (ICRA) , 2013
2013
Earlier work this paper cites.
S. Levine and V. Koltun, “Guided policy search,” International Conference on Machine Learning (ICML) , 2013
2013
Earlier work this paper cites.
S. Levine and P. Abbeel, “Learning neural network policies with guided policy search under unknown dynamics,” Neural Information Processing Systems (NIPS) , 2014
2014
Earlier work this paper cites.
W. Zaremba and I. Sutskever, “Learning to execute,” arXiv:1410.4615 , 2014
2014
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , 2015
2015
Cited alongside, same era.
M. Althoff, “An Introduction to CORA 2015,” Proc. Workshop on Applied Verification for Continuous and Hybrid Systems , 2015
2015
Cited alongside, same era.
L. Janson, E. Schmerling, A. Clark, and M. Pavone, “Fast marching tree: A fast marching sampling-based method for optimal motion planning in many dimensions,” International Journal of Robotics Research , 2015
2015
Cited alongside, same era.
2017
Later among the works it cites.
C. Florensa, D. Held, M. Wulfmeier, M. Zhang, and P. Abbeel, “Reverse curriculum generation for reinforcement learning,” Conference on Robot Learning (CoRL) , 2017
2017
Later among the works it cites.
S. Bansal, R. Calandra, S. Levine, and C. Tomlin, “MBMF: Model-based priors for model-free reinforcement learning,” Conference on Robot Learning (CoRL) , 2017
2017
Later among the works it cites.
Y. Chebotar, K. Hausman, M. Zhang, G. Sukhatme, S. Schaal, and S. Levine, “Combining model-based and model-free updates for trajectory-centric reinforcement learning,” International Conference on Machine Learning (ICML) , 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , 2016
2016
Cited alongside, same era.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” Journal of Machine Learning Research , 2016
2016
Cited alongside, same era.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” International Conference on Learning Representations (ICLR) , 2016
2016
Cited alongside, same era.
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy, “Deep exploration via bootstrapped DQN,” Neural Information Processing Systems (NIPS) , 2016
2016
Cited alongside, same era.
S. Gu, T. Lillicrap, I. Sutskever, and S. Levine, “Continuous deep Q-learning with model-based acceleration,” International Conference on Machine Learning (ICML) , 2016
2016
Cited alongside, same era.
R. Houthooft, X. Chen, Y. Duan, J. Schulman, F. De Turck, and P. Abbeel, “Vime: Variational information maximizing exploration,” Neural Information Processing Systems (NIPS) , 2016
2016
Cited alongside, same era.
C. Fan, B. Qi, S. Mitra, M. Viswanathan, and P. S. Duggirala, “Automatic Reachability Analysis for Nonlinear Hybrid Models with C2E2,” Computer Aided Verification , 2016
2016
Cited alongside, same era.
M. Chen, S. Herbert, and C. J. Tomlin, “Fast reachable set approximations via state decoupling disturbances,” IEEE Conf. Decision and Control , 2016
2016
Cited alongside, same era.
2017
Later among the works it cites.
A. Graves, M. G. Bellemare, J. Menick, R. Munos, and K. Kavukcuoglu, “Automated curriculum learning for neural networks,” International Conference on Machine Learning (ICML) , 2017
2017
Later among the works it cites.
S. Bansal, M. Chen, S. Herbert, and C. J. Tomlin, “Hamilton-Jacobi reachability: A brief overview and recent advances,” IEEE Conference on Decision and Control (CDC) , 2017
2017
Later among the works it cites.
S. Singh, A. Majumdar, J.-J. Slotine, and M. Pavone, “Robust online motion planning via contraction theory and convex optimization,” IEEE International Conference on Robotics and Automation (ICRA) , 2017
2017
Later among the works it cites.
B. Ichter, E. Schmerling, and M. Pavone, “Group marching tree: Sampling-based approximately optimal motion planning on gpus,” IEEE International Conference on Robotic Computing (IRC) , 2017
2017
Later among the works it cites.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” Robotics: Science and Systems (RSS) , 2018
2018
Closest in time.
G. Thomas, M. Chien, A. Tamar, J. A. Ojea, and P. Abbeel, “Learning robotic assembly from CAD,” IEEE International Conference on Robotics and Automation (ICRA) , 2018
2018
Closest in time.
M. Chen and C. J. Tomlin, “Hamilton-Jacobi Reachability: Some Recent Theoretical Advances and Applications in Unmanned Airspace Management,” Annual Review of Control, Robotics, and Autonomous Systems , 2018
2018
Closest in time.
M. Chen, S. L. Herbert, M. Vashishtha, S. Bansal, and C. J. Tomlin, “Decomposition of Reachable Sets and Tubes for a Class of Nonlinear Systems,” IEEE Transactions on Automatic Control , 2018
2018
Closest in time.
J. Harrison, A. Sharma, and M. Pavone, “Meta-learning priors for efficient online bayesian regression,” Workshop on the Algorithmic Foundations of Robotics (WAFR) , 2018
2018
Closest in time.