Fetching the paper…
Reading the bibliography…
The proven efficacy of learning-based control schemes strongly motivates their application to robotic systems operating in the physical world.
“Theory of ordinary differential equations”
Earl Coddington and Norman Levinson · 1955
Earlier work this paper cites.
“Differential games and representation formulas for solutions of Hamilton-Jacobi-Isaacs equations”
L.. Evans and P.. Souganidis · 1984
Earlier work this paper cites.
“Efficient implementation of essentially non-oscillatory shock-capturing schemes”
Chi-Wang Shu and Stanley Osher · 1988
Earlier work this paper cites.
“Differential Games with Maximum Cost”
E.N. Barron · 1990
Earlier work this paper cites.
“Numerical computation of multivariate normal probabilities”
A Genz · 1992
Earlier work this paper cites.
“Level set methods and dynamic implicit surfaces”
Stanley Osher and Ronald Fedkiw · 2003
Earlier work this paper cites.
“Lyapunov design for safe reinforcement learning”
TJ Perkins and AG Barto · 2003
Earlier work this paper cites.
“Safety verification of hybrid systems using barrier certificates”
Stephen Prajna and Ali Jadbabaie · 2004
Earlier work this paper cites.
“Risk-sensitive reinforcement learning applied to control under constraints”
P. Geibel and F. Wysotzki · 2005
Earlier work this paper cites.
“A time-dependent Hamilton-Jacobi formulation of reachable sets for continuous dynamic games”
Ian. Mitchell, A.. Bayen and C.. Tomlin · 2005
Earlier work this paper cites.
“A Toolbox of Hamilton-Jacobi Solvers for Analysis of Nondeterministic Continuous and Hybrid Systems”
Ian Mitchell and Jeremy Templeton · 2005
Earlier work this paper cites.
“Gaussian processes for machine learning”
Carl Rasmussen and Christopher Williams · 2006
Earlier work this paper cites.
“An application of reinforcement learning to aerobatic helicopter flight”
Pieter Abbeel, Adam Coates, Morgan Quigley and Andrew Ng · 2007
Earlier work this paper cites.
“Learning for control from multiple demonstrations”
Adam Coates, Pieter Abbeel and Andrew. Ng · 2008
Earlier work this paper cites.
“Policy search via the signed derivative.”
JZ Kolter and AY Ng · 2009
Cited alongside, same era.
“Unmanned aircraft systems”
Alan Hobbs · 2010
Cited alongside, same era.
“A probabilistic approach to mixed open-loop and closed-loop control, with application to extreme autonomous driving”
J Kolter, Christian Plagemann, David Jackson, Andrew Ng and Sebastian Thrun · 2010
Cited alongside, same era.
“A simple learning strategy for high-speed quadrocopter multi-flips”
Sergei Lupashin, Angela Sch“”ollig, Michael Sherback and Raffaello D’Andrea · 2010
Cited alongside, same era.
“Toward reachability-based controller design for hybrid systems in robotics”
Jerry Ding, Jeremy Gillula, Haomiao Huang, Michael Vitus, Wei Zhang and Claire Tomlin · 2011
Cited alongside, same era.
“Feedback controller parameterizations for Reinforcement Learning”
John. Roberts, Ian. Manchester and Russ Tedrake · 2011
“Reach-avoid problems with time-varying dynamics, targets and constraints”
Jaime. Fisac, Mo Chen, Claire. Tomlin and S. Sastry · 2015
Later among the works it cites.
“A Comprehensive Survey on Safe Reinforcement Learning”
Javier Garc“’a and Fernando Fern“’andez · 2015
Later among the works it cites.
“Scalable safety-preserving robust control synthesis for continuous-time linear systems”
Shahab Kaynama, Ian. Mitchell, Meeko Oishi and Guy. Dumont · 2015
Later among the works it cites.
“Human-level control through deep reinforcement learning”
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei Rusu, Joel Veness, Marc Bellemare, Alex Graves, Martin Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg and Demis Hassabis · 2015
Later among the works it cites.
“Trust Region Policy Optimization”
John Schulman, Sergey Levine, Michael Jordan and Pieter Abbeel · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Guaranteed Safe Online Learning via Reachability: tracking a ground target using a quadrotor”
Jeremy. Gillula and Claire. Tomlin · 2012
Cited alongside, same era.
“Reducing Conservativeness in Safety Guarantees by Learning Disturbances Online: Iterated Guaranteed Safe Online Learning.”
JH Gillula and CJ Tomlin · 2012
Cited alongside, same era.
“Safe exploration in Markov decision processes”
TM Moldovan and P Abbeel · 2012
Cited alongside, same era.
“Compositional safety analysis using barrier certificates”
Christoffer Sloth, George. Pappas and Rafael Wisniewski · 2012
Cited alongside, same era.
“Provably safe and robust learning-based model predictive control”
Anil Aswani, Humberto Gonzalez, S. Sastry and Claire Tomlin · 2013
Cited alongside, same era.
“A modified Riccati transformation for decentralized computation of the viability kernel under LTI dynamics”
Shahab Kaynama and Meeko Oishi · 2013
Cited alongside, same era.
“Safe learning of regions of attraction for uncertain, nonlinear systems with Gaussian processes”
Felix Berkenkamp, Riccardo Moriconi, Angela. Schoellig and Andreas Krause · 2016
Later among the works it cites.
“Fast Reachable Set Approximations via State Decoupling Disturbances”
Mo Chen, Sylvia Herbert and Claire. Tomlin · 2016
Later among the works it cites.
“Transfer from Simulation to Real World through Learning Deep Inverse Dynamics Model”
Paul Christiano, Zain Shah, Igor Mordatch, Jonas Schneider, Trevor Blackwell, Joshua Tobin, Pieter Abbeel and Wojciech Zaremba · 2016
Later among the works it cites.
“Algorithms for overcoming the curse of dimensionality for certain Hamilton-Jacobi equations arising in control theory and elsewhere”
J“’er“ˆome Darbon and Stanley Osher · 2016
Later among the works it cites.
“Safe Model-based Reinforcement Learning with Stability Guarantees”
Felix Berkenkamp, Matteo Turchetta, Angela. Schoellig and Andreas Krause · 2017
Closest in time.
“FaSTrack: a Modular Framework for Fast and Guaranteed Safe Motion Planning”
Sylvia Herbert, Mo Chen, Soojean Han, Somil Bansal, Jaime Fisac and Claire Tomlin · 2017
Closest in time.
“Adversarial Attacks on Neural Network Policies”
Sandy Huang, Nicolas Papernot, Ian Goodfellow, Yan Duan and Pieter Abbeel · 2017
Closest in time.
“Time-Optimal Collaborative Guidance using the Generalized Hopf Formula”
Matthew. Kirchner, Robert Mar, Gary Hewer, J“’er“ˆome Darbon, Stanley Osher and Y.. Chow · 2017
Closest in time.