Fetching the paper…
Reading the bibliography…
Reach-avoid optimal control problems, in which the system must reach certain goal conditions while staying clear of unacceptable failure modes, are central to safety and liveness assurance for autonomous robotic systems, but their exact solutions are intractable for complex dynamics and environments.
“Theory of ordinary differential equations”
Earl Coddington and Norman Levinson · 1955
Earlier work this paper cites.
“Differential Games”
Rufus Isaacs · 1965
Earlier work this paper cites.
“Defining Liveness”
Bowen Alpern and Fred. Schneider · 1985
Earlier work this paper cites.
“Reinforcement learning is direct adaptive optimal control”
R.. Sutton, A.. Barto and R.. Williams · 1992
Earlier work this paper cites.
“Q-learning”
Christopher… Watkins and Peter Dayan · 1992
Earlier work this paper cites.
“Asynchronous Stochastic Approximation and Q-Learning”
John. Tsitsiklis · 1994
Earlier work this paper cites.
“The flexible, extensible and efficient Toolbox of Level Set Methods”
Ian. Mitchell · 2008
Earlier work this paper cites.
“DeepReach: A Deep Learning Approach to High-Dimensional Reachability”, 2020
Somil Bansal and Claire Tomlin · 2011
Earlier work this paper cites.
“Reach-Avoid Problems with Time-Varying Dynamics, Targets and Constraints”
Jaime. Fisac, Mo Chen, Claire. Tomlin and S. Sastry · 2015
Earlier work this paper cites.
“Deep Reinforcement Learning with Double Q-Learning”
Hado Van, Arthur Guez and David Silver · 2015
Cited alongside, same era.
“Continuous control with deep reinforcement learning.”
Timothy. Lillicrap et al · 2016
Cited alongside, same era.
Greg Brockman et al · 2016
Cited alongside, same era.
“Collision Between a Car Operating With Automated Vehicle Control Systems and a Tractor-Semitrailer Truck Near Williston, Florida, May 7, 2016.”, 2017
NTSB · 2017
Cited alongside, same era.
“Hamilton-Jacobi Reachability: A Brief Overview and Recent Advances”
S. Bansal, M. Chen, S. Herbert and C.. Tomlin · 2017
Cited alongside, same era.
“In the Air with Zipline’s Medical Delivery Drones” (Accessed on 2021/02/08.)
Evan Ackerman and Michael Koziol · 2019
Later among the works it cites.
“A Classification-based Approach for Approximate Reachability”
V. Rubies-Royo, D. Fridovich-Keil, S. Herbert and C.. Tomlin · 2019
Later among the works it cites.
“Bridging hamilton-jacobi safety analysis and reinforcement learning”
Jaime Fisac et al · 2019
Later among the works it cites.
“Safe reinforcement learning via online shielding”
Osbert Bastani · 2019
Later among the works it cites.
“Waymo Becomes First Company to Launch Driverless Ride-Hailing to Public - The Washington Post”
Faiz Siddiqui · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“FaSTrack: a Modular Framework for Fast and Guaranteed Safe Motion Planning”
Sylvia Herbert et al · 2017
Cited alongside, same era.
“Recursive Regression with Neural Networks: Approximating the HJI PDE Solution”, 2017
Vicenç Rubies-Royo and Claire Tomlin · 2017
Cited alongside, same era.
“A general safety framework for learning-based control in uncertain robotic systems”
Jaime. Fisac* et al · 2018
Cited alongside, same era.
“Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor”
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel and Sergey Levine · 2018
Cited alongside, same era.
NTSB · 2020
Later among the works it cites.
“Collision Between Car Operating with Partial Driving Automation and Truck-Tractor Semitrailer, Delray Beach, Florida, March 1, 2019” (Accessed on 2021/02/08.), 2020
NTSB · 2020
Later among the works it cites.
“Robust Model Predictive Shielding for Safe Reinforcement Learning with Stochastic Dynamics”
S. Li and O. Bastani · 2020
Later among the works it cites.
“Algorithms for Verifying Deep Neural Networks”
Changliu Liu et al · 2021
Closest in time.