Fetching the paper…
Reading the bibliography…
Safety is one of the biggest concerns to applying reinforcement learning (RL) to the physical world.
2742. a mean value theorem
Thomas Muirhead Flett · 1958
Earlier work this paper cites.
On hölder’s inequality
EF Beckenbach · 1966
Earlier work this paper cites.
Real-time obstacle avoidance for manipulators and mobile robots
Oussama Khatib · 1986
Earlier work this paper cites.
Noise reduction in chaotic time-series data: A survey of common methods
Eric J Kostelich and Thomas Schreiber · 1993
Earlier work this paper cites.
Gaussian processes for machine learning
Christopher K Williams and Carl Edward Rasmussen · 2006
Earlier work this paper cites.
Gaussian process optimization in the bandit setting: No regret and experimental design
Niranjan Srinivas, Andreas Krause, Sham M Kakade, and Matthias Seeger · 2009
Earlier work this paper cites.
Information-theoretic regret bounds for gaussian process optimization in the bandit setting
Niranjan Srinivas, Andreas Krause, Sham M Kakade, and Matthias W Seeger · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Reactive sliding-mode algorithm for collision avoidance in robotic systems
Luis Gracia, Fabricio Garelli, and Antonio Sala · 2013
Earlier work this paper cites.
Control barrier function based quadratic programs with application to adaptive cruise control
Aaron D Ames, Jessy W Grizzle, and Paulo Tabuada · 2014
Earlier work this paper cites.
Control in a safe set: Addressing safety in human-robot interactions
Changliu Liu and Masayoshi Tomizuka · 2014
Earlier work this paper cites.
Kinematic and dynamic vehicle models for autonomous driving control design
Jason Kong, Mark Pfeiffer, Georg Schildbach, and Francesco Borrelli · 2015
Earlier work this paper cites.
Dropout as a bayesian approximation: Representing model uncertainty in deep learning
Yarin Gal and Zoubin Ghahramani · 2016
Earlier work this paper cites.
Constrained policy optimization
Joshua Achiam, David Held, Aviv Tamar, and Pieter Abbeel · 2017
Earlier work this paper cites.
Safe model-based reinforcement learning with stability guarantees
Felix Berkenkamp, Matteo Turchetta, Angela Schoellig, and Andreas Krause · 2017
Cited alongside, same era.
Risk-constrained reinforcement learning with percentile risk criteria
Yinlam Chow, Mohammad Ghavamzadeh, Lucas Janson, and Marco Pavone · 2017
Cited alongside, same era.
On kernelized multi-armed bandits
Sayak Ray Chowdhury and Aditya Gopalan · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Later among the works it cites.
Stochastic variance reduction for deep q-learning
Wei-Ye Zhao, Xi-Ya Guan, Yang Liu, Xiaoming Zhao, and Jian Peng · 2019
Later among the works it cites.
nuscenes: A multimodal dataset for autonomous driving
Holger Caesar, Varun Bankiti, Alex H Lang, Sourabh Vora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom · 2020
Later among the works it cites.
Shieldnn: A provably safe nn filter for unsafe nn controllers
James Ferlez, Mahmoud Elnaggar, Yasser Shoukry, and Cody Fleming · 2020
Later among the works it cites.
Generating robust supervision for learning-based visual navigation using hamilton-jacobi reachability
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gal Dalal, Krishnamurthy Dvijotham, Matej Vecerik, Todd Hester, Cosmin Paduraru, and Yuval Tassa · 2018
Cited alongside, same era.
A general safety framework for learning-based control in uncertain robotic systems
Jaime F Fisac, Anayo K Akametalu, Melanie N Zeilinger, Shahab Kaynama, Jeremy Gillula, and Claire J Tomlin · 2018
Cited alongside, same era.
The apolloscape dataset for autonomous driving
Xinyu Huang, Xinjing Cheng, Qichuan Geng, Binbin Cao, Dingfu Zhou, Peng Wang, Yuanqing Lin, and Ruigang Yang · 2018
Cited alongside, same era.
Gaussian processes and kernel methods: A review on connections and equivalences
Motonobu Kanagawa, Philipp Hennig, Dino Sejdinovic, and Bharath K Sriperumbudur · 2018
Cited alongside, same era.
Safe exploration and optimization of constrained mdps using gaussian processes
Akifumi Wachi, Yanan Sui, Yisong Yue, and Masahiro Ono · 2018
Cited alongside, same era.
Human motion prediction using semi-adaptable neural networks
Yujiao Cheng, Weiye Zhao, Changliu Liu, and Masayoshi Tomizuka · 2019
Cited alongside, same era.
Safely learning to control the constrained linear quadratic regulator
Sarah Dean, Stephen Tu, Nikolai Matni, and Benjamin Recht · 2019
Cited alongside, same era.
Anjian Li, Somil Bansal, Georgios Giovanis, Varun Tolani, Claire Tomlin, and Mo Chen · 2020
Later among the works it cites.
Experimental evaluation of human motion prediction toward safe and efficient human robot collaboration
Weiye Zhao, Liting Sun, Changliu Liu, and Masayoshi Tomizuka · 2020
Later among the works it cites.
Armin Lederer, Jonas Umlauft, and Sandra Hirche · 2021
Later among the works it cites.
Safe adaptation with multiplicative uncertainties using robust safe set algorithm
Charles Noren, Weiye Zhao, and Changliu Liu · 2021
Later among the works it cites.
Wcsac: Worst-case soft actor critic for safety-constrained reinforcement learning
Qisong Yang, Thiago D Simão, Simon H Tindemans, and Matthijs TJ Spaan · 2021
Later among the works it cites.
Model-free safe control for zero-violation reinforcement learning
Weiye Zhao, Tairan He, and Changliu Liu · 2021
Later among the works it cites.
Persistently feasible robust safe control by safety index synthesis and convex semi-infinite programming
Tianhao Wei, Shucheng Kang, Weiye Zhao, and Changliu Liu · 2022
Closest in time.
Hybrid task constrained planner for robot manipulator in confined environment
Yifan Sun, Weiye Zhao, and Changliu Liu · 2023
Closest in time.
State-wise safe reinforcement learning: A survey
Weiye Zhao, Tairan He, Rui Chen, Tianhao Wei, and Changliu Liu · 2023
Closest in time.