Fetching the paper…
Reading the bibliography…
Despite the tremendous success of Reinforcement Learning (RL) algorithms in simulation environments, applying RL to real-world applications still faces many challenges.
Nonlinear Systems
H.K. Khalil and J.W. Grizzle · 2002
Earlier work this paper cites.
Pac model-free reinforcement learning
Alexander L. Strehl, Lihong Li, Eric Wiewiora, John Langford, and Michael L. Littman · 2006
Earlier work this paper cites.
Gaussian processes for machine learning
Christopher K Williams and Carl Edward Rasmussen · 2006
Earlier work this paper cites.
Control barrier function based quadratic programs with application to adaptive cruise control
Aaron D Ames, Jessy W Grizzle, and Paulo Tabuada · 2014
Earlier work this paper cites.
Control in a safe set: Addressing safety in human-robot interactions
Changliu Liu and Masayoshi Tomizuka · 2014
Earlier work this paper cites.
Kinematic and dynamic vehicle models for autonomous driving control design
Jason Kong, Mark Pfeiffer, Georg Schildbach, and Francesco Borrelli · 2015
Earlier work this paper cites.
Constrained policy optimization
Joshua Achiam, David Held, Aviv Tamar, and Pieter Abbeel · 2017
Earlier work this paper cites.
Safe model-based reinforcement learning with stability guarantees
Felix Berkenkamp, Matteo Turchetta, Angela Schoellig, and Andreas Krause · 2017
Earlier work this paper cites.
On kernelized multi-armed bandits
Sayak Ray Chowdhury and Aditya Gopalan · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Superhuman ai for heads-up no-limit poker: Libratus beats top professionals
Noam Brown and Tuomas Sandholm · 2018
Earlier work this paper cites.
A lyapunov-based approach to safe reinforcement learning
Yinlam Chow, Ofir Nachum, Edgar Duenez-Guzman, and Mohammad Ghavamzadeh · 2018
Earlier work this paper cites.
Safe exploration in continuous action spaces
Gal Dalal, Krishnamurthy Dvijotham, Matej Vecerik, Todd Hester, Cosmin Paduraru, and Yuval Tassa · 2018
Earlier work this paper cites.
A general safety framework for learning-based control in uncertain robotic systems
Jaime F Fisac, Anayo K Akametalu, Melanie N Zeilinger, Shahab Kaynama, Jeremy Gillula, and Claire J Tomlin · 2018
Earlier work this paper cites.
Accelerated primal-dual policy optimization for safe reinforcement learning
Qingkai Liang, Fanyu Que, and Eytan Modiano · 2018
Earlier work this paper cites.
arXiv preprint arXiv:1805.11074
Chen Tessler, Daniel J Mankowitz, and Shie Mannor · 2018
Earlier work this paper cites.
Safe exploration and optimization of constrained mdps using gaussian processes
Akifumi Wachi, Yanan Sui, Yisong Yue, and Masahiro Ono · 2018
Cited alongside, same era.
Value constrained model-free continuous control
Steven Bohez, Abbas Abdolmaleki, Michael Neunert, Jonas Buchli, Nicolas Heess, and Raia Hadsell · 2019
Cited alongside, same era.
End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks
Richard Cheng, Gábor Orosz, Richard M Murray, and Joel W Burdick · 2019
Cited alongside, same era.
Lyapunov-based safe policy optimization for continuous control
Yinlam Chow, Ofir Nachum, Aleksandra Faust, Edgar Duenez-Guzman, and Mohammad Ghavamzadeh · 2019
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Reachability-based trajectory safeguard (rts): A safe and fast reinforcement learning safety layer for continuous control
Yifei Simon Shao, Chao Chen, Shreyas Kousik, and Ram Vasudevan · 2021
Later among the works it cites.
Recovery rl: Safe reinforcement learning with learned recovery zones
Brijen Thananjeyan, Ashwin Balakrishna, Suraj Nair, Michael Luo, Krishnan Srinivasan, Minho Hwang, Joseph E Gonzalez, Julian Ibarz, Chelsea Finn, and Ken Goldberg · 2021
Later among the works it cites.
Model-free safe control for zero-violation reinforcement learning
Weiye Zhao, Tairan He, and Changliu Liu · 2021
Later among the works it cites.
Safe learning in robotics: From learning-based control to safe reinforcement learning
Lukas Brunke, Melissa Greeff, Adam W Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P Schoellig · 2022
Later among the works it cites.
A review of safe reinforcement learning: Methods, theory and applications
Shangding Gu, Long Yang, Yali Du, Guang Chen, Florian Walter, Jun Wang, Yaodong Yang, and Alois Knoll · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Approximation gradient error variance reduced optimization
W Zhao, Yang Liu, Xiaoming Zhao, J Qiu, and Jian Peng · 2019
Cited alongside, same era.
Stochastic variance reduction for deep q-learning
Wei-Ye Zhao, Xi-Ya Guan, Yang Liu, Xiaoming Zhao, and Jian Peng · 2019
Cited alongside, same era.
Conservative safety critics for exploration
Homanga Bharadhwaj, Aviral Kumar, Nicholas Rhinehart, Sergey Levine, Florian Shkurti, and Animesh Garg · 2020
Cited alongside, same era.
Shieldnn: A provably safe nn filter for unsafe nn controllers
James Ferlez, Mahmoud Elnaggar, Yasser Shoukry, and Cody Fleming · 2020
Cited alongside, same era.
Safe reinforcement learning in constrained markov decision processes
Akifumi Wachi and Yanan Sui · 2020
Cited alongside, same era.
Contact-rich trajectory generation in confined environments using iterative convex optimization
Wei-Ye Zhao, Suqin He, Chengtao Wen, and Changliu Liu · 2020
Cited alongside, same era.
Safe and sample-efficient reinforcement learning for clustered dynamic environments
Hongyi Chen and Changliu Liu · 2021
Cited alongside, same era.
Later among the works it cites.
Reinforcement learning with automated auxiliary loss search
Tairan He, Yuge Zhang, Kan Ren, Minghuan Liu, Che Wang, Weinan Zhang, Yuqing Yang, and Dongsheng Li · 2022
Later among the works it cites.
Joint synthesis of safety certificate and safe control policy using constrained reinforcement learning
Haitong Ma, Changliu Liu, Shengbo Eben Li, Sifa Zheng, and Jianyu Chen · 2022
Later among the works it cites.
Safe control with neural network dynamic models
Tianhao Wei and Changliu Liu · 2022
Later among the works it cites.
Persistently feasible robust safe control by safety index synthesis and convex semi-infinite programming
Tianhao Wei, Shucheng Kang, Weiye Zhao, and Changliu Liu · 2022
Later among the works it cites.
Evaluating model-free reinforcement learning toward safety-critical tasks
Linrui Zhang, Qin Zhang, Li Shen, Bo Yuan, Xueqian Wang, and Dacheng Tao · 2022
Later among the works it cites.
Provably safe tolerance estimation for robot arms via sum-of-squares programming
Weiye Zhao, Suqin He, and Changliu Liu · 2022
Later among the works it cites.
Probabilistic safeguard for reinforcement learning using safety index guided gaussian process models
Weiye Zhao, Tairan He, and Changliu Liu · 2022
Later among the works it cites.
Safety index synthesis via sum-of-squares programming
Weiye Zhao, Tairan He, Tianhao Wei, Simin Liu, and Changliu Liu · 2022
Later among the works it cites.
A hierarchical long short term safety framework for efficient robot manipulation under uncertainty
Suqin He, Weiye Zhao, Chuxiong Hu, Yu Zhu, and Changliu Liu · 2023
Closest in time.
Autocost: Evolving intrinsic cost for zero-violation reinforcement learning
Tairan He, Weiye Zhao, and Changliu Liu · 2023
Closest in time.