Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) has demonstrated remarkable performance in many continuous control tasks.
“Minimization of functions having Lipschitz continuous first partial derivatives”
Larry Armijo · 1966
Earlier work this paper cites.
“Real-time obstacle avoidance for manipulators and mobile robots”
Oussama Khatib · 1986
Earlier work this paper cites.
“Mujoco: A physics engine for model-based control”
Emanuel Todorov, Tom Erez and Yuval Tassa · 2012
Earlier work this paper cites.
“Reactive sliding-mode algorithm for collision avoidance in robotic systems”
Luis Gracia, Fabricio Garelli and Antonio Sala · 2013
Earlier work this paper cites.
“Safe policy iteration”
Matteo Pirotta, Marcello Restelli, Alessio Pecorino and Daniele Calandriello · 2013
Earlier work this paper cites.
“Control barrier function based quadratic programs with application to adaptive cruise control”
Aaron Ames, Jessy Grizzle and Paulo Tabuada · 2014
Earlier work this paper cites.
“Control in a safe set: Addressing safety in human-robot interactions”
Changliu Liu and Masayoshi Tomizuka · 2014
Earlier work this paper cites.
“A comprehensive survey on safe reinforcement learning”
Javier Garcıa and Fernando Fernández · 2015
Earlier work this paper cites.
“Kinematic and dynamic vehicle models for autonomous driving control design”
Jason Kong, Mark Pfeiffer, Georg Schildbach and Francesco Borrelli · 2015
Earlier work this paper cites.
“Constrained policy optimization”
Joshua Achiam, David Held, Aviv Tamar and Pieter Abbeel · 2017
Earlier work this paper cites.
“Safe Model-based Reinforcement Learning with Stability Guarantees”
Felix Berkenkamp, Matteo Turchetta, Angela Schoellig and Andreas Krause · 2017
Earlier work this paper cites.
“Proximal Policy Optimization Algorithms”, 2017
John Schulman et al · 2017
Earlier work this paper cites.
“Proximal Policy Optimization Algorithms”
John Schulman et al · 2017
Earlier work this paper cites.
“Bio-inspired genetic algorithms with formalized crossover operators for robotic applications”
Jie Zhang, Man Kang, Xiaojuan Li and Geng-yang Liu · 2017
Earlier work this paper cites.
“Safe exploration in continuous action spaces”
Gal Dalal et al · 2018
Earlier work this paper cites.
“A general safety framework for learning-based control in uncertain robotic systems”
Jaime Fisac et al · 2018
Earlier work this paper cites.
“Optlayer-practical constrained optimization for deep reinforcement learning in the real world”
Tu-Hoa Pham, Giovanni De and Ryuki Tachibana · 2018
Cited alongside, same era.
“End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks”
Richard Cheng, Gábor Orosz, Richard Murray and Joel Burdick · 2019
Cited alongside, same era.
“Lyapunov-based safe policy optimization for continuous control”
Yinlam Chow et al · 2019
Cited alongside, same era.
“Smoothing Policies and Safe Policy Gradients”
Matteo Papini, Matteo Pirotta and Marcello Restelli · 2019
Cited alongside, same era.
“Benchmarking safe exploration in deep reinforcement learning”
Alex Ray, Joshua Achiam and Dario Amodei · 2019
“Learning predictive safety filter via decomposition of robust invariant set”
Zeyang Li, Chuxiong Hu, Weiye Zhao and Changliu Liu · 2023
Later among the works it cites.
“State-wise safe reinforcement learning: a survey”
Weiye Zhao et al · 2023
Later among the works it cites.
“Safety Filtering While Training: Improving the Performance and Sample Efficiency of Reinforcement Learning Agents”
Federico Bejarano, Lukas Brunke and Angela. Schoellig · 2024
Closest in time.
“Multimodal Safe Control for Human-Robot Interaction”, 2024
Ravi Pandya, Tianhao Wei and Changliu Liu · 2024
Closest in time.
“State-wise Constrained Policy Optimization”, 2024
Weiye Zhao et al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Shieldnn: A provably safe nn filter for unsafe nn controllers”
James Ferlez, Mahmoud Elnaggar, Yasser Shoukry and Cody Fleming · 2020
Cited alongside, same era.
“Adaptive safety with control barrier functions”
Andrew Taylor and Aaron Ames · 2020
Cited alongside, same era.
“Review of digital twin about concepts, technologies, and industrial applications”
Mengnan Liu, Shuiliang Fang, Huiyue Dong and Cunzhi Xu · 2021
Cited alongside, same era.
“Learning to Build High-Fidelity and Robust Environment Models”
Weinan Zhang et al · 2021
Cited alongside, same era.
“Model-free safe control for zero-violation reinforcement learning”
Weiye Zhao, Tairan He and Changliu Liu · 2021
Cited alongside, same era.
“Edge-assisted Collaborative Digital Twin for Safety-Critical Robotics in Industrial IoT”, 2022
Sumit. Das, Mohammad Uddin and Sabur Baidya · 2022
Cited alongside, same era.
“Safe Reinforcement Learning Using Robust Control Barrier Functions”, 2022
Yousef Emam et al · 2022
Cited alongside, same era.
“Absolute Policy Optimization: Enhancing Lower Probability Bound of Performance with High Confidence”
Weiye Zhao et al · 2024
Closest in time.
Weiye Zhao et al · 2024
Closest in time.
“GUARD: A Safe Reinforcement Learning Benchmark”, 2024
Weiye Zhao et al · 2024
Closest in time.
“Real-is-Sim: Bridging the Sim-to-Real Gap with a Dynamic Digital Twin”, 2025
Jad Abou-Chakra et al · 2025
Closest in time.
“Towards General-Purpose Model-Free Reinforcement Learning”, 2025
Scott Fujimoto et al · 2025
Closest in time.
“A Survey on Deep Reinforcement Learning Applications in Autonomous Systems: Applications, Open Challenges, and Future Directions”
Shruti Govinda, Bouziane Brik and Saad Harous · 2025
Closest in time.
“Continual Learning and Lifting of Koopman Dynamics for Linear Control of Legged Robots”, 2025
Feihan Li et al · 2025
Closest in time.
“Risk-Sensitive Reinforcement Learning With Exponential Criteria”
Erfaun Noorani, Christos. Mavridis and John. Baras · 2025
Closest in time.
“SPARK: A Modular Benchmark for Humanoid Robot Safety”, 2025
Yifan Sun et al · 2025
Closest in time.
“Passivity-Centric Safe Reinforcement Learning for Contact-Rich Robotic Tasks”, 2025
Heng Zhang, Gokhan Solak, Sebastian Hjorth and Arash Ajoudani · 2025
Closest in time.