Fetching the paper…
Reading the bibliography…
Reinforcement learning algorithms need exploration to learn.
Constrained Markov decision processes
Eitan Altman. 1999 · 1999
Earlier work this paper cites.
Iterative Linear Quadratic Regulator Design for Nonlinear Biological Movement Systems. In International Conference on Informatics in Control, Automation and Robotics
Weiwei Li and Emanuel Todorov. 2004 · 2004
Earlier work this paper cites.
Ellipsoidal Toolbox (ET). In IEEE Conference on Decision and Control (CDC, Vol. 45) . 1498–1503
Alex A. Kurzhanskiy and Pravin Varaiya. 2006 · 2006
Earlier work this paper cites.
Learning to be safe: Deep rl with a safety critic
Krishnan Srinivasan, Benjamin Eysenbach, Sehoon Ha, Jie Tan, and Chelsea Finn. 2020 · 2010
Earlier work this paper cites.
Safe Exploration in Markov Decision Processes. In Proceedings of the 29th International Coference on Machine Learning (ICML) . 1451–1458
Teodor Mihai Moldovan and Pieter Abbeel. 2012 · 2012
Earlier work this paper cites.
Safe Exploration Techniques for Reinforcement Learning – An Overview. In Modelling and Simulation for Autonomous Systems (MESAS) , Jan Hodicky (Ed.). Springer International Publishing, Cham, 357–375
Martin Pecka and Tomas Svoboda. 2014 · 2014
Earlier work this paper cites.
A Comprehensive Survey on Safe Reinforcement Learning
Javier García, Fern, and o Fernández. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis. 2015 · 2015
Earlier work this paper cites.
Concrete Problems in AI Safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul F. Christiano, John Schulman, and Dan Mané. 2016 · 2016
Earlier work this paper cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis. 2016 · 2016
Earlier work this paper cites.
Constrained Policy Optimization. In Proceedings of the 34th International Conference on Machine Learning (ICML) , Doina Precup and Yee Whye Teh (Eds.). 22–31
Joshua Achiam, David Held, Aviv Tamar, and Pieter Abbeel. 2017 · 2017
Earlier work this paper cites.
Safe Model-based Reinforcement Learning with Stability Guarantees. In Advances in Neural Information Processing Systems 30 (NeurIPS) . 908–918
Felix Berkenkamp, Matteo Turchetta, Angela Schoellig, and Andreas Krause. 2017 · 2017
Earlier work this paper cites.
Risk-constrained reinforcement learning with percentile risk criteria
Yinlam Chow, Mohammad Ghavamzadeh, Lucas Janson, and Marco Pavone. 2017 · 2017
Earlier work this paper cites.
Jan Leike, Miljan Martic, Victoria Krakovna, Pedro A. Ortega, Tom Everitt, Andrew Lefrancq, Laurent Orseau, and Shane Legg. 2017 · 2017
Cited alongside, same era.
Model Predictive Control: Theory, Computation, and Design (2nd edition ed.)
James Blake Rawlings, David Q Mayne, and Moritz Diehl. 2017 · 2017
Cited alongside, same era.
Deep Reinforcement Learning in a Handful of Trials Using Probabilistic Dynamics Models. In Advances in Neural Information Processing Systems 31 (NeurIPS) . 4759–4770
Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine. 2018 · 2018
Cited alongside, same era.
Safe Exploration in Continuous Action Spaces
Gal Dalal, Krishnamurthy Dvijotham, Matej Vecerik, Todd Hester, Cosmin Paduraru, and Yuval Tassa. 2018 · 2018
Cited alongside, same era.
Learning-Based Model Predictive Control for Safe Exploration. In IEEE Conference on Decision and Control (CDC) . 6059–6066
Benchmarking Safe Exploration in Deep Reinforcement Learning
Alex Ray, Joshua Achiam, and Dario Amodei. 2019 · 2019
Later among the works it cites.
Reward Constrained Policy Optimization. In 7th International Conference on Learning Representations (ICLR)
Chen Tessler, Daniel J. Mankowitz, and Shie Mannor. 2019 · 2019
Later among the works it cites.
Learning control barrier functions from expert demonstrations. In 2020 59th IEEE Conference on Decision and Control (CDC) . IEEE, 3717–3724
Alexander Robey, Haimin Hu, Lars Lindemann, Hanwen Zhang, Dimos V Dimarogonas, Stephen Tu, and Nikolai Matni. 2020 · 2020
Later among the works it cites.
Projection-Based Constrained Policy Optimization. In International Conference on Learning Representations
Tsung-Yen Yang, Justinian Rosca, Karthik Narasimhan, and Peter J. Ramadge. 2020 · 2020
Later among the works it cites.
Challenges of real-world reinforcement learning: definitions, benchmarks and analysis
Gabriel Dulac-Arnold, Nir Levine, Daniel J. Mankowitz, Jerry Li, Cosmin Paduraru, Sven Gowal, and Todd Hester. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Torsten Koller, Felix Berkenkamp, Matteo Turchetta, and Andreas Krause. 2018 · 2018
Cited alongside, same era.
Accelerated primal-dual policy optimization for safe reinforcement learning
Qingkai Liang, Fanyu Que, and Eytan Modiano. 2018 · 2018
Cited alongside, same era.
Neural Network Dynamics for Model-Based Deep Reinforcement Learning with Model-Free Fine-Tuning. In IEEE International Conference on Robotics and Automation (ICRA) . 7559–7566
Anusha Nagabandi, Gregory Kahn, Ronald S. Fearing, and Sergey Levine. 2018 · 2018
Cited alongside, same era.
Linear Model Predictive Safety Certification for Learning-Based Control. In IEEE Conference on Decision and Control (CDC) . 7130–7135
K. P. Wabersich and M. N. Zeilinger. 2018 · 2018
Cited alongside, same era.
Control barrier functions: Theory and applications. In European Control Conference (ECC) . 3420–3431
Aaron D Ames, Samuel Coogan, Magnus Egerstedt, Gennaro Notomista, Koushil Sreenath, and Paulo Tabuada. 2019 · 2019
Cited alongside, same era.
End-to-End Safe Reinforcement Learning through Barrier Functions for Safety-Critical Continuous Control Tasks. In Proceedings of the 33rd AAAI Conference on Artificial Intelligence (AAAI, Vol. 33) . 3387–3395
Richard Cheng, Gábor Orosz, Richard M. Murray, and Joel W. Burdick. 2019 · 2019
Cited alongside, same era.
Lyapunov-based safe policy optimization for continuous control
Yinlam Chow, Ofir Nachum, Aleksandra Faust, Edgar Duenez-Guzman, and Mohammad Ghavamzadeh. 2019 · 2019
Cited alongside, same era.
When to Trust Your Model: Model-Based Policy Optimization. In Advances in Neural Information Processing Systems 32 (NeurIPS) , H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alché-Buc, E. Fox, and R. Garnett (Eds.). 12519–12530
Michael Janner, Justin Fu, Marvin Zhang, and Sergey Levine. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Learning Barrier Certificates: Towards Safe Reinforcement Learning with Zero Training-time Violations. In Advances in Neural Information Processing Systems 34 (NeurIPS) , M. Ranzato, A. Beygelzimer, Y. Dauphin, P.S. Liang, and J. Wortman Vaughan (Eds.). 25621–25632
Yuping Luo and Tengyu Ma. 2021 · 2021
Later among the works it cites.
Lyapunov Barrier Policy Optimization
Harshit Sikchi, Wenxuan Zhou, and David Held. 2021 · 2021
Later among the works it cites.
Recovery RL: Safe Reinforcement Learning With Learned Recovery Zones
Brijen Thananjeyan, Ashwin Balakrishna, Suraj Nair, Michael Luo, Krishnan Srinivasan, Minho Hwang, Joseph E. Gonzalez, Julian Ibarz, Chelsea Finn, and Ken Goldberg. 2021 · 2021
Later among the works it cites.
Probabilistic Model Predictive Safety Certification for Learning-Based Control
Kim P. Wabersich, Lukas Hewing, Andrea Carron, and Melanie N. Zeilinger. 2022 · 2021
Later among the works it cites.
A predictive safety filter for learning-based control of constrained nonlinear dynamical systems
Kim Peter Wabersich and Melanie N. Zeilinger. 2021 · 2021
Later among the works it cites.
Safe Learning in Robotics: From Learning-Based Control to Safe Reinforcement Learning
Lukas Brunke, Melissa Greeff, Adam W. Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P. Schoellig. 2022 · 2022
Later among the works it cites.
Bullet-Safety-Gym: A Framework for Constrained Reinforcement Learning
Sven Gronauer. 2022 · 2022
Later among the works it cites.
Safe Reinforcement Learning with Chance-constrained Model Predictive Control. In Proceedings of The 4th Annual Learning for Dynamics and Control Conference , Vol. 168. PMLR, 291–303
Samuel Pfrommer, Tanmay Gautam, Alec Zhou, and Somayeh Sojoudi. 2022 · 2022
Later among the works it cites.