Fetching the paper…
Reading the bibliography…
Reinforcement learning (RL) controllers are flexible and performant but rarely guarantee safety.
A. Y. Ng, D. Harada, and S. Russell, “Policy invariance under reward transformations: Theory and application to reward shaping,” in International Conference on Machine Learning , 1999
1999
Earlier work this paper cites.
F. Blanchini, “Set invariance in control,” Automatica , 1999
1999
Earlier work this paper cites.
D. Silver et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , 2016
2016
Earlier work this paper cites.
C. E. Luis and J. L. Ny, “Design of a trajectory tracking controller for a nanoquadcopter,” Technical Report, École Polytechnique de Montréal, 2016
2016
Earlier work this paper cites.
J. Achiam, D. Held, A. Tamar, and P. Abbeel, “Constrained policy optimization,” in International conference on machine learning , 2017
2017
Earlier work this paper cites.
J. B. Rawlings, D. Q. Mayne, and M. Diehl, Model predictive control: theory, computation, and design , 2nd ed. Madison, Wisconsin: Nob Hill Publishing, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. A. Shvets, A. Rakhlin, A. A. Kalinin, and V. I. Iglovikov, “Automatic instrument segmentation in robot-assisted surgery using deep learning,” in IEEE International Conference on Machine Learning and Applications , 2018
2018
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction , 2nd ed. The MIT Press, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
T.-H. Pham, G. De Magistris, and R. Tachibana, “Optlayer - practical constrained optimization for deep reinforcement learning in the real world,” in 2018 IEEE International Conference on Robotics and Automation , 2018
2018
Earlier work this paper cites.
Z. Li, U. Kalabić, and T. Chu, “Safe reinforcement learning: Learning with supervision using a constraint-admissible set,” in American Control Conference , 2018
2018
Earlier work this paper cites.
A. D. Ames, S. Coogan, M. Egerstedt, G. Notomista, K. Sreenath, and P. Tabuada, “Control barrier functions: Theory and applications,” in European Control Conference , 2019
2019
Cited alongside, same era.
R. Cheng, G. Orosz, R. M. Murray, and J. W. Burdick, “End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks,” in Proceedings of the AAAI Conference on Artificial Intelligence , 2019
2019
Cited alongside, same era.
K. Burnett et al. , “Zeus: A system description of the two-time winner of the collegiate SAE AutoDrive competition,” Journal of Field Robotics , 2021
2021
Cited alongside, same era.
K. P. Wabersich and M. N. Zeilinger, “A predictive safety filter for learning-based control of constrained nonlinear dynamical systems,” Automatica , 2021
2021
Cited alongside, same era.
E. Altman, Constrained Markov decision processes . Routledge, 2021
K.-C. Hsu, H. Hu, and J. F. Fisac, “The safety filter: A unified view of safety-critical control in autonomous systems,” Annual Review of Control, Robotics, and Autonomous Systems , 2023
2023
Later among the works it cites.
K. P. Wabersich, A. J. Taylor, J. J. Choi, K. Sreenath, C. J. Tomlin, A. D. Ames, and M. N. Zeilinger, “Data-driven safety filters: Hamilton-jacobi reachability, control barrier functions, and predictive methods for uncertain systems,” IEEE Control Systems Magazine , 2023
2023
Later among the works it cites.
H. Krasowski, J. Thumm, M. Müller, L. Schäfer, X. Wang, and M. Althoff, “Provably safe reinforcement learning: Conceptual analysis, survey, and benchmarking,” Transactions on Machine Learning Research , 2023
2023
Later among the works it cites.
F. Pizarro Bejarano, L. Brunke, and A. P. Schoellig, “Multi-step model predictive safety filters: Reducing chattering by increasing the prediction horizon,” in IEEE Conference on Decision and Control , 2023
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
2021
Cited alongside, same era.
J. Köhler, R. Soloperto, M. A. Müller, and F. Allgöwer, “A computationally efficient robust model predictive control framework for uncertain nonlinear systems – extended version,” IEEE Transactions on Automatic Control , 2021
2021
Cited alongside, same era.
R. Verschueren, G. Frison, D. Kouzoupis, J. Frey, N. van Duijkeren, A. Zanelli, B. Novoselnik, T. Albin, R. Quirynen, and M. Diehl, “acados: a modular open-source framework for fast embedded optimal control,” Mathematical Programming Computation , 2021
2021
Cited alongside, same era.
L. Brunke, M. Greeff, A. W. Hall, Z. Yuan, S. Zhou, J. Panerati, and A. P. Schoellig, “Safe learning in robotics: From learning-based control to safe reinforcement learning,” Annual Review of Control, Robotics, and Autonomous Systems , 2022
2022
Cited alongside, same era.
Z. Yuan, A. W. Hall, S. Zhou, L. Brunke, M. Greeff, J. Panerati, and A. P. Schoellig, “Safe-control-gym: A unified benchmark suite for safe learning-based control and reinforcement learning in robotics,” IEEE Robotics and Automation Letters , 2022
2022
Cited alongside, same era.
X. Wang, “Ensuring safety of learning-based motion planners using control barrier functions,” IEEE Robotics and Automation Letters , 2022
2022
Cited alongside, same era.
S. Pfrommer, T. Gautam, A. Zhou, and S. Sojoudi, “Safe reinforcement learning with chance-constrained model predictive control,” in Learning for Dynamics and Control Conference , 2022
2022
Cited alongside, same era.
Later among the works it cites.
N. Kochdumper, H. Krasowski, X. Wang, S. Bak, and M. Althoff, “Provably safe reinforcement learning via action projection using reachability analysis and polynomial zonotopes,” IEEE Open Journal of Control Systems , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
A. Norouzi, S. Shahpouri, D. Gordon, M. Shahbakhti, and C. R. Koch, “Safe deep reinforcement learning in diesel engine emission control,” Journal of Systems and Control Engineering , 2023
2023
Later among the works it cites.
K. Dunlap, M. Mote, K. Delsing, and K. L. Hobbs, “Run time assured reinforcement learning for safe satellite docking,” Journal of Aerospace Information Systems , 2023
2023
Later among the works it cites.
W. Xiao, T.-H. Wang, R. Hasani, M. Chahine, A. Amini, X. Li, and D. Rus, “BarrierNet: Differentiable control barrier functions for learning of safe robot control,” IEEE Transactions on Robotics , 2023
2023
Later among the works it cites.
S. Yang, S. Chen, V. M. Preciado, and R. Mangharam, “Differentiable safe controller design through control barrier functions,” IEEE Control Systems Letters , 2023
2023
Later among the works it cites.
X. Wang and M. Althoff, “Safe reinforcement learning for automated vehicles via online reachability analysis,” IEEE Transactions on Intelligent Vehicles , 2023
2023
Later among the works it cites.