Fetching the paper…
Reading the bibliography…
Safety stands as the primary obstacle preventing the widespread adoption of learning-based robotic systems in our daily lives.
Constructive safety using control barrier functions
P. Wieland · 2007
Earlier work this paper cites.
A comprehensive survey on safe reinforcement learning
J. García and F. Fernández · 2015
Earlier work this paper cites.
Constrained policy optimization
J. Achiam, D. Held, A. Tamar, and P. Abbeel · 2017
Earlier work this paper cites.
Control barrier function based quadratic programs for safety critical systems
A. D. Ames, X. Xu, J. W. Grizzle, and P. Tabuada · 2017
Earlier work this paper cites.
Safe reinforcement learning via shielding
M. Alshiekh, R. Bloem, R. Ehlers, B. Könighofer, S. Niekum, and U. Topcu · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Earlier work this paper cites.
Constrained exploration and recovery from experience shaping
T.-H. Pham, G. De Magistris, D. J. Agravante, S. Chaudhury, A. Munawar, and R. Tachibana · 2018
Earlier work this paper cites.
Reinforcement Learning: An Introduction
R. S. Sutton and A. G. Barto · 2018
Earlier work this paper cites.
Reward constrained policy optimization, 2018
C. Tessler, D. J. Mankowitz, and S. Mannor · 2018
Earlier work this paper cites.
Control barrier functions: Theory and applications, 2019
A. D. Ames, S. Coogan, M. Egerstedt, G. Notomista, K. Sreenath, and P. Tabuada · 2019
Earlier work this paper cites.
End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks
R. Cheng, G. Orosz, R. M. Murray, and J. W. Burdick · 2019
Earlier work this paper cites.
Reinforcement learning with convex constraints, 2019
S. Miryoosefi, K. Brantley, H. D. I. au2, M. Dudik, and R. Schapire · 2019
Cited alongside, same era.
Learning for safety-critical control with control barrier functions, 2019
A. Taylor, A. Singletary, Y. Yue, and A. Ames · 2019
Cited alongside, same era.
Exploration-exploitation in constrained mdps, 2020
Y. Efroni, S. Mannor, and M. Pirotta · 2020
Cited alongside, same era.
Robust model predictive shielding for safe reinforcement learning with stochastic dynamics
S. Li and O. Bastani · 2020
Cited alongside, same era.
Safe learning for control using control lyapunov functions and control barrier functions: A review
A. Anand, K. Seel, V. Gjærum, A. Håkansson, H. Robinson, and A. Saad · 2021
Cited alongside, same era.
A simple reward-free approach to constrained reinforcement learning, 2021
Champion-level drone racing using deep reinforcement learning
E. Kaufmann, L. Bauersfeld, A. Loquercio, M. Müller, V. Koltun, and D. Scaramuzza · 2023
Later among the works it cites.
Intentionally-underestimated value function at terminal state for temporal-difference learning with mis-designed reward, 2023
T. Kobayashi · 2023
Later among the works it cites.
Can a bayesian oracle prevent harm from an agent?, 2024
Y. Bengio, M. K. Cohen, N. Malkin, M. MacDermott, D. Fornasiere, P. Greiner, and Y. Kaddar · 2024
Later among the works it cites.
International scientific report on the safety of advanced ai (interim report), 2024
Y. Bengio, S. Mindermann, D. Privitera, T. Besiroglu, R. Bommasani, S. Casper, Y. Choi, D. Goldfarb, H. Heidari, L. Khalatbari, S. Longpre, V. Mavroudis, M. Mazeika, K. Y. Ng, C. T. Okolo, D. Raji, T. Skeadas, F. Tramèr, B. Adekanmbi, P. Christiano, D. Dalrymple, T. G. Dietterich, E. Felten, P. Fung, P.-O. Gourinchas, N. Jennings, A. Krause, P. Liang, T. Ludermir, V. Marda, H. Margetts, J. A. McDermid, A. Narayanan, A. Nelson, A. Oh, G. Ramchurn, S. Russell, M. Schaake, D. Song, A. Soto, L. Tiedrich, G. Varoquaux, A. Yao, and Y.-Q. Zhang · 2024
Later among the works it cites.
Rl, but don’t do anything i wouldn’t do, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Miryoosefi and C. Jin · 2021
Cited alongside, same era.
Reward is enough
D. Silver, S. Singh, D. Precup, and R. S. Sutton · 2021
Cited alongside, same era.
A predictive safety filter for learning-based control of constrained nonlinear dynamical systems, 2021
K. P. Wabersich and M. N. Zeilinger · 2021
Cited alongside, same era.
Safe reinforcement learning using robust control barrier functions, 2022
Y. Emam, G. Notomista, P. Glotfelter, Z. Kira, and M. Egerstedt · 2022
Cited alongside, same era.
Sim-to-lab-to-real: Safe reinforcement learning with shielding and generalization guarantees
K. Hsu, A. Z. Ren, D. P. Nguyen, A. Majumdar, and J. F. Fisac · 2022
Cited alongside, same era.
M. K. Cohen, M. Hutter, Y. Bengio, and S. Russell · 2024
Later among the works it cites.
Building guardrails for large language models, 2024
Y. Dong, R. Mu, G. Jin, Y. Qi, J. Hu, X. Zhao, J. Meng, W. Ruan, and X. Huang · 2024
Later among the works it cites.
A review of safe reinforcement learning: Methods, theory and applications, 2024
S. Gu, L. Yang, Y. Du, G. Chen, F. Walter, J. Wang, and A. Knoll · 2024
Later among the works it cites.
Learning control barrier functions and their application in reinforcement learning: A survey, 2024
M. Guerrier, H. Fouad, and G. Beltrame · 2024
Later among the works it cites.
Gymnasium: A standard interface for reinforcement learning environments
M. Towers, A. Kwiatkowski, J. Terry, J. U. Balis, G. De Cola, T. Deleu, M. Goulão, A. Kallinteris, M. Krimmel, A. KG, et al · 2024
Later among the works it cites.
Open problems in machine unlearning for ai safety, 2025
F. Barez, T. Fu, A. Prabhu, S. Casper, A. Sanyal, A. Bibi, A. O’Gara, R. Kirk, B. Bucknall, T. Fist, L. Ong, P. Torr, K.-Y. Lam, R. Trager, D. Krueger, S. Mindermann, J. Hernandez-Orallo, M. Geva, and Y. Gal · 2025
Closest in time.