Fetching the paper…
Reading the bibliography…
Safety is a crucial property of every robotic platform: any control policy should always comply with actuator limits and avoid collisions with the environment and humans.
R. P. Singh and P. W. Likins, “Singular Value Decomposition for Constrained Dynamical Systems,” Journal of Applied Mechanics, Transactions ASME , vol. 52, no. 4, pp. 943–948, 1985
1985
Earlier work this paper cites.
S. S. Kim and M. J. Vanderploeg, “QR Decomposition for State Space Representation of Constrained Mechanical Dynamic Systems,” Journal of Mechanisms, Transmissions, and Automation in Design , vol. 108, pp. 183–188, 1986
1986
Earlier work this paper cites.
R. H. Byrd and R. B. Schnabel, “Continuity of the Null Space Basis and Constrained Optimization,” Mathematical Programming , vol. 35, no. 1, pp. 32–41, 1986
1986
Earlier work this paper cites.
E. Altman, “Constrained Markov Decision Processes with Total Cost Criteria: Lagrangian Approach and Dual Linear Program,” Mathematical methods of operations research , vol. 48, no. 3, pp. 387–417, 1998
1998
Earlier work this paper cites.
E. Altman, Constrained Markov decision processes: stochastic modeling . Routledge, 1999
1999
Earlier work this paper cites.
A. Hans, D. Schneegaß, A. M. Schäfer, and S. Udluft, “Safe Exploration for Reinforcement Learning,” in European Symposium on Artificial Neural Networks (ESANN) , 2008
2008
Earlier work this paper cites.
T. M. Moldovan and P. Abbeel, “Safe exploration in markov decision processes,” in International Conference on Machine Learning (ICML) , 2012, pp. 1711–1718
2012
Earlier work this paper cites.
J. Garcia and F. Fernandez, “Safe Exploration of State and Action Spaces in Reinforcement Learning,” Journal of Artificial Intelligence Research , vol. 45, pp. 515–564, 2012
2012
Earlier work this paper cites.
A. K. Akametalu, J. F. Fisac, J. H. Gillula, S. Kaynama, M. N. Zeilinger, and C. J. Tomlin, “Reachability-based safe learning with gaussian processes,” in 53rd IEEE Conference on Decision and Control . IEEE, 2014, pp. 1424–1431
2014
Earlier work this paper cites.
M. Pecka, K. Zimmermann, and T. Svoboda, “Safe exploration for reinforcement learning in real unstructured environments,” in Proc. of the Computer Vision Winter Workshop , 2015
2015
Earlier work this paper cites.
H. B. Ammar, R. Tutunov, and E. Eaton, “Safe policy search for lifelong reinforcement learning with sublinear regret,” in International Conference on Machine Learning . PMLR, 2015
2015
Earlier work this paper cites.
D. Martínez, G. Alenya, and C. Torras, “Safe robot execution in model-based reinforcement learning,” in 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2015, pp. 6422–6427
2015
Earlier work this paper cites.
M. Loper, N. Mahmood, J. Romero, G. Pons-Moll, and M. J. Black, “Smpl: A skinned multi-person linear model,” ACM TOG , 2015
2015
Earlier work this paper cites.
F. Berkenkamp, M. Turchetta, A. P. Schoellig, and A. Krause, “Safe Model-based Reinforcement Learning with Stability Guarantees,” in Conference on Neural Information Processing Systems (NIPS) , 2017
2017
Earlier work this paper cites.
J. Achiam, D. Held, A. Tamar, and P. Abbeel, “Constrained Policy Optimization,” in International Conference on Machine Learning (ICML) , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
K. M. Lynch and F. C. Park, Modern robotics . Cambridge University Press, 2017
2017
Earlier work this paper cites.
M. Alshiekh, R. Bloem, R. Ehlers, B. Könighofer, S. Niekum, and U. Topcu, “Safe Reinforcement Learning via Shielding,” in AAAI Conference on Artificial Intelligence (AAAI) , 2018
2018
Earlier work this paper cites.
A. Wachi, Y. Sui, Y. Yue, and M. Ono, “Safe Exploration and Optimization of Constrained MDPs Using Gaussian Processes,” in AAAI Conference on Artificial Intelligence (AAAI) , 2018
2018
Earlier work this paper cites.
T. Koller, F. Berkenkamp, M. Turchetta, and A. Krause, “Learning-Based Model Predictive Control for Safe Exploration,” in IEEE Conference on Decision and Control , 2018
2018
Earlier work this paper cites.
Y. Chow, O. Nachum, E. Duenez-Guzman, and M. Ghavamzadeh, “A Lyapunov-based Approach to Safe Reinforcement Learning,” in Conference on Neural Information Processing Systems (NIPS) , 2018
2018
Cited alongside, same era.
J. F. Fisac, A. K. Akametalu, M. N. Zeilinger, S. Kaynama, J. Gillula, and C. J. Tomlin, “A general safety framework for learning-based control in uncertain robotic systems,” IEEE Transactions on Automatic Control , vol. 64, no. 7, pp. 2737–2752, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
T.-H. Pham, G. De Magistris, and R. Tachibana, “Optlayer-practical constrained optimization for deep reinforcement learning in the real world,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 6236–6243
2018
Cited alongside, same era.
J. Ibarz, J. Tan, C. Finn, M. Kalakrishnan, P. Pastor, and S. Levine, “How to train your robot with deep reinforcement learning: lessons we have learned,” The International Journal of Robotics Research , vol. 40, no. 4-5, pp. 698–721, 2021
2021
Later among the works it cites.
D. Ding, X. Wei, Z. Yang, Z. Wang, and M. R. Jovanovic, “Provably Efficient Safe Exploration via Primal-Dual Policy Optimization,” in International Conference on Artificial Intelligence and Statistics (AISTATS) , vol. 130, 2021
2021
Later among the works it cites.
Y. S. Shao, C. Chen, S. Kousik, and R. Vasudevan, “Reachability-based trajectory safeguard (rts): A safe and fast reinforcement learning safety layer for continuous control,” IEEE Robotics and Automation Letters , vol. 6, no. 2, pp. 3663–3670, 2021
2021
Later among the works it cites.
A. Marco, D. Baumann, M. Khadiv, P. Hennig, L. Righetti, and S. Trimpe, “Robot learning with crash constraints,” IEEE Robotics and Automation Letters , vol. 6, no. 2, pp. 1439–1446, 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
C. Tessler, D. J. Mankowitz, and S. Mannor, “Reward Constrained Policy Optimization,” in International Conference on Learning Representations (ICLR) , 2019
2019
Cited alongside, same era.
F. Berkenkamp, “Safe exploration in reinforcement learning: Theory and applications in robotics,” Ph.D. dissertation, ETH Zurich, 2019
2019
Cited alongside, same era.
Y. Chow, O. Nachum, A. Faust, E. Duenez-Guzman, and M. Ghavamzadeh, “Lyapunov-based Safe Policy Optimization for Continuous Control,” in Reinforcement Learning for Real Life (RL4RealLife) Workshop in the 36 th International Conference on Machine Learning , 2019
2019
Cited alongside, same era.
R. Cheng, G. Orosz, R. M. Murray, and J. W. Burdick, “End-to-End Safe Reinforcement Learning through Barrier Functions for Safety-Critical Continuous Control Tasks,” in AAAI Conference on Artificial Intelligence . AAAI Press, 2019, pp. 3387–3395
2019
Cited alongside, same era.
A. Bajcsy, S. Bansal, E. Bronstein, V. Tolani, and C. J. Tomlin, “An efficient reachability-based framework for provably safe autonomous navigation in unknown environments,” in 2019 IEEE 58th Conference on Decision and Control (CDC) . IEEE, 2019, pp. 1758–1765
2019
Cited alongside, same era.
B. Lütjens, M. Everett, and J. P. How, “Safe reinforcement learning with model uncertainty estimates,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 8662–8668
2019
Cited alongside, same era.
G. Chalvatzaki, X. S. Papageorgiou, P. Maragos, and C. S. Tzafestas, “Learn to adapt to human walking: A model-based reinforcement learning approach for a robotic assistant rollator,” IEEE Robotics and Automation Letters , vol. 4, no. 4, pp. 3774–3781, 2019
2019
Cited alongside, same era.
2021
Later among the works it cites.
Y. L. Pang, A. Xompero, C. Oh, and A. Cavallaro, “Towards safe human-to-robot handovers of unknown containers,” in 2021 30th IEEE International Conference on Robot & Human Interactive Communication (RO-MAN) . IEEE, 2021, pp. 51–58
2021
Later among the works it cites.
B. Thananjeyan, A. Balakrishna, S. Nair, M. Luo, K. Srinivasan, M. Hwang, J. E. Gonzalez, J. Ibarz, C. Finn, and K. Goldberg, “Recovery rl: Safe reinforcement learning with learned recovery zones,” IEEE Robotics and Automation Letters , vol. 6, no. 3, pp. 4915–4922, 2021
2021
Later among the works it cites.
C. D’Eramo, D. Tateo, A. Bonarini, M. Restelli, and J. Peters, “Mushroomrl: Simplifying reinforcement learning research,” Journal of Machine Learning Research , vol. 22, no. 131, pp. 1–5, 2021
2021
Later among the works it cites.
L. Brunke, M. Greeff, A. W. Hall, Z. Yuan, S. Zhou, J. Panerati, and A. P. Schoellig, “Safe learning in robotics: From learning-based control to safe reinforcement learning,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 5, pp. 411–444, 2022
2022
Closest in time.
A. I. Cowen-Rivers, D. Palenicek, V. Moens, M. A. Abdullah, A. Sootla, J. Wang, and H. Bou-Ammar, “Samba: Safe model-based & active reinforcement learning,” Machine Learning , pp. 1–31, 2022
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
P. Liu, D. Tateo, H. B. Ammar, and J. Peters, “Robot reinforcement learning on the constraint manifold,” in Conference on Robot Learning . PMLR, 2022, pp. 1357–1366
2022
Closest in time.
2022
Closest in time.
K. Weerakoon, A. J. Sathyamoorthy, U. Patel, and D. Manocha, “Terp: Reliable planning in uneven outdoor environments using deep reinforcement learning,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 9447–9453
2022
Closest in time.
2022
Closest in time.
J. Thumm and M. Althoff, “Provably safe deep reinforcement learning for robotic manipulation in human environments,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 6344–6350
2022
Closest in time.
S. R. Schepp, J. Thumm, S. B. Liu, and M. Althoff, “Sara: A tool for safe human-robot coexistence and collaboration through reachability analysis,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 4312–4317
2022
Closest in time.
2022
Closest in time.
R. Kaushik, K. Arndt, and V. Kyrki, “Safeapt: Safe simulation-to-real robot learning using diverse policies learned in simulation,” IEEE Robotics and Automation Letters , 2022
2022
Closest in time.