Fetching the paper…
Reading the bibliography…
Safety constraints and optimality are important, but sometimes conflicting criteria for controllers.
G. E. Box, “ Science and statistics ,” Journal of the American Statistical Association , 1976
1976
Earlier work this paper cites.
M. Bardi and I. C. Dolcetta, Optimal control and viscosity solutions of Hamilton-Jacobi-Bellman equations . Springer, 1997
1997
Earlier work this paper cites.
E. Altman, Constrained Markov Decision Processes . Routledge, 1999
1999
Earlier work this paper cites.
P. Geibel, “ Reinforcement Learning with Bounded Risk ,” in International Conference on Machine Learning , 2001
2001
Earlier work this paper cites.
S. Boyd and L. Vandenberghe, Convex optimization . Cambridge university press, 2004
2004
Earlier work this paper cites.
P. Geibel and F. Wysotzki, “ Risk-Sensitive Reinforcement Learning Applied to Control under Constraints ,” in Journal of Artificial Intelligence Research , 2005, pp. 81–108
2005
Earlier work this paper cites.
D. P. Bertsekas, Dynamic Programming and Optimal Control , 3rd ed. Athena Scientific Belmont, 2005, vol. 1
2005
Earlier work this paper cites.
W. H. Fleming and H. M. Soner, Controlled Markov processes and viscosity solutions . Springer, 2006
2006
Earlier work this paper cites.
J.-P. Aubin, A. M. Bayen, and P. Saint-Pierre, Viability theory: new directions , 2nd ed. Springer Science & Business Media, 2011
2011
Earlier work this paper cites.
I. R. Manchester, M. M. Tobenkin, M. Levashov, and R. Tedrake, “ Regions of attraction for hybrid limit cycles of walking robots ,” IFAC World Congress , vol. 44, 2011
2011
Earlier work this paper cites.
T. M. Moldovan and P. Abbeel, “ Safe exploration in markov decision processes ,” in Proceedings of the 29th International Conference on Machine Learning , 2012, pp. 1451–1458
2012
Earlier work this paper cites.
T. Koolen, T. De Boer, J. Rebula, A. Goswami, and J. Pratt, “ Capturability-based analysis and control of legged locomotion, part 1: Theory and application to three simple gait models ,” The international journal of robotics research , 2012
2012
Earlier work this paper cites.
J. Schulman, Y. Duan, J. Ho, A. Lee, I. Awwal, H. Bradlow, J. Pan, S. Patil, K. Goldberg, and P. Abbeel, “ Motion planning with sequential convex optimization and convex collision checking ,” The International Journal of Robotics Research , 2014
2014
Earlier work this paper cites.
A. Boccia, L. Grüne, and K. Worthmann, “ Stability and feasibility of state constrained MPC without stabilizing terminal constraints ,” Systems & control letters , 2014
2014
Earlier work this paper cites.
A. K. Akametalu, J. F. Fisac, J. H. Gillula, S. Kaynama, M. N. Zeilinger, and C. J. Tomlin, “ Reachability-based safe learning with Gaussian processes ,” in IEEE Conference on Decision and Control , 2014, pp. 1424–1431
2014
Earlier work this paper cites.
J. Garcıa and F. Fernández, “ A comprehensive survey on safe reinforcement learning ,” Journal of Machine Learning Research , vol. 16, 2015
2015
Earlier work this paper cites.
P. Zaytsev, S. J. Hasaneini, and A. Ruina, “ Two steps is enough: No need to plan far ahead for walking balance ,” in IEEE International Conference on Robotics and Automation , 2015
2015
Earlier work this paper cites.
R. Postoyan, L. Buşoniu, D. Nešić, and J. Daafouz, “ Stability analysis of discrete-time infinite-horizon optimal control with discounted cost ,” IEEE Transactions on Automatic Control , 2016
2016
Earlier work this paper cites.
L. Grüne and J. Pannek, Nonlinear Model Predictive Control , 2nd ed. Springer, 2017
2017
Earlier work this paper cites.
S. Bansal, M. Chen, S. Herbert, and C. J. Tomlin, “ Hamilton-jacobi reachability: A brief overview and recent advances ,” in IEEE Conference on Decision and Control , 2017
2017
Earlier work this paper cites.
J. Achiam, D. Held, A. Tamar, and P. Abbeel, “ Constrained policy optimization ,” in International Conference on Machine Learning . PMLR, 2017, pp. 22–31
2017
Earlier work this paper cites.
M. A. Posa, T. Koolen, and R. Tedrake, “ Balancing and step recovery capturability via sums-of-squares optimization ,” Robotics: Science and Systems , 2017
2017
Cited alongside, same era.
J. F. Fisac, A. K. Akametalu, M. N. Zeilinger, S. Kaynama, J. Gillula, and C. J. Tomlin, “ A general safety framework for learning-based control in uncertain robotic systems ,” IEEE Transactions on Automatic Control , vol. 64, 2018
2018
Cited alongside, same era.
M. Chen and C. J. Tomlin, “ Hamilton–Jacobi Reachability: Some Recent Theoretical Advances and Applications in Unmanned Airspace Management ,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 1, pp. 333–358, 2018
2018
Cited alongside, same era.
A. Rai, R. Antonova, S. Song, W. Martin, H. Geyer, and C. Atkeson, “ Bayesian optimization using domain knowledge on the atrias biped ,” in IEEE International Conference on Robotics and Automation , 2018
2018
Cited alongside, same era.
J. Nubert, J. Köhler, V. Berenz, F. Allgöwer, and S. Trimpe, “ Safe and fast tracking on a robot manipulator: Robust mpc and neural network control ,” IEEE Robotics and Automation Letters , 2020
2020
Later among the works it cites.
N. M. Boffi, S. Tu, N. Matni, J.-J. E. Slotine, and V. Sindhwani, “ Learning stability certificates from data ,” in Conference on Robot Learning , 2020, pp. 5914–5920
2020
Later among the works it cites.
A. Robey, H. Hu, L. Lindemann, H. Zhang, D. V. Dimarogonas, S. Tu, and N. Matni, “ Learning control barrier functions from expert demonstrations ,” in IEEE Conference on Decision and Control , 2020
2020
Later among the works it cites.
M. El-Shamouty, X. Wu, S. Yang, M. Albus, and M. F. Huber, “ Towards safe human-robot collaboration using deep reinforcement learning ,” in IEEE International Conference on Robotics and Automation , 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction , 2nd ed. MIT press, 2018
2018
Cited alongside, same era.
P. Zaytsev, W. Wolfslag, and A. Ruina, “ The boundaries of walking stability: Viability and controllability of simple models ,” IEEE Transactions on Robotics , 2018
2018
Cited alongside, same era.
J. S. Matthis, J. L. Yates, and M. M. Hayhoe, “ Gaze and the control of foot placement when walking in natural terrain ,” Current Biology , 2018
2018
Cited alongside, same era.
A. D. Ames, S. Coogan, M. Egerstedt, G. Notomista, K. Sreenath, and P. Tabuada, “ Control barrier functions: Theory and applications ,” in IEEE European Control Conference , 2019
2019
Cited alongside, same era.
S. Heim, A. Rohr, S. Trimpe, and A. Badri-Spröwitz, “ A learnable safety measure ,” in Conference on Robot Learning , 2019
2019
Cited alongside, same era.
J. F. Fisac, N. F. Lugovoy, V. Rubies-Royo, S. Ghosh, and C. J. Tomlin, “ Bridging hamilton-jacobi safety analysis and reinforcement learning ,” in 2019 International Conference on Robotics and Automation , 2019
2019
Cited alongside, same era.
M. Turchetta, F. Berkenkamp, and A. Krause, “ Safe exploration for interactive machine learning ,” in Advances in Neural Information Processing Systems , 2019, pp. 905–916
2019
Cited alongside, same era.
M. H. Yeganegi, M. Khadiv, S. A. A. Moosavian, J.-J. Zhu, A. Del Prete, and L. Righetti, “ Robust humanoid locomotion using trajectory optimization and sample-efficient learning ,” in IEEE International Conference on Humanoid Robots , 2019, pp. 170–177
2019
Cited alongside, same era.
L. Zheng and L. J. Ratliff, “ Constrained Upper Confidence Reinforcement Learning with Known Dynamics ,” IEEE International Conference on Robotics and Automation , pp. 620–629, 2020
2020
Later among the works it cites.
M. Korda, “ Computing controlled invariant sets from data using convex optimization ,” SIAM Journal on Control and Optimization , vol. 58, 2020
2020
Later among the works it cites.
M. Khadiv, A. Herzog, S. A. A. Moosavian, and L. Righetti, “ Walking control based on step timing adaptation ,” IEEE Transactions on Robotics , vol. 36, 2020
2020
Later among the works it cites.
M. V. Srinivasan, S.-W. Zhang, J. S. Chahl, E. Barth, and S. Venkatesh, “ How honeybees make grazing landings on flat surfaces ,” Biological cybernetics , 2000
2020
Later among the works it cites.
J. J. Choi, D. Lee, K. Sreenath, C. J. Tomlin, and S. L. Herbert, “ Robust control barrier-value functions for safety-critical control ,” in IEEE Conference on Decision and Control , 2021
2021
Closest in time.
S. Herbert, J. J. Choi, S. Sanjeev, M. Gibson, K. Sreenath, and C. J. Tomlin, “ Scalable learning of safety guarantees for autonomous systems using hamilton-jacobi reachability ,” in IEEE International Conference on Robotics and Automation , 2021
2021
Closest in time.
G. Chou, D. Berenson, and N. Ozay, “ Learning constraints from demonstrations with grid and parametric representations ,” The International Journal of Robotics Research , pp. 627–639, 2021
2021
Closest in time.
P.-F. Massiani, S. Heim, and S. Trimpe, “ On exploration requirements for learning safety constraints ,” in Learning for Dynamics and Control , 2021, pp. 8550–8556
2021
Closest in time.
A. Marco, D. Baumann, M. Khadiv, P. Hennig, L. Righetti, and S. Trimpe, “ Robot learning with crash constraints ,” IEEE Robotics and Automation Letters , vol. 6, pp. 1439–1446, 2021
2021
Closest in time.
A. Meduri, M. Khadiv, and L. Righetti, “ Deepq stepper: A framework for reactive dynamic walking on uneven terrain ,” in IEEE International Conference on Robotics and Automation , 2021
2021
Closest in time.
N. Funk, D. Baumann, V. Berenz, and S. Trimpe, “ Learning event-triggered control from data through joint optimization ,” IFAC Journal of Systems and Control , vol. 16, 2021
2021
Closest in time.
J. Li, D. Fridovich-Keil, S. Sojoudi, and C. J. Tomlin, “ Augmented lagrangian method for instantaneously constrained reinforcement learning problems ,” in IEEE Conference on Decision and Control , 2021
2021
Closest in time.
J. R. Martins and A. Ning, Engineering design optimization . Cambridge University Press, 2021
2021
Closest in time.
K.-C. Hsu, V. Rubies-Royo, C. Tomlin, and J. Fisac, “ Safety and Liveness Guarantees through Reach-Avoid Reinforcement Learning ,” in Robotics: Science and Systems , 2021
2021
Closest in time.
L. Brunke, M. Greeff, A. W. Hall, Z. Yuan, S. Zhou, J. Panerati, and A. P. Schoellig, “ Safe learning in robotics: From learning-based control to safe reinforcement learning ,” Annual Review of Control, Robotics, and Autonomous Systems , 2022
2022
Closest in time.
S. Paternain, M. Calvo-Fullana, L. F. Chamon, and A. Ribeiro, “ Safe policies for reinforcement learning via primal-dual methods ,” IEEE Transactions on Automatic Control , 2022
2022
Closest in time.