Fetching the paper…
Reading the bibliography…
Reinforcement Learning (RL) has been shown to be effective in many scenarios.
E. Altman, “Constrained markov decision processes with total cost criteria: Lagrangian approach and dual linear program,” Mathematical methods of operations research , vol. 48, no. 3, pp. 387–417, 1998
1998
Earlier work this paper cites.
E. Altman, Constrained Markov decision processes . CRC Press, 1999, vol. 7
1999
Earlier work this paper cites.
C. E. Rasmussen, “Gaussian processes in machine learning,” in Summer School on Machine Learning . Springer, 2003, pp. 63–71
2003
Earlier work this paper cites.
A. Zavala-Rio, I. Fantoni, and R. Lozano, “Global stabilization of a pvtol aircraft model with bounded inputs,” International Journal of Control , vol. 76, no. 18, pp. 1833–1844, 2003
2003
Earlier work this paper cites.
P. Geibel and F. Wysotzki, “Risk-sensitive reinforcement learning applied to control under constraints,” Journal of Artificial Intelligence Research , vol. 24, pp. 81–108, 2005
2005
Earlier work this paper cites.
P. Ogren, A. Backlund, T. Harryson, L. Kristensson, and P. Stensson, “Autonomous ucav strike missions using behavior control lyapunov functions,” in AIAA Guidance, Navigation, and Control Conference and Exhibit , 2006, p. 6197
2006
Earlier work this paper cites.
J. Cortes, “Discontinuous dynamical systems,” IEEE Control Systems Magazine , vol. 28, no. 3, pp. 36–73, 2008
2008
Earlier work this paper cites.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement learning in robotics: A survey,” The International Journal of Robotics Research , vol. 32, no. 11, pp. 1238–1274, 2013
2013
Earlier work this paper cites.
A. D. Ames, J. W. Grizzle, and P. Tabuada, “Control barrier function based quadratic programs with application to adaptive cruise control,” in 53rd IEEE Conference on Decision and Control , 12 2014, pp. 6271–6278
2014
Earlier work this paper cites.
J. Garcıa and F. Fernández, “A comprehensive survey on safe reinforcement learning,” Journal of Machine Learning Research , vol. 16, no. 1, pp. 1437–1480, 2015
2015
Earlier work this paper cites.
X. Xu, P. Tabuada, J. W. Grizzle, and A. D. Ames, “Robustness of control barrier functions for safety critical control,” IFAC-PapersOnLine , vol. 48, no. 27, pp. 54 – 61, 2015, analysis and Design of Hybrid Systems ADHS
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. D. Ames, X. Xu, J. W. Grizzle, and P. Tabuada, “Control barrier function based quadratic programs for safety critical systems,” IEEE Transactions on Automatic Control , vol. 62, no. 8, pp. 3861–3876, 2016
2016
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Cited alongside, same era.
Y. Chow, M. Ghavamzadeh, L. Janson, and M. Pavone, “Risk-constrained reinforcement learning with percentile risk criteria,” The Journal of Machine Learning Research , vol. 18, no. 1, pp. 6070–6120, 2017
2017
Cited alongside, same era.
F. Berkenkamp, M. Turchetta, A. P. Schoellig, and A. Krause, “Safe model-based reinforcement learning with stability guarantees,” in Proceedings of the 31st International Conference on Neural Information Processing Systems , 2017, pp. 908–919
2017
Cited alongside, same era.
B. Amos and J. Z. Kolter, “Optnet: Differentiable optimization as a layer in neural networks,” in International Conference on Machine Learning . PMLR, 2017, pp. 136–145
2017
Cited alongside, same era.
A. Agrawal, B. Amos, S. Barratt, S. Boyd, S. Diamond, and J. Z. Kolter, “Differentiable convex optimization layers,” Advances in Neural Information Processing Systems , vol. 32, pp. 9562–9574, 2019
2019
Later among the works it cites.
Y. Emam, P. Glotfelter, and M. Egerstedt, “Robust barrier functions for a fully autonomous, remotely accessible swarm-robotics testbed,” in 2019 IEEE 58th Conference on Decision and Control (CDC) . IEEE, 2019, pp. 3984–3990
2019
Later among the works it cites.
Y. Li and C. Liu, “Applications of multirotor drone technologies in construction management,” International Journal of Construction Management , vol. 19, no. 5, pp. 401–412, 2019
2019
Later among the works it cites.
A. D. Ames, S. Coogan, M. Egerstedt, G. Notomista, K. Sreenath, and P. Tabuada, “Control barrier functions: Theory and applications,” in 2019 18th European Control Conference (ECC) . IEEE, 2019, pp. 3420–3431
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Wang, A. D. Ames, and M. Egerstedt, “Safety barrier certificates for collisions-free multirobot systems,” IEEE Transactions on Robotics , vol. 33, no. 3, pp. 661–674, 2017
2017
Cited alongside, same era.
Y. Chow, O. Nachum, E. Duenez-Guzman, and M. Ghavamzadeh, “A lyapunov-based approach to safe reinforcement learning,” in Proceedings of the 32nd International Conference on Neural Information Processing Systems , 2018, pp. 8103–8112
2018
Cited alongside, same era.
J. F. Fisac, A. K. Akametalu, M. N. Zeilinger, S. Kaynama, J. Gillula, and C. J. Tomlin, “A general safety framework for learning-based control in uncertain robotic systems,” IEEE Transactions on Automatic Control , vol. 64, no. 7, pp. 2737–2752, 2018
2018
Cited alongside, same era.
Z. Li, U. Kalabić, and T. Chu, “Safe reinforcement learning: Learning with supervision using a constraint-admissible set,” in 2018 Annual American Control Conference (ACC) . IEEE, 2018, pp. 6390–6395
2018
Cited alongside, same era.
S. J. Kim, Y. Jeong, S. Park, K. Ryu, and G. Oh, “A survey of drone use for entertainment and avr (augmented and virtual reality),” in Augmented reality and virtual reality . Springer, 2018, pp. 339–352
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning . PMLR, 2018, pp. 1861–1870
2018
Cited alongside, same era.
E. Squires, P. Pierpaoli, and M. Egerstedt, “Constructive barrier certificates with applications to fixed-wing aircraft collision avoidance,” in 2018 IEEE Conference on Control Technology and Applications (CCTA) . IEEE, 2018, pp. 1656–1661
2018
Cited alongside, same era.
R. Cheng, G. Orosz, R. M. Murray, and J. W. Burdick, “End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, no. 01, 2019, pp. 3387–3395
2019
Cited alongside, same era.
M. Janner, J. Fu, M. Zhang, and S. Levine, “When to trust your model: Model-based policy optimization,” Advances in Neural Information Processing Systems , vol. 32, pp. 12 519–12 530, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2020
Later among the works it cites.
G. Notomista and M. Egerstedt, “Persistification of robotic tasks,” IEEE Transactions on Control Systems Technology , 2020
2020
Later among the works it cites.
B. Thananjeyan, A. Balakrishna, S. Nair, M. Luo, K. Srinivasan, M. Hwang, J. E. Gonzalez, J. Ibarz, C. Finn, and K. Goldberg, “Recovery rl: Safe reinforcement learning with learned recovery zones,” IEEE Robotics and Automation Letters , vol. 6, no. 3, pp. 4915–4922, 2021
2021
Closest in time.
V. Makoviychuk, L. Wawrzyniak, Y. Guo, M. Lu, K. Storey, M. Macklin, D. Hoeller, N. Rudin, A. Allshire, A. Handa et al. , “Isaac gym: High performance gpu based physics simulation for robot learning,” in Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2) , 2021
2021
Closest in time.
Y. Emam, P. Glotfelter, S. Wilson, G. Notomista, and M. Egerstedt, “Data-driven robust barrier functions for safe, long-term operation,” IEEE Transactions on Robotics , pp. 1–15, 2021
2021
Closest in time.
M. Ohnishi, G. Notomista, M. Sugiyama, and M. Egerstedt, “Constraint learning for control tasks with limited duration barrier functions,” Automatica , vol. 127, p. 109504, 2021
2021
Closest in time.