Fetching the paper…
Reading the bibliography…
Reinforcement learning (RL) exhibits impressive performance when managing complicated control tasks for robots.
H. L. Royden and P. Fitzpatrick, Real analysis . Macmillan New York, 1988, vol. 32
1988
Earlier work this paper cites.
D. P. Bertsekas, “Nonlinear programming,” Journal of the Operational Research Society , vol. 48, no. 3, pp. 334–334, 1997
1997
Earlier work this paper cites.
E. Altman, Constrained Markov decision processes: stochastic modeling . Routledge, 1999
1999
Earlier work this paper cites.
P. Geibel and F. Wysotzki, “Risk-sensitive reinforcement learning applied to control under constraints,” Journal of Artificial Intelligence Research , vol. 24, pp. 81–108, 2005
2005
Earlier work this paper cites.
Y. Shen, M. J. Tobia, T. Sommer, and K. Obermayer, “Risk-sensitive reinforcement learning,” Neural computation , vol. 26, no. 7, pp. 1298–1328, 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
A. D. Ames, J. W. Grizzle, and P. Tabuada, “Control barrier function based quadratic programs with application to adaptive cruise control,” in 53rd IEEE Conference on Decision and Control . IEEE, 2014, pp. 6271–6278
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
Q. Nguyen, A. Hereid, J. W. Grizzle, A. D. Ames, and K. Sreenath, “3d dynamic walking on stepping stones with control barrier functions,” in 2016 IEEE 55th Conference on Decision and Control (CDC) , 2016, pp. 827–834
2016
Earlier work this paper cites.
M. Z. Romdlony and B. Jayawardhana, “Stabilization with guaranteed safety using control lyapunov–barrier function,” Automatica , vol. 66, pp. 39–47, 2016
2016
Earlier work this paper cites.
M. Turchetta, F. Berkenkamp, and A. Krause, “Safe exploration in finite markov decision processes with gaussian processes,” Advances in Neural Information Processing Systems , vol. 29, 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Achiam, D. Held, A. Tamar, and P. Abbeel, “Constrained policy optimization,” in International conference on machine learning . PMLR, 2017, pp. 22–31
2017
Earlier work this paper cites.
F. Berkenkamp, M. Turchetta, A. Schoellig, and A. Krause, “Safe model-based reinforcement learning with stability guarantees,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
A. Agrawal and K. Sreenath, “Discrete control barrier functions for safety-critical control of discrete systems with application to bipedal robot navigation.” in Robotics: Science and Systems , vol. 13. Cambridge, MA, USA, 2017
2017
Cited alongside, same era.
2018
Cited alongside, same era.
T.-H. Pham, G. De Magistris, and R. Tachibana, “Optlayer-practical constrained optimization for deep reinforcement learning in the real world,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 6236–6243
2018
Cited alongside, same era.
T. Gurriet, A. Singletary, J. Reher, L. Ciarletta, E. Feron, and A. Ames, “Towards a framework for realizable safety critical control through active set invariance,” in 2018 ACM/IEEE 9th International Conference on Cyber-Physical Systems (ICCPS) . IEEE, 2018, pp. 98–106
J. Choi, F. Castañeda, C. J. Tomlin, and K. Sreenath, “Reinforcement learning for safety-critical control under model uncertainty, using control lyapunov functions and control barrier functions,” in Robotics: Science and Systems (RSS) , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
Z.-P. Jiang, T. Bian, W. Gao et al. , “Learning-based control: A tutorial and some recent results,” Foundations and Trends® in Systems and Control , vol. 8, no. 3, pp. 176–284, 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Y. Chow, O. Nachum, E. Duenez-Guzman, and M. Ghavamzadeh, “A lyapunov-based approach to safe reinforcement learning,” Advances in neural information processing systems , vol. 31, 2018
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning . PMLR, 2018, pp. 1861–1870
2018
Cited alongside, same era.
N. O. Lambert, D. S. Drew, J. Yaconelli, S. Levine, R. Calandra, and K. S. Pister, “Low-level control of a quadrotor with deep model-based reinforcement learning,” IEEE Robotics and Automation Letters , vol. 4, no. 4, pp. 4224–4230, 2019
2019
Cited alongside, same era.
R. Cheng, G. Orosz, R. M. Murray, and J. W. Burdick, “End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, no. 01, 2019, pp. 3387–3395
2019
Cited alongside, same era.
M. Ohnishi, L. Wang, G. Notomista, and M. Egerstedt, “Barrier-certified adaptive reinforcement learning with applications to brushbot navigation,” IEEE Transactions on robotics , vol. 35, no. 5, pp. 1186–1205, 2019
2019
Cited alongside, same era.
A. J. Taylor, V. D. Dorobantu, H. M. Le, Y. Yue, and A. D. Ames, “Episodic learning with control lyapunov functions for uncertain robotic systems,” in 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2019, pp. 6878–6884
2019
Cited alongside, same era.
2019
Cited alongside, same era.
A. Nagabandi, K. Konolige, S. Levine, and V. Kumar, “Deep dynamics models for learning dexterous manipulation,” in Conference on Robot Learning . PMLR, 2020, pp. 1101–1112
2020
Cited alongside, same era.
M. Han, L. Zhang, J. Wang, and W. Pan, “Actor-critic reinforcement learning for control with stability guarantee,” IEEE Robotics and Automation Letters , vol. 5, no. 4, pp. 6217–6224, 2020
2020
Later among the works it cites.
S. Belkhale, R. Li, G. Kahn, R. McAllister, R. Calandra, and S. Levine, “Model-based meta-reinforcement learning for flight with suspended payloads,” IEEE Robotics and Automation Letters , vol. 6, no. 2, pp. 1471–1478, 2021
2021
Later among the works it cites.
G. Thomas, Y. Luo, and T. Ma, “Safe reinforcement learning by imagining the near future,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Later among the works it cites.
R. Grandia, A. J. Taylor, A. D. Ames, and M. Hutter, “Multi-layered safety for legged robots via control barrier functions and model predictive control,” in 2021 IEEE International Conference on Robotics and Automation (ICRA) , 2021, pp. 8352–8358
2021
Later among the works it cites.
B. Thananjeyan, A. Balakrishna, S. Nair, M. Luo, K. Srinivasan, M. Hwang, J. E. Gonzalez, J. Ibarz, C. Finn, and K. Goldberg, “Recovery rl: Safe reinforcement learning with learned recovery zones,” IEEE Robotics and Automation Letters , vol. 6, no. 3, pp. 4915–4922, 2021
2021
Later among the works it cites.
L. Brunke, M. Greeff, A. W. Hall, Z. Yuan, S. Zhou, J. Panerati, and A. P. Schoellig, “Safe learning in robotics: From learning-based control to safe reinforcement learning,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 5, 2021
2021
Later among the works it cites.
M. Han, Y. Tian, L. Zhang, J. Wang, and W. Pan, “Reinforcement learning control of constrained dynamic systems with uniformly ultimate boundedness stability guarantee,” Automatica , vol. 129, p. 109689, 2021
2021
Later among the works it cites.
J. Panerati, H. Zheng, S. Zhou, J. Xu, A. Prorok, and A. P. Schoellig, “Learning to fly—a gym environment with pybullet physics for reinforcement learning of multi-agent quadcopter control,” in 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2021
2021
Later among the works it cites.
C. Dawson, Z. Qin, S. Gao, and C. Fan, “Safe nonlinear control using robust neural lyapunov-barrier functions,” in Conference on Robot Learning . PMLR, 2022, pp. 1724–1735
2022
Later among the works it cites.
Y. Meng, Y. Li, M. Fitzsimmons, and J. Liu, “Smooth converse lyapunov-barrier theorems for asymptotic stability with safety constraints and reach-avoid-stay specifications,” Automatica , vol. 144, p. 110478, 2022
2022
Later among the works it cites.