Fetching the paper…
Reading the bibliography…
This paper presents an end-to-end framework for safe learning-based control (LbC) using nonlinear stochastic MPC and distributionally robust optimization (DRO).
K. J. Astrom, Introduction to Stochastic Control Theory . Courier Corporation, 1970
1970
Earlier work this paper cites.
M. Doyle, T. Fuller, and J. Newman, “Modeling of galvanostatic charge and discharge of the lithium/polymer/insertion cell,” Journal of the Electrochemical Society , vol. 140, no. 6, pp. 1526–1533, 1993
1993
Earlier work this paper cites.
M. V. Kothare, V. Balakrishnan, and M. Morari, “Robust constrained model predictive control using linear matrix inequalities,” Automatica , vol. 32, no. 10, pp. 1361–1379, 1996
1996
Earlier work this paper cites.
H. Beyer and H. Schwefel, “Evolution strategies - a comprehensive introduction,” Natural Computing , vol. 1, no. 1, pp. 3–52, March 2002
2002
Earlier work this paper cites.
T. J. Perkins and A. G. Barto, “Lyapunov design for safe reinforcement learning,” J. Mach. Learn. Res. , vol. 3, no. null, p. 803–832, mar 2003
2003
Earlier work this paper cites.
A. Nilim and L. E. Ghaoui, “Robust control of markov decision processes with uncertain transition matrices,” Operations Research , vol. 53, no. 5, 2005
2005
Earlier work this paper cites.
2010
Earlier work this paper cites.
M. Tanaskovic, L. Fagiano, R. Smith, and M. Morari, “Adaptive receding horizon control for constrained mimo systems,” Automatica , vol. 50, pp. 3019–3029, 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
J. Garcia and F. Fernandes, “A comprehensive survey on safe reinforcement learning,” Journal of Machine Learning Research , vol. 16, pp. 1437–1480, 2016
2016
Earlier work this paper cites.
B. P. V. Parys, D. Kuhn, P. J. Goulart, and M. Morari, “Distributionally robust control of constrained stochastic systems,” IEEE Transactions on Automatic Control , vol. 61, no. 2, pp. 430–442, 2016
2016
Earlier work this paper cites.
R. Gao and A. J. Kleywegt, “Distributionally robust stochastic optimization with wasserstein distance,” arXiv , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
U. Rosolia and F. Borrelli, “Learning model predictive control for iterative tasks. a data-driven control framework,” IEEE Transactions on Automatic Control , vol. 63, no. 7, pp. 1883–1896, 2017
2017
Cited alongside, same era.
J. Paulson, E. Buehler, and A. Mesbah, “Arbitrary polynomial chaos for uncertainty propagation of correlated random variables in dynamic systems,” IFAC PapersOnLine , vol. 50, no. 1, pp. 3548–3553, 2017
2017
Cited alongside, same era.
I. Yang, “A convex optimization approach to distributionally robust markov decision processes with wasserstein distance,” IEEE Control Systems Letters , vol. 1, no. 1, pp. 164–169, 2017
2017
Cited alongside, same era.
F. Berkenkamp, M. Turchetta, A. Schoellig, and A. Krause, “Safe model-based reinforcement learning with stability guarantees,” in Advances in Neural Information Processing Systems , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, Eds., vol. 30. Curran Associates, Inc., 2017
2017
S. Dean, S. Tu, N. Matni, and B. Recht, “Safely learning to control the constrained linear quadratic regulator,” in Proceedings of the 2019 American Control Conference . Philadelphia, PA, USA: IEEE, 2019
2019
Later among the works it cites.
R. Cheng, G. Orosz, R. Murray, and J. Burdick, “End-to-end safe reinforcement learning through barrier functions for safety-critical continuous control tasks,” in AAAI , 2019
2019
Later among the works it cites.
E. Lecarpentier and E. Rachelson, “Non-stationary markov decision processes, a worst-case approach using model-based reinforcement learning,” in Advances in Neural Information Processing Systems 32 . Curran Associates, Inc., 2019, pp. 7216–7225
2019
Later among the works it cites.
I. Akbar, “Uncertainty estimation in continuous models applied to reinforcement learning,” Ph.D. dissertation, UC San Diego, 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
J. Achiam, D. Held, A. Tamar, and P. Abbeel, “Constrained policy optimization,” in Proceedings of the 2017 International Conference on Machine Learning (ICML) . Sydney, Australia: PMLR, 2017
2017
Cited alongside, same era.
H. Perez, X. Hu, S. Dey, and S. Moura, “Optimal charging of li-ion batteries with coupled electro-thermal-aging dynamics,” IEEE Transactions on Vehicular Technology , vol. 66, no. 7, pp. 7761–7770, 2017
2017
Cited alongside, same era.
M. Bujarbaruah, X. Zhang, and F. Borrelli, “Adaptive mpc with chance constraints for fir systems,” 2018
2018
Cited alongside, same era.
T. Koller, F. Berkenkamp, M. Turchetta, and A. Krause, “Learning-based model predictive control for safe exploration,” arXiv , 2018
2018
Cited alongside, same era.
Y. Chow, O. Nachum, E. Duenez-Guzman, and M. Ghavamzadeh, “A lyapunov-based approach to safe reinforcement learning,” in Advances in Neural Information Processing Systems , S. Bengio, H. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, and R. Garnett, Eds., vol. 31. Curran Associates, Inc., 2018
2018
Cited alongside, same era.
P. Esfahani and D. Kuhn, “Data-driven distributionally robust optimization using the wasserstein metric: Performance guarantees and tractable reformulations,” Mathematical Programming , vol. 171, no. 1–2, pp. 115–166, 2018
2018
Cited alongside, same era.
C. Zhao and Y. Guan, “Data-driven risk-averse stochastic optimization with wasserstein metric,” Operations Research Letters , vol. 46, no. 2, pp. 262–267, 2018
2018
Cited alongside, same era.
I. Yang, “Wasserstein distributionally robust stochastic control: A data-driven approach,” arXiv , 2018
2018
Cited alongside, same era.
L. Hewing, K. P. Wabersich, M. Menner, and M. N. Zeilinger, “Learning-based model predictive control: Toward safe learning in control,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 3, no. 1, pp. 269–296, 2020. [Online]. Available: https://doi.org/10.1146/annurev-control-090419-075625
2020
Closest in time.
D. D. Fan, J. Nguyen, R. Thakker, N. Alatur, A. akbar Agha-mohammadi, and E. Theodorou, “Bayesian learning-based adaptive control for safety critical systems,” 2020 IEEE International Conference on Robotics and Automation (ICRA) , pp. 4093–4099, 2020
2020
Closest in time.
J. Choi, F. Castañeda, C. J. Tomlin, and K. Sreenath, “Reinforcement learning for safety-critical control under model uncertainty, using control lyapunov functions and control barrier functions,” 2020
2020
Closest in time.
M. J. Khojasteh, V. Dhiman, M. Franceschetti, and N. Atanasov, “Probabilistic safety constraints for learned high relative degree system dynamics,” in L4DC , 2020
2020
Closest in time.
A. Kandel and S. Moura, “Safe wasserstein constrained deep q-learning,” arXiv , 2020
2020
Closest in time.
T. Westenbroek, A. Agrawal, F. Castaneda, S. Sastry, and K. Sreenath, “Combining model-based design and model-free policy optimization to learn safe, stabilizing controllers,” in Proceedings of the 7th IFAC Conference on Analysis and Design of Hybrid Systems , 2021
2021
Closest in time.
J. Coulson, J. Lygeros, and F. Dorfler, “Distributionally robust chance constrained data-enabled predictive control,” IEEE Transactions on Automatic Control , pp. 1–1, 2021
2021
Closest in time.
2021
Closest in time.
A. Kandel, S. Park, and S. Moura, “Distributionally robust surrogate optimal control for high-dimensional systems,” 2021
2021
Closest in time.
2021
Closest in time.
A. Kandel, “Wasserstein Nonlinear MPC,” Aug. 2023. [Online]. Available: https://github.com/aaronkandel/Wasserstein-Nonlinear-MPC/tree/main
2023
Closest in time.