Fetching the paper…
Reading the bibliography…
In this paper, we seek to learn a robot policy guaranteed to satisfy state constraints.
Numerical methods for solving linear least squares problems
Gene Golub · 1965
Earlier work this paper cites.
System identification
Lennart Ljung · 1998
Earlier work this paper cites.
Fundamental limitations and differences of robust and adaptive control
Le Yi Wang and Ji-Feng Zhang · 2001
Earlier work this paper cites.
Convex Polytopes
Branko Grünbaum, Volker Kaibel, Victor Klee, and Gtinter M. Ziegler · 2003
Earlier work this paper cites.
A multi-model structure for model predictive control
Federico Di Palma and Lalo Magni · 2004
Earlier work this paper cites.
A control problem for affine dynamical systems on a full-dimensional polytope
L. Habets and J. H. Van Schuppen · 2004
Earlier work this paper cites.
Basic Real Analysis
Anthony W Knapp · 2005
Earlier work this paper cites.
Robust model predictive control: A survey
Alberto Bemporad and Manfred Morari · 2007
Earlier work this paper cites.
Optimal Control of Constrained Piecewise Affine Systems
Frank J Christophersen · 2007
Earlier work this paper cites.
A fully automated framework for control of linear systems from temporal logic specifications
M. Kloetzer and C. Belta · 2008
Earlier work this paper cites.
δ \delta -complete decision procedures for satisfiability over the reals
Sicun Gao, Jeremy Avigad, and Edmund M Clarke · 2012
Earlier work this paper cites.
Synthesis for constrained nonlinear systems using hybridization and robust controllers on simplices
A. Girard and S. Martin · 2012
Earlier work this paper cites.
Next generation airborne collision avoidance system
Mykel J Kochenderfer, Jessica E Holland, and James P Chryssanthacopoulos · 2012
Earlier work this paper cites.
MuJoCo: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Provably safe and robust learning-based model predictive control
Anil Aswani, Humberto Gonzalez, S Shankar Sastry, and Claire Tomlin · 2013
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2016
Earlier work this paper cites.
Stochastic model predictive control: An overview and perspectives for future research
Ali Mesbah · 2016
Earlier work this paper cites.
Exponential control barrier functions for enforcing high relative-degree safety-critical constraints
Quan Nguyen and Koushil Sreenath · 2016
Earlier work this paper cites.
Constrained policy optimization
Joshua Achiam, David Held, Aviv Tamar, and Pieter Abbeel · 2017
Earlier work this paper cites.
OptNet: Differentiable optimization as a layer in neural networks
Brandon Amos and J Zico Kolter · 2017
Earlier work this paper cites.
Stochastic MPC with offline uncertainty sampling
Matthias Lorenzen, Fabrizio Dabbene, Roberto Tempo, and Frank Allgöwer · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
A spline theory of deep learning
Randall Balestriero and Richard Baraniuk · 2018
Cited alongside, same era.
Safe exploration in continuous action spaces
Gal Dalal, Krishnamurthy Dvijotham, Matej Vecerik, Todd Hester, Cosmin Paduraru, and Yuval Tassa · 2018
Cited alongside, same era.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke Hoof, and David Meger · 2018
Cited alongside, same era.
Optlayer - practical constrained optimization for deep reinforcement learning in the real world
Tu-Hoa Pham, Giovanni De Magistris, and Ryuki Tachibana · 2018
Cited alongside, same era.
Challenges of real-world reinforcement learning: Definitions, benchmarks and analysis
Gabriel Dulac-Arnold, Nir Levine, Daniel J Mankowitz, Jerry Li, Cosmin Paduraru, Sven Gowal, and Todd Hester · 2021
Later among the works it cites.
A sample-efficient algorithm for episodic finite-horizon MDP with constraints
Krishna C Kalagarla, Rahul Jain, and Pierluigi Nuzzo · 2021
Later among the works it cites.
Planning with learned dynamics: Probabilistic guarantees on safety and reachability via Lipschitz constants
Craig Knuth, Glen Chou, Necmiye Ozay, and Dmitry Berenson · 2021
Later among the works it cites.
Model-based constrained reinforcement learning using generalized control barrier function
Haitong Ma, Jianyu Chen, Shengbo Eben, Ziyu Lin, Yang Guan, Yangang Ren, and Sifa Zheng · 2021
Later among the works it cites.
Reachable polyhedral marching (RPM): A safety verification algorithm for robotic systems with deep neural network components
Joseph A Vincent and Mac Schwager · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, et al · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
Towards fast computation of certified robustness for ReLU networks
Lily Weng, Huan Zhang, Hongge Chen, Zhao Song, Cho-Jui Hsieh, Luca Daniel, Duane Boning, and Inderjit Dhillon · 2018
Cited alongside, same era.
Efficient neural network robustness certification with general activation functions
Huan Zhang, Tsui-Wei Weng, Pin-Yu Chen, Cho-Jui Hsieh, and Luca Daniel · 2018
Cited alongside, same era.
Differentiable convex optimization layers
Akshay Agrawal, Brandon Amos, Shane Barratt, Stephen Boyd, Steven Diamond, and J Zico Kolter · 2019
Cited alongside, same era.
Control barrier functions: Theory and applications
Aaron D Ames, Samuel Coogan, Magnus Egerstedt, Gennaro Notomista, Koushil Sreenath, and Paulo Tabuada · 2019
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemysław Dębiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, et al · 2019
Cited alongside, same era.
Probabilistic model predictive safety certification for learning-based control
Kim P Wabersich, Lukas Hewing, Andrea Carron, and Melanie N Zeilinger · 2021
Later among the works it cites.
A predictive safety filter for learning-based control of constrained nonlinear dynamical systems
Kim Peter Wabersich and Melanie N Zeilinger · 2021
Later among the works it cites.
Safe learning in robotics: From learning-based control to safe reinforcement learning
Lukas Brunke, Melissa Greeff, Adam W Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P Schoellig · 2022
Later among the works it cites.
Robust safe control synthesis with disturbance observer-based control barrier functions
Ersin Daş and Richard M Murray · 2022
Later among the works it cites.
A review of safe reinforcement learning: Methods, theory and applications
Shangding Gu, Long Yang, Yali Du, Guang Chen, Florian Walter, Jun Wang, Yaodong Yang, and Alois Knoll · 2022
Later among the works it cites.
Joint synthesis of safety certificate and safe control policy using constrained reinforcement learning
Haitong Ma, Changliu Liu, Shengbo Eben Li, Sifa Zheng, and Jianyu Chen · 2022
Later among the works it cites.
Sablas: Learning safe control for black-box dynamical systems
Zengyi Qin, Dawei Sun, and Chuchu Fan · 2022
Later among the works it cites.
Model-free safe control for zero-violation reinforcement learning
Weiye Zhao, Tairan He, and Changliu Liu · 2022
Later among the works it cites.
POLICE: Provably optimal linear constraint enforcement for deep neural networks
Randall Balestriero and Yann LeCun · 2023
Later among the works it cites.
A new computationally simple approach for implementing neural networks with output hard constraints
Andrei V Konstantinov and Lev V Utkin · 2023
Later among the works it cites.
Constrained decision transformer for offline safe reinforcement learning
Zuxin Liu, Zijian Guo, Yihang Yao, Zhepeng Cen, Wenhao Yu, Tingnan Zhang, and Ding Zhao · 2023
Later among the works it cites.
ConBaT: Control barrier transformer for safe policy learning
Yue Meng, Sai Vemprala, Rogerio Bonatti, Chuchu Fan, and Ashish Kapoor · 2023
Later among the works it cites.
Backward reachability analysis of neural feedback loops: Techniques for linear and nonlinear systems
Nicholas Rober, Sydney M Katz, Chelsea Sidrane, Esen Yel, Michael Everett, Mykel J Kochenderfer, and Jonathan P How · 2023
Later among the works it cites.
RAYEN: Imposition of hard convex constraints on neural networks
Jesus Tordesillas, Jonathan P How, and Marco Hutter · 2023
Later among the works it cites.
Enforcing hard constraints with soft barriers: Safe reinforcement learning in unknown stochastic environments
Yixuan Wang, Simon Sinong Zhan, Ruochen Jiao, Zhilu Wang, Wanxin Jin, Zhuoran Yang, Zhaoran Wang, Chao Huang, and Qi Zhu · 2023
Later among the works it cites.
BarrierNet: Differentiable control barrier functions for learning of safe robot control
Wei Xiao, Tsun-Hsuan Wang, Ramin Hasani, Makram Chahine, Alexander Amini, Xiao Li, and Daniela Rus · 2023
Later among the works it cites.
Model-free safe reinforcement learning through neural barrier certificate
Yujie Yang, Yuxuan Jiang, Yichen Liu, Jianyu Chen, and Shengbo Eben Li · 2023
Later among the works it cites.