Fetching the paper…
Reading the bibliography…
A common pipeline in learning-based control is to iteratively estimate a model of system dynamics, and apply a trajectory optimization algorithm - e.g.~$\mathtt{iLQR}$ - on the learned model to minimize a target cost.
Differential dynamic programming
David H Jacobson and David Q Mayne · 1970
Earlier work this paper cites.
On the perturbation of pseudo-inverses, projections and linear least squares problems
Gilbert W Stewart · 1977
Earlier work this paper cites.
Bettering operation of robots by learning
Suguru Arimoto, Sadao Kawamura, and Fumio Miyazaki · 1984
Earlier work this paper cites.
Model predictive control: past, present and future
Manfred Morari and Jay H Lee · 1999
Earlier work this paper cites.
On the stability of unconstrained receding horizon control with a general terminal cost
Ali Jadbabaie and John Hauser · 2001
Earlier work this paper cites.
Gaussian process model based predictive control
Juš Kocijan, Roderick Murray-Smith, Carl Edward Rasmussen, and Agathe Girard · 2004
Earlier work this paper cites.
Iterative linear quadratic regulator design for nonlinear biological movement systems
Weiwei Li and Emanuel Todorov · 2004
Earlier work this paper cites.
A generalized iterative lqg method for locally-optimal feedback control of constrained nonlinear stochastic systems
Emanuel Todorov and Weiwei Li · 2005
Earlier work this paper cites.
Optimal control: linear quadratic methods
Brian DO Anderson and John B Moore · 2007
Earlier work this paper cites.
Optimization: algorithms and consistent approximations , volume 124
Elijah Polak · 2012
Earlier work this paper cites.
Synthesis and stabilization of complex behaviors through online trajectory optimization
Yuval Tassa, Tom Erez, and Emanuel Todorov · 2012
Earlier work this paper cites.
User-friendly tail bounds for sums of random matrices
Joel A Tropp · 2012
Earlier work this paper cites.
Provably safe and robust learning-based model predictive control
Anil Aswani, Humberto Gonzalez, S Shankar Sastry, and Claire Tomlin · 2013
Earlier work this paper cites.
Concentration inequalities: A nonasymptotic theory of independence
Stéphane Boucheron, Gábor Lugosi, and Pascal Massart · 2013
Earlier work this paper cites.
Guided policy search
Sergey Levine and Vladlen Koltun · 2013
Earlier work this paper cites.
Learning neural network policies with guided policy search under unknown dynamics
Sergey Levine and Pieter Abbeel · 2014
Earlier work this paper cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition
Hamed Karimi, Julie Nutini, and Mark Schmidt · 2016
Cited alongside, same era.
On the sample complexity of the linear quadratic regulator working draft
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2017
Cited alongside, same era.
Information theoretic mpc for model-based reinforcement learning
Grady Williams, Nolan Wagener, Brian Goldfain, Paul Drews, James M Rehg, Byron Boots, and Evangelos A Theodorou · 2017
Cited alongside, same era.
JAX: composable transformations of Python+NumPy programs, 2018
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, George Necula, Adam Paszke, Jake VanderPlas, Skye Wanderman-Milne, and Qiao Zhang · 2018
Cited alongside, same era.
Learning-based model predictive control for safe exploration
Torsten Koller, Felix Berkenkamp, Matteo Turchetta, and Andreas Krause · 2018
Cited alongside, same era.
Learning nonlinear dynamical systems from a single trajectory
Dylan Foster, Tuhin Sarkar, and Alexander Rakhlin · 2020
Later among the works it cites.
Haiku: Sonnet for JAX, 2020
Tom Hennigan, Trevor Cai, Tamara Norman, and Igor Babuschkin · 2020
Later among the works it cites.
Active learning for nonlinear system identification with guarantees
Horia Mania, Michael I Jordan, and Benjamin Recht · 2020
Later among the works it cites.
Learning the linear quadratic regulator from nonlinear observations
Zakaria Mhammedi, Dylan J Foster, Max Simchowitz, Dipendra Misra, Wen Sun, Akshay Krishnamurthy, Alexander Rakhlin, and John Langford · 2020
Later among the works it cites.
Control of unknown nonlinear systems with linear time-varying mpc
Dimitris Papadimitriou, Ugo Rosolia, and Francesco Borrelli · 2020
Later among the works it cites.
Naive exploration is optimal for online lqr
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning without mixing: Towards a sharp analysis of linear system identification
Max Simchowitz, Horia Mania, Stephen Tu, Michael I Jordan, and Benjamin Recht · 2018
Cited alongside, same era.
High-dimensional probability: An introduction with applications in data science , volume 47
Roman Vershynin · 2018
Cited alongside, same era.
Online control with adversarial disturbances
Naman Agarwal, Brian Bullins, Elad Hazan, Sham Kakade, and Karan Singh · 2019
Cited alongside, same era.
Learning-based model predictive control for autonomous racing
Juraj Kabzan, Lukas Hewing, Alexander Liniger, and Melanie N Zeilinger · 2019
Cited alongside, same era.
Deep dynamics models for learning dexterous manipulation. arxiv
A Nagabandi, K Konoglie, S Levine, and V Kumar · 2019
Cited alongside, same era.
Non-asymptotic identification of lti systems from a single trajectory
Samet Oymak and Necmiye Ozay · 2019
Cited alongside, same era.
Learning how to autonomously race a car: a predictive control approach
Ugo Rosolia and Francesco Borrelli · 2019
Cited alongside, same era.
Max Simchowitz and Dylan Foster · 2020
Later among the works it cites.
On the perturbation of the moore–penrose inverse of a matrix
Xuefeng Xu · 2020
Later among the works it cites.
Black-box control for linear dynamical systems
Xinyi Chen and Elad Hazan · 2021
Later among the works it cites.
Lyapunov-stable neural-network control
Hongkai Dai, Benoit Landry, Lujie Yang, Marco Pavone, and Russ Tedrake · 2021
Later among the works it cites.
Certainty equivalent perception-based control
Sarah Dean and Benjamin Recht · 2021
Later among the works it cites.
trajax: differentiable optimal control on accelerators, 2021
Roy Frostig, Vikas Sindhwani, Sumeet Singh, and Stephen Tu · 2021
Later among the works it cites.
Data-driven mpc for quadrotors
Guillem Torrente, Elia Kaufmann, Philipp Föhn, and Davide Scaramuzza · 2021
Later among the works it cites.
On the stability of nonlinear receding horizon control: a geometric perspective
Tyler Westenbroek, Max Simchowitz, Michael I Jordan, and S Shankar Sastry · 2021
Later among the works it cites.
Tasil: Taylor series imitation learning
Daniel Pfrommer, Thomas TCK Zhang, Stephen Tu, and Nikolai Matni · 2022
Later among the works it cites.
Non-asymptotic and accurate learning of nonlinear dynamical systems
Yahya Sattar and Samet Oymak · 2022
Later among the works it cites.
Learning to control linear systems can be hard
Anastasios Tsiamis, Ingvar M Ziemann, Manfred Morari, Nikolai Matni, and George J Pappas · 2022
Later among the works it cites.