Fetching the paper…
Reading the bibliography…
We consider adaptive control of the Linear Quadratic Regulator (LQR), where an unknown linear system is controlled subject to quadratic costs.
Nonlinear and adaptive control design
Miroslav Krstic, Ioannis Kanellakopoulos, and Peter V Kokotovic · 1995
Earlier work this paper cites.
K. Zhou, J. C. Doyle, and K. Glover · 1995
Earlier work this paper cites.
Robust adaptive control , volume 1
Petros A Ioannou and Jing Sun · 1996
Earlier work this paper cites.
PAC Adaptive Control of Linear Systems
Claude-Nicolas Fiechter · 1997
Earlier work this paper cites.
Adaptive control of linear time invariant systems: the “bet on the best” principle
S. Bittanti and M. C. Campi · 2006
Earlier work this paper cites.
Relaxing dynamic programming
Bo Lincoln and Anders Rantzer · 2006
Earlier work this paper cites.
Positive trigonometric polynomials and signal processing applications
Bogdan Dumitrescu · 2007
Earlier work this paper cites.
Regret Bounds for the Adaptive Control of Linear Quadratic Systems
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2011
Earlier work this paper cites.
Hanson-Wright inequality and sub-gaussian concentration
Mark Rudelson and Roman Vershynin · 2011
Earlier work this paper cites.
Online Learning for Linearly Parametrized Control Problems
Yasin Abbasi-Yadkori · 2012
Cited alongside, same era.
Efficient Reinforcement Learning for High Dimensional Linear Quadratic Systems
Morteza Ibrahimi, Adel Javanmard, and Benjamin Van Roy · 2012
Cited alongside, same era.
Bayesian Optimal Control of Smoothly Parameterized Systems: The Lazy Posterior Sampling Algorithm
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2015
Cited alongside, same era.
CVXPY: A Python-embedded modeling language for convex optimization
Steven Diamond and Stephen Boyd · 2016
Cited alongside, same era.
Conic Optimization via Operator Splitting and Homogeneous Self-Dual Embedding
Brendan O’Donoghue, Eric Chu, Neal Parikh, and Stephen Boyd · 2016
Cited alongside, same era.
Posterior Sampling for Reinforcement Learning Without Episodes
On the Sample Complexity of the Linear Quadratic Regulator
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2017
Later among the works it cites.
Finite Time Analysis of Optimal Adaptive Policies for Linear-Quadratic Systems
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2017
Later among the works it cites.
Scalable system level synthesis for virtually localizable systems
Nikolai Matni, Yuh-Shyang Wang, and James Anderson · 2017
Later among the works it cites.
Learning-based Control of Unknown Linear Systems with Thompson Sampling
Yi Ouyang, Mukul Gagrani, and Rahul Jain · 2017
Later among the works it cites.
Least-Squares Temporal Difference Learning for the Linear Quadratic Regulator
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ian Osband and Benjamin Van Roy · 2016
Cited alongside, same era.
A System Level Approach to Controller Synthesis
Yuh-Shyang Wang, Nikolai Matni, and John C Doyle · 2016
Cited alongside, same era.
Thompson Sampling for Linear-Quadratic Control Problems
Marc Abeille and Alessandro Lazaric · 2017
Cited alongside, same era.
Structured State Space Realizations for SLS Distributed Controllers
James Anderson and Nikolai Matni · 2017
Cited alongside, same era.
Stephen Tu and Benjamin Recht · 2017
Later among the works it cites.
Regret Bounds for Model-Free Linear Quadratic Control
Yasin Abbasi-Yadkori, Nevena Lazic, and Csaba Szepesvári · 2018
Closest in time.
Global Convergence of Policy Gradient Methods for Linearized Control Problems
Maryam Fazel, Rong Ge, Sham M. Kakade, and Mehran Mesbahi · 2018
Closest in time.
Learning Without Mixing: Towards A Sharp Analysis of Linear System Identification
Max Simchowitz, Horia Mania, Stephen Tu, Michael I Jordan, and Benjamin Recht · 2018
Closest in time.