Fetching the paper…
Reading the bibliography…
We consider the problem of controlling a possibly unknown linear dynamical system with adversarial perturbations, adversarially chosen convex loss functions, and partially observed states, known as non-stochastic control.
Stability of discrete linear feedback systems
Vladimír Kučera · 1975
Earlier work this paper cites.
Modern wiener-hopf design of optimal controllers–part ii: The multivariable case
Dante Youla, Hamid Jabr, and Jr Bongiorno · 1976
Earlier work this paper cites.
Feedback and optimal sensitivity: Model reference transformations, multiplicative seminorms, and approximate inverses
George Zames · 1981
Earlier work this paper cites.
Stable lqg controllers
Yoram Halevi · 1994
Earlier work this paper cites.
System identification
Lennart Ljung · 1999
Earlier work this paper cites.
Lecture 10: Q-parametrization
Alexander Megretski · 2004
Earlier work this paper cites.
Smoothed analysis of algorithms: Why the simplex algorithm usually takes polynomial time
Daniel A Spielman and Shang-Hua Teng · 2004
Earlier work this paper cites.
Dynamic programming and optimal control , volume 1
Dimitri Bertsekas · 2005
Earlier work this paper cites.
A characterization of convex problems in decentralized control
Michael Rotkowitz and Sanjay Lall · 2005
Earlier work this paper cites.
Prediction, learning, and games
Nicolo Cesa-Bianchi and Gábor Lugosi · 2006
Earlier work this paper cites.
Optimization over state feedback policies for robust control with constraints
Paul J Goulart, Eric C Kerrigan, and Jan M Maciejowski · 2006
Earlier work this paper cites.
Optimal control: linear quadratic methods
Brian DO Anderson and John B Moore · 2007
Earlier work this paper cites.
H-infinity optimal control and related minimax design problems: a dynamic game approach
Tamer Başar and Pierre Bernhard · 2008
Earlier work this paper cites.
Online Markov decision processes
Eyal Even-Dar, Sham M Kakade, and Yishay Mansour · 2009
Earlier work this paper cites.
An efficient projection for ℓ 1 , ∞ \ell_{1,\infty} regularization
Ariadna Quattoni, Xavier Carreras, Michael Collins, and Trevor Darrell · 2009
Earlier work this paper cites.
Regret bounds for the adaptive control of linear quadratic systems
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2011
Earlier work this paper cites.
Online least squares estimation with self-normalized processes: An application to bandit problems
Yasin Abbasi-Yadkori, Dávid Pál, and Csaba Szepesvári · 2011
Earlier work this paper cites.
Adaptive control: stability, convergence and robustness
Shankar Sastry and Marc Bodson · 2011
Cited alongside, same era.
Robust adaptive control
Petros A Ioannou and Jing Sun · 2012
Cited alongside, same era.
Online learning and online convex optimization
Shai Shalev-Shwartz et al · 2012
Cited alongside, same era.
Better rates for any adversarial deterministic MDP
Ofer Dekel and Elad Hazan · 2013
Cited alongside, same era.
Online learning in episodic markovian decision processes by relative entropy policy search
Alexander Zimin and Gergely Neu · 2013
Cited alongside, same era.
Tracking adversarial targets
Yasin Abbasi-Yadkori, Peter Bartlett, and Varun Kanade · 2014
Cited alongside, same era.
Regret bounds for robust adaptive control of the linear quadratic regulator
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2018
Later among the works it cites.
Input perturbations for adaptive regulation and learning
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2018
Later among the works it cites.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham M Kakade, and Mehran Mesbahi · 2018
Later among the works it cites.
Spectral filtering for general linear dynamical systems
Elad Hazan, Holden Lee, Karan Singh, Cyril Zhang, and Yi Zhang · 2018
Later among the works it cites.
Learning without mixing: Towards a sharp analysis of linear system identification
Max Simchowitz, Horia Mania, Stephen Tu, Michael I Jordan, and Benjamin Recht · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Smoothed analysis of tensor decompositions
Aditya Bhaskara, Moses Charikar, Ankur Moitra, and Aravindan Vijayaraghavan · 2014
Cited alongside, same era.
First-order methods of smooth convex optimization with inexact oracle
Olivier Devolder, François Glineur, and Yurii Nesterov · 2014
Cited alongside, same era.
Online learning for adversaries with memory: price of past mistakes
Oren Anava, Elad Hazan, and Shie Mannor · 2015
Cited alongside, same era.
Introduction to online convex optimization
Elad Hazan · 2016
Cited alongside, same era.
Introduction to online convex optimization
Elad Hazan et al · 2016
Cited alongside, same era.
On the complexity of best-arm identification in multi-armed bandit models
Emilie Kaufmann, Olivier Cappé, and Aurélien Garivier · 2016
Cited alongside, same era.
High-dimensional probability: An introduction with applications in data science , volume 47
Roman Vershynin · 2018
Later among the works it cites.
Learning linear-quadratic regulators efficiently with only 𝒪 ( T ) \mathcal{O}(\sqrt{T}) regret
Alon Cohen, Tomer Koren, and Yishay Mansour · 2019
Later among the works it cites.
An input–output parametrization of stabilizing controllers: Amidst youla and system level synthesis
Luca Furieri, Yang Zheng, Antonis Papachristodoulou, and Maryam Kamgarpour · 2019
Later among the works it cites.
The nonstochastic control problem
Elad Hazan, Sham M. Kakade, and Karan Singh · 2019
Later among the works it cites.
Certainty equivalent control of lqr is efficient
Horia Mania, Stephen Tu, and Benjamin Recht · 2019
Later among the works it cites.
Non-asymptotic identification of lti systems from a single trajectory
Samet Oymak and Necmiye Ozay · 2019
Later among the works it cites.
Finite-time system identification for partially observed lti systems of unknown order
Tuhin Sarkar, Alexander Rakhlin, and Munther A Dahleh · 2019
Later among the works it cites.
Learning linear dynamical systems with semi-parametric least squares
Max Simchowitz, Ross Boczar, and Benjamin Recht · 2019
Later among the works it cites.
Finite sample analysis of stochastic system identification
Anastasios Tsiamis and George J Pappas · 2019
Later among the works it cites.
A system level approach to controller synthesis
Yuh-Shyang Wang, Nikolai Matni, and John C Doyle · 2019
Later among the works it cites.
Naive exploration is optimal for online lqr
Max Simchowitz and Dylan J. Foster · 2020
Closest in time.