Fetching the paper…
Reading the bibliography…
We introduce algorithms for learning nonlinear dynamical systems of the form $x_{t+1}=\sigma(\Theta^{\star}x_t)+\varepsilon_t$, where $\Theta^{\star}$ is a weight matrix, $\sigma$ is a nonlinear link function, and $\varepsilon_t$ is a mean-zero noise process.
System identification—a survey
Karl Johan Åström and Peter Eykhoff · 1971
Earlier work this paper cites.
Power distribution inequalities in optimization and robustness of uncertain systems
Alexandre Megretski and Sergei Treil · 1993
Earlier work this paper cites.
Robust stability with time-varying structured uncertainty
Jeff S Shamma · 1994
Earlier work this paper cites.
Robust performance against time-varying structured perturbations
K Poola and Ashok Tikku · 1995
Earlier work this paper cites.
System Identification: Theory for the User
Lennart Ljung · 1998
Earlier work this paper cites.
Empirical Processes in M-Estimation
Sara A. van de Geer · 2000
Earlier work this paper cites.
Adaptive estimation in autoregression or-mixing regression via model selection
Yannick Baraud, Fabienne Comte, and Gabrielle Viennet · 2001
Earlier work this paper cites.
Finite sample properties of system identification methods
Marco C Campi and Erik Weyer · 2002
Earlier work this paper cites.
On Hoeffding’s inequality for dependent random variables
Sara A. van de Geer · 2002
Earlier work this paper cites.
Local rademacher complexities
Peter L Bartlett, Olivier Bousquet, and Shahar Mendelson · 2005
Earlier work this paper cites.
Local rademacher complexities and oracle inequalities in risk minimization
Vladimir Koltchinskii · 2006
Earlier work this paper cites.
A learning theory approach to system identification and stochastic adaptive control
Mathukumalli Vidyasagar and Rajeeva L Karandikar · 2006
Earlier work this paper cites.
Self-normalized processes: Limit theory and Statistical Applications
Victor H de la Peña, Tze Leung Lai, and Qi-Man Shao · 2008
Earlier work this paper cites.
Introduction to Nonparametric Estimation
Alexandre B Tsybakov · 2008
Earlier work this paper cites.
Kernel methods for deep learning
Youngmin Cho and Lawrence K Saul · 2009
Earlier work this paper cites.
The isotron algorithm: High-dimensional isotonic regression
Adam Tauman Kalai and Ravi Sastry · 2009
Earlier work this paper cites.
Regret bounds for the adaptive control of linear quadratic systems
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2011
Cited alongside, same era.
Efficient learning of generalized linear and single index models with isotonic regression
Sham M Kakade, Varun Kanade, Ohad Shamir, and Adam Kalai · 2011
Cited alongside, same era.
Distributed control of positive systems
Anders Rantzer · 2011
Cited alongside, same era.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2012
Cited alongside, same era.
Active and passive learning of linear separators under log-concave distributions
Maria-Florina Balcan and Phil Long · 2013
Cited alongside, same era.
Taming the monster: A fast and simple algorithm for contextual bandits
Alekh Agarwal, Daniel Hsu, Satyen Kale, John Langford, Lihong Li, and Robert Schapire · 2014
Learning high-dimensional generalized linear autoregressive models
Eric C Hall, Garvesh Raskutti, and Rebecca M Willett · 2018
Later among the works it cites.
Spectral filtering for general linear dynamical systems
Elad Hazan, Holden Lee, Karan Singh, Cyril Zhang, and Yi Zhang · 2018
Later among the works it cites.
Regularization and the small-ball method I: sparse recovery
Guillaume Lecué and Shahar Mendelson · 2018
Later among the works it cites.
The landscape of empirical risk for nonconvex losses
Song Mei, Yu Bai, and Andrea Montanari · 2018
Later among the works it cites.
Learning without mixing: Towards a sharp analysis of linear system identification
Max Simchowitz, Horia Mania, Stephen Tu, Michael I Jordan, and Benjamin Recht · 2018
Later among the works it cites.
Least-squares temporal difference learning for the linear quadratic regulator
Stephen Tu and Benjamin Recht · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Online nonparametric regression
Alexander Rakhlin and Karthik Sridharan · 2014
Cited alongside, same era.
Online learning via sequential complexities
Alexander Rakhlin, Karthik Sridharan, and Ambuj Tewari · 2014
Cited alongside, same era.
High-Dimensional Statistics: A Non-Asymptotic Viewpoint
Martin J. Wainwright · 2014
Cited alongside, same era.
Statistical learning with sparsity: the lasso and generalizations
Trevor Hastie, Robert Tibshirani, and Martin Wainwright · 2015
Cited alongside, same era.
Learning with square loss: Localization through offset rademacher complexity
Tengyuan Liang, Alexander Rakhlin, and Karthik Sridharan · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Later among the works it cites.
Convex programming for estimation in nonlinear recurrent models
Sohail Bahmani and Justin Romberg · 2019
Later among the works it cites.
On the sample complexity of the linear quadratic regulator
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2019
Later among the works it cites.
Certainty equivalent control of LQR is efficient
Horia Mania, Stephen Tu, and Benjamin Recht · 2019
Later among the works it cites.
Stochastic gradient descent learns state equations with nonlinear activations
Samet Oymak · 2019
Later among the works it cites.
Near optimal finite time identification of arbitrary linear dynamical systems
Tuhin Sarkar and Alexander Rakhlin · 2019
Later among the works it cites.
Finite-time system identification for partially observed LTI systems of unknown order
Tuhin Sarkar, Alexander Rakhlin, and Munther A. Dahleh · 2019
Later among the works it cites.
Learning linear dynamical systems with semi-parametric least squares
Max Simchowitz, Ross Boczar, and Benjamin Recht · 2019
Later among the works it cites.
Non-asymptotic and accurate learning of nonlinear dynamical systems
Yahya Sattar and Samet Oymak · 2020
Closest in time.
Naive exploration is optimal for online lqr
Max Simchowitz and Dylan J Foster · 2020
Closest in time.