Fetching the paper…
Reading the bibliography…
We consider the problem of learning stabilizable systems governed by nonlinear state equation $h_{t+1}=\phi(h_t,u_t;\theta)+w_t$.
Effective construction of linear state-variable models from input/output functions
BL Ho and Rudolf E Kálmán · 1966
Earlier work this paper cites.
System identification—a survey
Karl Johan Åström and Peter Eykhoff · 1971
Earlier work this paper cites.
Non-linear system identification using neural networks
Sheng Chen, SA Billings, and PM Grant · 1990
Earlier work this paper cites.
Rates of convergence for empirical processes of stationary mixing sequences
Bin Yu · 1994
Earlier work this paper cites.
PID controllers: theory, design, and tuning
Karl Johan Åström and Tore Hägglund · 1995
Earlier work this paper cites.
An introduction to the kalman filter
Greg Welch, Gary Bishop, et al · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
System analysis via integral quadratic constraints
Alexandre Megretski and Anders Rantzer · 1997
Earlier work this paper cites.
System identification
Lennart Ljung · 1999
Earlier work this paper cites.
Empirical Processes in M-estimation
Sara A Geer and Sara van de Geer · 2000
Earlier work this paper cites.
The concentration of measure phenomenon
Michel Ledoux · 2001
Earlier work this paper cites.
Introducing sostools: A general purpose sum of squares programming solver
Stephen Prajna, Antonis Papachristodoulou, and Pablo A Parrilo · 2002
Earlier work this paper cites.
A tutorial on sum of squares techniques for systems analysis
Antonis Papachristodoulou and Stephen Prajna · 2005
Earlier work this paper cites.
Stability bounds for non-iid processes
Mehryar Mohri and Afshin Rostamizadeh · 2008
Earlier work this paper cites.
Rademacher complexity bounds for non-iid processes
Mehryar Mohri and Afshin Rostamizadeh · 2009
Earlier work this paper cites.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur · 2010
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2010
Earlier work this paper cites.
Efficient learning of generalized linear and single index models with isotonic regression
Sham M Kakade, Varun Kanade, Ohad Shamir, and Adam Kalai · 2011
Earlier work this paper cites.
System identification: a frequency domain approach
Rik Pintelon and Johan Schoukens · 2012
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton · 2013
Earlier work this paper cites.
Accelerating a recurrent neural network to finite-time convergence for solving time-varying sylvester equation by using a sign-bi-power activation function
Shuai Li, Sanfeng Chen, and Bo Liu · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Long short-term memory recurrent neural network architectures for large scale acoustic modeling
Haşim Sak, Andrew Senior, and Françoise Beaufays · 2014
Cited alongside, same era.
Sample complexity of episodic fixed-horizon reinforcement learning
Christoph Dann and Emma Brunskill · 2015
Cited alongside, same era.
Regret analysis for adaptive linear-quadratic policies
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2017
Cited alongside, same era.
Learning linear dynamical systems via spectral filtering
Elad Hazan, Karan Singh, and Cyril Zhang · 2017
Cited alongside, same era.
Generalization bounds for non-stationary mixing processes
Vitaly Kuznetsov and Mehryar Mohri · 2017
Cited alongside, same era.
Convergence analysis of two-layer neural networks with relu activation
Fitting relus via sgd and quantized sgd
Seyed Mohammadreza Mousavi Kalan, Mahdi Soltanolkotabi, and A Salman Avestimehr · 2019
Later among the works it cites.
Finite-time analysis of approximate policy iteration for the linear quadratic regulator
Karl Krauth, Stephen Tu, and Benjamin Recht · 2019
Later among the works it cites.
Derivative-free methods for policy optimization: Guarantees for linear quadratic systems
Dhruv Malik, Ashwin Pananjady, Kush Bhatia, Koulik Khamaru, Peter Bartlett, and Martin Wainwright · 2019
Later among the works it cites.
Certainty equivalent control of lqr is efficient
Horia Mania, Stephen Tu, and Benjamin Recht · 2019
Later among the works it cites.
Stochastic gradient descent learns state equations with nonlinear activations
Samet Oymak · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yuanzhi Li and Yang Yuan · 2017
Cited alongside, same era.
Nonparametric risk bounds for time-series forecasting
Daniel J McDonald, Cosma Rohilla Shalizi, and Mark Schervish · 2017
Cited alongside, same era.
On the convergence rate of training recurrent neural networks
Zeyuan Allen-Zhu, Yuanzhi Li, and Zhao Song · 2018
Cited alongside, same era.
Regret bounds for robust adaptive control of the linear quadratic regulator
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2018
Cited alongside, same era.
Finite time identification in unstable linear systems
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2018
Cited alongside, same era.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham Kakade, and Mehran Mesbahi · 2018
Cited alongside, same era.
Uniform convergence of gradients for non-convex learning and optimization
Dylan J Foster, Ayush Sekhari, and Karthik Sridharan · 2018
Cited alongside, same era.
Later among the works it cites.
Non-asymptotic identification of lti systems from a single trajectory
Samet Oymak and Necmiye Ozay · 2019
Later among the works it cites.
A tour of reinforcement learning: The view from continuous control
Benjamin Recht · 2019
Later among the works it cites.
Near optimal finite time identification of arbitrary linear dynamical systems
Tuhin Sarkar and Alexander Rakhlin · 2019
Later among the works it cites.
Data driven estimation of stochastic switched linear systems of unknown order
Tuhin Sarkar, Alexander Rakhlin, and Munther A Dahleh · 2019
Later among the works it cites.
Finite-time system identification for partially observed lti systems of unknown order
Tuhin Sarkar, Alexander Rakhlin, and Munther A Dahleh · 2019
Later among the works it cites.
A simple framework for learning stabilizable systems
Yahya Sattar and Samet Oymak · 2019
Later among the works it cites.
Learning linear dynamical systems with semi-parametric least squares
Max Simchowitz, Ross Boczar, and Benjamin Recht · 2019
Later among the works it cites.
Finite sample analysis of stochastic system identification
Anastasios Tsiamis and George J Pappas · 2019
Later among the works it cites.
Finite-sample analysis for sarsa and q-learning with linear function approximation
Shaofeng Zou, Tengyu Xu, and Yingbin Liang · 2019
Later among the works it cites.
Learning nonlinear dynamical systems from a single trajectory
Dylan J Foster, Alexander Rakhlin, and Tuhin Sarkar · 2020
Closest in time.
Convex nonparametric formulation for identification of gradient flows
Mohammad Khosravi and Roy S Smith · 2020
Closest in time.
Nonlinear system identification with prior knowledge on the region of attraction
Mohammad Khosravi and Roy S Smith · 2020
Closest in time.
Active learning for nonlinear system identification with guarantees
Horia Mania, Michael I Jordan, and Benjamin Recht · 2020
Closest in time.
Learning stabilizable nonlinear dynamics with contraction-based regularization
Sumeet Singh, Spencer M Richards, Vikas Sindhwani, Jean-Jacques E Slotine, and Marco Pavone · 2020
Closest in time.
Sample complexity of kalman filtering for unknown systems
Anastasios Tsiamis, Nikolai Matni, and George Pappas · 2020
Closest in time.
Active learning for identification of linear dynamical systems
Andrew Wagenmaker and Kevin Jamieson · 2020
Closest in time.