Fetching the paper…
Reading the bibliography…
Learning how to effectively control unknown dynamical systems is crucial for intelligent autonomous systems.
On tail probabilities for martingales
David A Freedman · 1975
Earlier work this paper cites.
Adaptive estimation and identification for discrete systems with markov jump parameters
Jitendra Tugnait · 1982
Earlier work this paper cites.
Discrete-time markovian-jump linear quadratic optimal control
Howard J Chizeck, Alan S Willsky, and D Castanon · 1986
Earlier work this paper cites.
A probabilistic approach to dynamic power system security
KA Loparo and F Abdel-Malek · 1990
Earlier work this paper cites.
On the adaptive control of jump parameter systems via nonlinear filtering
Peter E Caines and Ji-Feng Zhang · 1995
Earlier work this paper cites.
Adaptive linear quadratic gaussian control: the cost-biased approach revisited
Marco C Campi and PR Kumar · 1998
Earlier work this paper cites.
System identification
Lennart Ljung · 1999
Earlier work this paper cites.
Necessary and sufficient conditions for adaptive stablizability of jump linear systems
F Xue and L Guo · 2001
Earlier work this paper cites.
Stochastic optimal control of jumping Markov parameter processes with applications to finance
DO Cajueiro · 2002
Earlier work this paper cites.
Switching in systems and control
Daniel Liberzon · 2003
Earlier work this paper cites.
Combining stochastic and greedy search in hybrid estimation
Lars Blackmore, Stanislav Funiak, and Brian C Williams · 2005
Earlier work this paper cites.
Decentralized control of power systems via robust control of uncertain markov jump parameter systems
Valery Ugrinovskii* and Hemanshu R Pota · 2005
Earlier work this paper cites.
Discrete-time Markov jump linear systems
Oswaldo Luiz Valle Costa, Marcelo Dutra Fragoso, and Ricardo Paulino Marques · 2006
Earlier work this paper cites.
Uniform stabilization of discrete-time switched and markovian jump linear systems
Ji-Woong Lee and Geir E. Dullerud · 2006
Earlier work this paper cites.
Stability bounds for non-iid processes
Mehryar Mohri and Afshin Rostamizadeh · 2008
Earlier work this paper cites.
Optimal monetary policy under uncertainty: a markov jump-linear-quadratic approach
Lars EO Svensson, Noah Williams, et al · 2008
Earlier work this paper cites.
Bayesian nonparametric methods for learning markov switching processes
Emily B Fox, Erik B Sudderth, Michael I Jordan, and Alan S Willsky · 2010
Earlier work this paper cites.
Regret bounds for the adaptive control of linear quadratic systems
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2011
Earlier work this paper cites.
A sparsification approach to set membership identification of switched affine systems
Necmiye Ozay, Mario Sznaier, Constantino M Lagoa, and Octavia I Camps · 2011
Earlier work this paper cites.
A tail inequality for quadratic forms of subgaussian random vectors
Daniel Hsu, Sham Kakade, Tong Zhang, et al · 2012
Earlier work this paper cites.
Efficient reinforcement learning for high dimensional linear quadratic systems
Morteza Ibrahimi, Adel Javanmard, and Benjamin Van Roy · 2012
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2012
Earlier work this paper cites.
Stochastic processes: theory for applications
Robert G Gallager · 2013
Earlier work this paper cites.
Online learning and optimization of markov jump affine models
Sevi Baltaoglu, Lang Tong, and Qing Zhao · 2016
Cited alongside, same era.
Generalization bounds for non-stationary mixing processes
Vitaly Kuznetsov and Mehryar Mohri · 2017
Cited alongside, same era.
Markov chains and mixing times
David A Levin and Yuval Peres · 2017
Cited alongside, same era.
Non-asymptotic analysis of robust control from coarse-grained identification
Stephen Tu, Ross Boczar, Andrew Packard, and Benjamin Recht · 2017
Cited alongside, same era.
Improved regret bounds for thompson sampling in linear quadratic control problems
Marc Abeille and Alessandro Lazaric · 2018
Cited alongside, same era.
Finite sample analysis of stochastic system identification
Anastasios Tsiamis and George J Pappas · 2019
Later among the works it cites.
Spectral state compression of markov processes
Anru Zhang and Mengdi Wang · 2019
Later among the works it cites.
Efficient optimistic exploration in linear-quadratic regulators via lagrangian relaxation
Marc Abeille and Alessandro Lazaric · 2020
Later among the works it cites.
Logarithmic regret for learning linear quadratic regulators efficiently
Asaf Cassel, Alon Cohen, and Tomer Koren · 2020
Later among the works it cites.
On adaptive linear–quadratic regulators
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2020
Later among the works it cites.
Optimism-based adaptive regulation of linear-quadratic systems
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2018
Cited alongside, same era.
Finite-time adaptive stabilization of linear systems
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2018
Cited alongside, same era.
Finite time identification in unstable linear systems
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2018
Cited alongside, same era.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham Kakade, and Mehran Mesbahi · 2018
Cited alongside, same era.
Hybrid system identification: Theory and algorithms for learning switching models, vol. 478
F Lauer and G Bloch · 2018
Cited alongside, same era.
Learning without mixing: Towards a sharp analysis of linear system identification
Max Simchowitz, Horia Mania, Stephen Tu, Michael I Jordan, and Benjamin Recht · 2018
Cited alongside, same era.
Model-free linear quadratic control via reduction to expert prediction
Yasin Abbasi-Yadkori, Nevena Lazic, and Csaba Szepesvári · 2019
Cited alongside, same era.
Later among the works it cites.
The nonstochastic control problem
Elad Hazan, Sham Kakade, and Karan Singh · 2020
Later among the works it cites.
Statistical consistency of set-membership estimator for linear systems
Pedro Hespanhol and Anil Aswani · 2020
Later among the works it cites.
Policy learning of mdps with mixed continuous/discrete variables: A case study on model-free control of markovian jump systems
Joao Paulo Jansch-Porto, Bin Hu, and Geir Dullerud · 2020
Later among the works it cites.
Finite-time identification of stable linear systems optimality of the least-squares estimator
Yassir Jedra and Alexandre Proutiere · 2020
Later among the works it cites.
Explore more and improve regret in linear quadratic regulators
Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, and Anima Anandkumar · 2020
Later among the works it cites.
Logarithmic regret bound in partially observable linear dynamical systems
Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, and Anima Anandkumar · 2020
Later among the works it cites.
Computing stabilizing linear controllers via policy iteration
Andrew Lamperski · 2020
Later among the works it cites.
Non-asymptotic closed-loop system identification using autoregressive processes and hankel model reduction
Bruce Lee and Andrew Lamperski · 2020
Later among the works it cites.
On the linear convergence of random search for discrete-time lqr
Hesameddin Mohammadi, Mahdi Soltanolkotabi, and Mihailo R Jovanović · 2020
Later among the works it cites.
Non-asymptotic and accurate learning of nonlinear dynamical systems
Yahya Sattar and Samet Oymak · 2020
Later among the works it cites.
Naive exploration is optimal for online lqr
Max Simchowitz and Dylan Foster · 2020
Later among the works it cites.
Policy optimization for ℋ 2 \mathcal{H}_{2} linear control with ℋ ∞ \mathcal{H}_{\infty} robustness guarantee: Implicit regularization and global convergence
Kaiqing Zhang, Bin Hu, and Tamer Basar · 2020
Later among the works it cites.
Black-box control for linear dynamical systems
Xinyi Chen and Elad Hazan · 2021
Closest in time.
Certainty equivalent quadratic control for markov jump systems
Zhe Du, Yahya Sattar, Davoud Ataee Tarzanagh, Laura Balzano, Samet Oymak, and Necmiye Ozay · 2021
Closest in time.
Data-driven control of markov jump systems: Sample complexity and regret bounds
Zhe Du, Yahya Sattar, Davoud Ataee Tarzanagh, Laura Balzano, Necmiye Ozay, and Samet Oymak · 2021
Closest in time.
Stability and identification of random asynchronous linear time-invariant systems
Sahin Lale, Oguzhan Teke, Babak Hassibi, and Anima Anandkumar · 2021
Closest in time.
Analysis of the optimization landscape of linear quadratic gaussian (lqg) control
Yang Zheng, Yujie Tang, and Na Li · 2021
Closest in time.