Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershevlvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis · 2016
Later among the works it cites.
A System Level Approach to Controller Synthesis
Original
Y.-S. Wang, N. Matni, and J. C. Doyle · 2016
Later among the works it cites.
Thompson Sampling for Linear-Quadratic Control Problems
M. Abeille and A. Lazaric · 2017
Closest in time.
Structured State Space Realizations for SLS Distributed Controllers
J. Anderson and N. Matni · 2017
Closest in time.
Predictive Control for Linear and Hybrid Systems
F. Borrelli, A. Bemporad, and M. Morari · 2017
Closest in time.
Learning Linear Dynamical Systems via Spectral Filtering
E. Hazan, K. Singh, and C. Zhang · 2017
Closest in time.
Contextual Decision Processes with Low Bellman Rank are PAC-Learnable
N. Jiang, A. Krishnamurthy, A. Agarwal, J. Langford, and R. E. Schapire · 2017
Closest in time.
Occupy the Cloud: Distributed Computing for the 99%
E. Jonas, Q. Pu, S. Venkataraman, I. Stoica, and B. Recht · 2017
Closest in time.
Generalization bounds for non-stationary mixing processes
V. Kuznetsov and M. Mohri · 2017
Closest in time.
Scalable system level synthesis for virtually localizable systems
N. Matni, Y.-S. Wang, and J. Anderson · 2017
Closest in time.
Nonparametric Risk Bounds for Time-Series Forecasting
D. J. McDonald, C. R. Shalizi, and M. Schervish · 2017
Closest in time.
Control of Unknown Linear Systems with Thompson Sampling
Y. Ouyang, M. Gagrani, and R. Jain · 2017
Closest in time.
Stability Analysis of Discrete-Time Infinite-Horizon Optimal Control With Discounted Cost
R. Postoyan, L. Buşoniu, D. Nešić, and J. Daafouz · 2017
Closest in time.
A Tutorial on Thompson Sampling
Original
D. Russo, B. V. Roy, A. Kazerouni, and I. Osband · 2017
Closest in time.
Non-Asymptotic Analysis of Robust Control from Coarse-Grained Identification
Original
S. Tu, R. Boczar, A. Packard, and B. Recht · 2017
Closest in time.
Improved Regret Bounds for Thompson Sampling in Linear Quadratic Control Problems
M. Abeille and A. Lazaric · 2018
Closest in time.
Safely Learning to Control the Constrained Linear Quadratic Regulator
Original
S. Dean, S. Tu, N. Matni, and B. Recht · 2018
Closest in time.
Global Convergence of Policy Gradient Methods for the Linear Quadratic Regulator
M. Fazel, R. Ge, S. M. Kakade, and M. Mesbahi · 2018
Closest in time.
Spectral Filtering for General Linear Dynamical Systems
Original
E. Hazan, H. Lee, K. Singh, C. Zhang, and Y. Zhang · 2018
Closest in time.
Learning Without Mixing: Towards A Sharp Analysis of Linear System Identification
M. Simchowitz, H. Mania, S. Tu, M. I. Jordan, and B. Recht · 2018
Closest in time.
High-dimensional Statistics: A Non-Asymptotic Viewpoint
M. J. Wainwright · 2019
Closest in time.