Fetching the paper…
Reading the bibliography…
The linear quadratic regulator (LQR) problem has reemerged as an important theoretical benchmark for reinforcement learning-based control of complex dynamical systems with continuous state and action spaces.
Gradient methods for the minimisation of functionals
B.T. Polyak · 1963
Earlier work this paper cites.
Optimal stationary control of a linear system with state-dependent noise
W Murray Wonham · 1967
Earlier work this paper cites.
A survey of stability of stochastic systems
Frank Kozin · 1969
Earlier work this paper cites.
An iterative technique for the computation of the steady state gains for the discrete optimal regulator
G. Hewer · 1971
Earlier work this paper cites.
Feedback stabilizability for stochastic systems with state and control dependent noise
Jacques L Willems and Jan C Willems · 1976
Earlier work this paper cites.
The uncertainty threshold principle: Some fundamental limitations of optimal decision making under dynamic uncertainty
Michael Athans, Richard Ku, and Stanley Gershwin · 1977
Earlier work this paper cites.
Further results on the uncertainty threshold principle
Richard Ku and Michael Athans · 1977
Earlier work this paper cites.
An eigensystem realization algorithm for modal parameter identification and model reduction
Jer-Nan Juang and Richard S. Pappa · 1985
Earlier work this paper cites.
Robust static and dynamic output-feedback stabilization: Deterministic and stochastic perspectives
Dennis Bernstein · 1987
Earlier work this paper cites.
Linear Matrix Inequalities in System and Control Theory
S. Boyd, L. El Ghaoui, E. Feron, and V. Balakrishnan · 1994
Earlier work this paper cites.
Adaptive linear quadratic control using policy iteration
Steven J Bradtke, B Erik Ydstie, and Andrew G Barto · 1994
Earlier work this paper cites.
Dynamic programming and optimal control
Dimitri P Bertsekas · 1995
Earlier work this paper cites.
State-feedback control of systems with multiplicative noise via linear matrix inequalities
Laurent El Ghaoui · 1995
Earlier work this paper cites.
PAC adaptive control of linear systems
Claude-Nicolas Fiechter · 1997
Earlier work this paper cites.
Stochastic H ∞ {H}^{\infty}
Diederich Hinrichsen and Anthony J Pritchard · 1998
Earlier work this paper cites.
A natural policy gradient
Sham M Kakade · 2002
Earlier work this paper cites.
Properties of the solutions of rational matrix difference equations
G. Freiling and A. Hochhaus · 2003
Earlier work this paper cites.
Convex optimization
Stephen Boyd, Stephen P Boyd, and Lieven Vandenberghe · 2004
Earlier work this paper cites.
Rational matrix equations in stochastic control
Tobias Damm · 2004
Earlier work this paper cites.
Robust control design for linear systems via multiplicative noise
Benjamin Gravell, Peyman Mohajerin Esfahani, and Tyler Summers · 2004
Cited alongside, same era.
Online convex optimization in the bandit setting: Gradient descent without a gradient
Abraham D. Flaxman, Adam Tauman Kalai, Adam Tauman Kalai, and H. Brendan McMahan · 2005
Cited alongside, same era.
Estimation and control of systems with multiplicative noise via linear matrix inequalities
Weiwei Li, Emanuel Todorov, and Robert E Skelton · 2005
Cited alongside, same era.
Power-electronic systems for the grid integration of renewable energy sources: A survey
Juan Manuel Carrasco, Leopoldo García Franquelo, Jan T Bialasiewicz, Eduardo Galván, Ramón Carlos Portillo Guisado, María de los Ángeles Martín Prats, José Ignacio León, and Narciso Moreno-Alfonso · 2006
Cited alongside, same era.
Policy gradient methods for robotics
J. Peters and S. Schaal · 2006
Cited alongside, same era.
Data-driven control: A behavioral approach
T.M. Maupong and P. Rapisarda · 2017
Later among the works it cites.
Domain randomization for transferring deep neural networks from simulation to the real world
Josh Tobin, Rachel Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel · 2017
Later among the works it cites.
Least-squares temporal difference learning for the linear quadratic regulator
Stephen Tu and Benjamin Recht · 2017
Later among the works it cites.
Improved regret bounds for thompson sampling in linear quadratic control problems
Marc Abeille and Alessandro Lazaric · 2018
Later among the works it cites.
An input-output approach to structured stochastic uncertainty
Bassam Bamieh and Maurice Filo · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Special issue on technology of networked control systems
Panos Antsaklis and John Baillieul · 2007
Cited alongside, same era.
A survey of recent results in networked control systems
Joao P Hespanha, Payam Naghshtabrizi, and Yonggang Xu · 2007
Cited alongside, same era.
Properties of Stein (Lyapunov) iterations for solving a general Riccati equation
Ivan Ganchev Ivanov · 2007
Cited alongside, same era.
Stochastic tools in turbulence
John L Lumley · 2007
Cited alongside, same era.
Regret bounds for the adaptive control of linear quadratic systems
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2011
Cited alongside, same era.
Lyapunov equations, energy functionals, and model order reduction of bilinear and stochastic systems
Peter Benner and Tobias Damm · 2011
Cited alongside, same era.
Reinforcement learning and feedback control: Using natural decision methods to design optimal adaptive controllers
Frank L Lewis, Draguna Vrabie, and Kyriakos G Vamvoudakis · 2012
Cited alongside, same era.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham Kakade, and Mehran Mesbahi · 2018
Later among the works it cites.
Foundations and challenges of low-inertia systems
Federico Milano, Florian Dörfler, Gabriela Hug, David J Hill, and Gregor Verbič · 2018
Later among the works it cites.
A tour of reinforcement learning: The view from continuous control
Benjamin Recht · 2018
Later among the works it cites.
A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy Lillicrap, Karen Simonyan, and Demis Hassabis · 2018
Later among the works it cites.
Learning convex bounds for linear quadratic control policy synthesis
Jack Umenberger and Thomas B Schön · 2018
Later among the works it cites.
Recovering robustness in model-free reinforcement learning
Harish K Venkataraman and Peter J Seiler · 2018
Later among the works it cites.
Data-driven minimum-energy controls for linear systems
G. Baggio, V. Katewa, and F. Pasqualetti · 2019
Closest in time.
LQR through the lens of first order methods: Discrete-time case
Jingjing Bu, Afshin Mesbahi, Maryam Fazel, and Mehran Mesbahi · 2019
Closest in time.
Data-driven lqr control design
G. R. Gonçalves da Silva, A. S. Bazanella, C. Lorenzini, and L. Campestrini · 2019
Closest in time.
Certainty equivalent control of LQR is efficient
Horia Mania, Stephen Tu, and Benjamin Recht · 2019
Closest in time.
On persistency of excitation and formulas for data-driven control
C. D. Persis and P. Tesi · 2019
Closest in time.
Safe learning-based control of stochastic jump linear systems: a distributionally robust approach
M. Schuurmans, P. Sopasakis, and P. Patrinos · 2019
Closest in time.
Convergence guarantees of policy optimization methods for markovian jump linear systems
Joao Paulo Jansch-Porto, Bin Hu, and Geir Dullerud · 2020
Closest in time.
Linear system identification under multiplicative noise from multiple trajectory data
Yu Xing, Ben Gravell, Xingkang He, Karl Henrik Johansson, and Tyler Summers · 2020
Closest in time.