Fetching the paper…
Reading the bibliography…
We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem.
The implementation shortfall: Paper versus reality
Andre F Perold · 1988
Earlier work this paper cites.
PAC adaptive control of linear systems
Claude-Nicolas Fiechter · 1997
Earlier work this paper cites.
Optimal execution of portfolio transactions
Robert Almgren and Neil Chriss · 2001
Earlier work this paper cites.
Optimal execution with nonlinear impact functions and trading-enhanced risk
Robert Almgren · 2003
Earlier work this paper cites.
Introductory lectures on convex optimization: A basic course
Yurii Nesterov · 2003
Earlier work this paper cites.
Iterative linear quadratic regulator design for nonlinear biological movement systems
Weiwei Li and Emanuel Todorov · 2004
Earlier work this paper cites.
Direct estimation of equity market impact
Robert Almgren, Chee Thum, Emmanuel Hauptmann, and Hong Li · 2005
Earlier work this paper cites.
Dynamic Programming And Optimal Control
Dimitri Bertsekas · 2005
Earlier work this paper cites.
Online convex optimization in the bandit setting: Gradient descent without a gradient
Abraham D. Flaxman, Adam Tauman Kalai, and H. Brendan McMahan · 2005
Earlier work this paper cites.
Reinforcement learning for optimized trade execution
Yuriy Nevmyvaka, Yi Feng, and Michael Kearns · 2006
Earlier work this paper cites.
Optimal Control: Linear Quadratic Methods
Brian D. O. Anderson and John B Moore · 2007
Earlier work this paper cites.
Optimal execution strategies in limit order books with general shape functions
Aurélien Alfonsi, Antje Fruth, and Alexander Schied · 2010
Earlier work this paper cites.
Regret bounds for the adaptive control of linear quadratic systems
Yasin Abbasi-Yadkori and Csaba Szepesvári · 2011
Earlier work this paper cites.
Optimal trade execution under geometric Brownian motion in the Almgren and Chriss framework
Jim Gatheral and Alexander Schied · 2011
Earlier work this paper cites.
Recovering low-rank matrices from few coefficients in any basis
David Gross · 2011
Earlier work this paper cites.
Stochastic MPC for real-time market-based optimal power dispatch
Panagiotis Patrinos, Sergio Trimboli, and Alberto Bemporad · 2011
Earlier work this paper cites.
Efficient reinforcement learning for high dimensional linear quadratic systems
Morteza Ibrahimi, Adel Javanmard, and Benjamin V Roy · 2012
Cited alongside, same era.
Adaptive control
Karl J Åström and Björn Wittenmark · 2013
Cited alongside, same era.
The price impact of order book events
Rama Cont, Arseniy Kukanov, and Sasha Stoikov · 2014
Cited alongside, same era.
A reinforcement learning extension to the Almgren-Chriss framework for optimal trade execution
Dieter Hendricks and Diane Wilcox · 2014
Cited alongside, same era.
A note on the Hanson-Wright inequality for random vectors with dependencies
Radoslaw Adamczak · 2015
Cited alongside, same era.
LQG for portfolio optimization
Marc Abeille, Alessandro Lazaric, Xavier Brokmann, et al · 2016
Cited alongside, same era.
Linear-quadratic mean-field reinforcement learning: convergence of policy gradient methods
René Carmona, Mathieu Laurière, and Zongjun Tan · 2019
Later among the works it cites.
On the sample complexity of the linear quadratic regulator
Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, and Stephen Tu · 2019
Later among the works it cites.
Benjamin Gravell, Peyman Mohajerin Esfahani, and Tyler Summers · 2019
Later among the works it cites.
Derivative-free methods for policy optimization: guarantees for linear quadratic systems
Dhruv Malik, Ashwin Pananjady, Kush Bhatia, Koulik Khamaru, Peter Bartlett, and Martin Wainwright · 2019
Later among the works it cites.
A tour of reinforcement learning: The view from continuous control
Benjamin Recht · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Thompson sampling for linear-quadratic control problems
Marc Abeille and Alessandro Lazaric · 2017
Cited alongside, same era.
Control of unknown linear systems with Thompson sampling
Yi Ouyang, Mukul Gagrani, and Rahul Jain · 2017
Cited alongside, same era.
Differential game-based load frequency control for power networks and its integration with electricity market mechanisms
Yasuaki Wasa, Kengo Sakata, Kenji Hirata, and Kenko Uchida · 2017
Cited alongside, same era.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham M Kakade, and Mehran Mesbahi · 2018
Cited alongside, same era.
Double deep Q-learning for optimal execution
Brian Ning, Franco Ho Ting Ling, and Sebastian Jaimungal · 2018
Cited alongside, same era.
Least-squares temporal difference learning for the linear quadratic regulator
Stephen Tu and Benjamin Recht · 2018
Cited alongside, same era.
The gap between model-based and model-free methods on the linear quadratic regulator: An asymptotic viewpoint
Stephen Tu and Benjamin Recht · 2019
Later among the works it cites.
On the global convergence of actor-critic: a case for linear quadratic regulator with ergodic cost
Zhuoran Yang, Yongxin Chen, Mingyi Hong, and Zhaoran Wang · 2019
Later among the works it cites.
Policy optimization provably converges to Nash equilibria in zero-sum linear quadratic games
Kaiqing Zhang, Zhuoran Yang, and Tamer Basar · 2019
Later among the works it cites.
Policy gradient-based algorithms for continuous-time linear quadratic control
Jingjing Bu, Afshin Mesbahi, and Mehran Mesbahi · 2020
Closest in time.
Reinforcement learning in economics and finance
Arthur Charpentier, Romuald Elie, and Carl Remlinger · 2020
Closest in time.
Optimism-based adaptive regulation of linear-quadratic systems
Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, and George Michailidis · 2020
Closest in time.
Efficient learning of distributed linear-quadratic control policies
Salar Fattahi, Nikolai Matni, and Somayeh Sojoudi · 2020
Closest in time.
Entropy regularization for mean field games with learning
Xin Guo, Renyuan Xu, and Thaleia Zariphopoulou · 2020
Closest in time.
On the analysis of model-free methods for the linear quadratic regulator
Zeyu Jin, Johann Michael Schmitt, and Zaiwen Wen · 2020
Closest in time.
Learning a functional control for high-frequency finance
Laura Leal, Mathieu Laurière, and Charles-Albert Lehalle · 2020
Closest in time.
Deep reinforcement learning for trading
Zihao Zhang, Stefan Zohren, and Stephen Roberts · 2020
Closest in time.