Fetching the paper…
Reading the bibliography…
Gradient-based methods have been widely used for system design and optimization in diverse application domains.
1903
Earlier work this paper cites.
1903
Earlier work this paper cites.
1904
Earlier work this paper cites.
1907
Earlier work this paper cites.
1907
Earlier work this paper cites.
Lamperski A. 2020. Computing stabilizing linear controllers via policy iteration
1907
Earlier work this paper cites.
1907
Earlier work this paper cites.
1910
Earlier work this paper cites.
1911
Earlier work this paper cites.
Draper CS, Li YT. 1951. Principles of optimalizing control systems and an application to the internal combustion engine
1951
Earlier work this paper cites.
Whitaker HP, Yamron J, Kezer A. 1958. Design of model-reference adaptive control systems for aircraft
1958
Earlier work this paper cites.
Kalman RE. 1960. Contributions to the theory of optimal control. Bol. soc. mat. mexicana
1960
Earlier work this paper cites.
Talkin A. 1961. Adaptive servo tracking. IRE Transactions on Automatic Control
1961
Earlier work this paper cites.
Kleinman D. 1968. On an iterative technique for Riccati equation computations. IEEE Transactions on Automatic Control
1968
Earlier work this paper cites.
Levine W, Athans M. 1970. On the determination of the optimal constant output feedback gains for linear multivariable systems. IEEE Transactions on Automatic control
1970
Earlier work this paper cites.
Hewer G. 1971. An iterative technique for the computation of the steady state gains for the discrete optimal regulator. IEEE Transactions on Automatic Control
1971
Earlier work this paper cites.
Willems J. 1972. Dissipative dynamical systems part i: General theory. Archive for Rational Mech. and Analysis
1972
Earlier work this paper cites.
Jacobson D. 1973. Optimal stochastic linear systems with exponential performance criteria and their relation to deterministic differential games. IEEE Transactions on Automatic Control
1973
Earlier work this paper cites.
Sandell N, Varaiya P, Athans M. 1975. A survey of decentralized control methods for large scale systems. Systems Eng. for Power, US Dept. of Commerce
1975
Earlier work this paper cites.
Goldstein A. 1977. Optimization of Lipschitz continuous functions. Mathematical Programming
1977
Earlier work this paper cites.
Tsitsiklis JN. 1984. Problems in decentralized decision making and computation. Tech. rep., Massachusetts Inst of Tech Cambridge Lab for Information and Decision Systems
1984
Earlier work this paper cites.
Makila P, Toivonen H. 1987. Computational methods for parametric LQ problems–A survey. IEEE Transactions on Automatic Control
1987
Earlier work this paper cites.
Mustafa D. 1989. Relations between maximum-entropy/ ℋ ∞ \mathcal{H}_{\infty} control and combined ℋ ∞ \mathcal{H}_{\infty} /LQG control. Systems & Control Letters
1989
Earlier work this paper cites.
Mustafa D, Bernstein DS. 1991. LQG cost bounds in discrete-time ℋ 2 / ℋ ∞ \mathcal{H}_{2}/\mathcal{H}_{\infty} control. Transactions of the Institute of Measurement and Control
1991
Earlier work this paper cites.
Boyd S, El Ghaoui L, Feron E, Balakrishnan V. 1994. Linear Matrix Inequalities in System and Control Theory
1994
Earlier work this paper cites.
Gahinet P, Apkarian P. 1994. A linear matrix inequality approach to H ∞ {H}_{\infty} control. International Journal of Robust and Nonlinear Control
1994
Earlier work this paper cites.
Peres PL, Geromel JC. 1994. An alternate numerical solution to the linear quadratic problem. IEEE Transactions on Automatic Control
1994
Earlier work this paper cites.
Bradtke SJ, Ydstie BE, Barto AG. 1994. Adaptive linear quadratic control using policy iteration
1994
Earlier work this paper cites.
Başar T, Bernhard P. 1995. H ∞ H^{\infty} -Optimal Control and Related Minimax Design Problems
1995
Earlier work this paper cites.
Zhou K, Doyle JC, Glover K. 1996. Robust and optimal control
1996
Earlier work this paper cites.
Rantzer A. 1996. On the Kalman-Yakubovich-Popov Lemma. Systems & Control letters
1996
Earlier work this paper cites.
Rotkowitz M, Lall S. 2005. A characterization of convex problems in decentralized control. IEEE Transactions on Automatic Control
1996
Earlier work this paper cites.
Rautert T, Sachs EW. 1997. Computational design of optimal output feedback controllers. SIAM Journal on Optimization
1997
Earlier work this paper cites.
Bertsekas DP. 1997. Nonlinear programming. Journal of the Operational Research Society
1997
Earlier work this paper cites.
Scherer C, Gahinet P, Chilali M. 1997. Multiobjective output-feedback control via LMI optimization. IEEE Transactions on Automatic Control
1997
Earlier work this paper cites.
Megretski A, Rantzer A. 1997. System analysis via integral quadratic constraints. IEEE Transactions on Automatic Control
1997
Earlier work this paper cites.
Hjalmarsson H, Gevers M, Gunnarsson S, Lequin O. 1998. Iterative feedback tuning: theory and applications. IEEE control systems magazine
1998
Earlier work this paper cites.
Başar T, Olsder G. 1999. Dynamic Noncooperative Game Theory
1999
Earlier work this paper cites.
Dullerud G, Paganini F. 1999. A Course in Robust Control Theory: A Convex Approach
1999
Earlier work this paper cites.
Sutton RS, McAllester DA, Singh SP, Mansour Y. 2000. Policy gradient methods for reinforcement learning with function approximation
2000
Earlier work this paper cites.
Konda VR, Tsitsiklis JN. 2000. Actor-critic algorithms
2000
Earlier work this paper cites.
Van der Schaft A. 2000. L 2 L_{2} -Gain and Passivity Techniques in Nonlinear Control
2000
Earlier work this paper cites.
Hjalmarsson H. 2002. Iterative feedback tuning—an overview. International journal of adaptive control and signal processing
2002
Earlier work this paper cites.
Kakade SM. 2002. A natural policy gradient
2002
Earlier work this paper cites.
Lagoudakis MG, Parr R. 2003. Least-squares policy iteration. The Journal of Machine Learning Research
2003
Earlier work this paper cites.
Boyd S, Vandenberghe L. 2004. Convex Optimization
2004
Earlier work this paper cites.
Scherer C, Wieland S. 2004. Linear matrix inequalities in control. Lecture notes for a course of the dutch institute of systems and control, Delft University of Technology
2004
Earlier work this paper cites.
Prajna S, Papachristodoulou A, Seiler P, Parrilo PA. 2004. SOSTOOLS: Sum of squares optimization toolbox for MATLAB
2004
Earlier work this paper cites.
Noll D, Apkarian P. 2005. Spectral bundle methods for non-convex maximum eigenvalue functions: second-order methods. Mathematical Programming
2005
Earlier work this paper cites.
Morimoto J, Doya K. 2005. Robust reinforcement learning. Neural computation
2005
Earlier work this paper cites.
Apkarian P, Noll D. 2006. Nonsmooth ℋ ∞ \mathcal{H}_{\infty} synthesis. IEEE Transactions on Automatic Control
2006
Cited alongside, same era.
Saeki M. 2006. Static output feedback design for ℋ ∞ \mathcal{H}_{\infty} control by descent method
2006
Cited alongside, same era.
Nesterov Y, Polyak BT. 2006. Cubic regularization of Newton method and its global performance. Mathematical Programming
2006
Cited alongside, same era.
2006
Cited alongside, same era.
2007
Cited alongside, same era.
Furieri L, Zheng Y, Kamgarpour M. 2020. Learning the globally optimal distributed LQ regulator
2020
Later among the works it cites.
Mohammadi H, Soltanolkotabi M, Jovanović MR. 2020. On the linear convergence of random search for discrete-time LQR. IEEE Control Systems Letters
2020
Later among the works it cites.
Zhang K, Hu B, Başar T. 2020. On the stability and convergence of robust adversarial reinforcement learning: A case study on linear quadratic systems. Advances in Neural Information Processing Systems
2020
Later among the works it cites.
Gravell B, Esfahani PM, Summers T. 2020. Learning optimal controllers for linear systems with multiplicative noise via policy gradient. IEEE Transactions on Automatic Control
2020
Later among the works it cites.
Bu J, Mesbahi M. 2020. Global convergence of policy gradient algorithms for indefinite least squares stationary optimal control. IEEE Control Systems Letters
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Al-Tamimi A, Lewis FL, Abu-Khalaf M. 2007. Model-free 𝒬 \mathcal{Q} -learning designs for linear discrete-time zero-sum games with application to H-infinity control. Automatica
2007
Cited alongside, same era.
2007
Cited alongside, same era.
Apkarian P, Noll D, Rondepierre A. 2008. Mixed ℋ 2 / ℋ ∞ \mathcal{H}_{2}/\mathcal{H}_{\infty} control via nonsmooth optimization. SIAM Journal on Control and Optimization
2008
Cited alongside, same era.
Gumussoy S, Henrion D, Millstone M, Overton ML. 2009. Multiobjective robust control with HIFOO 2.0. IFAC Proceedings Volumes
2009
Cited alongside, same era.
Mårtensson K, Rantzer A. 2009. Gradient methods for iterative distributed control synthesis
2009
Cited alongside, same era.
Arzelier D, Deaconu G, Gumussoy S, Henrion D. 2011. H2 for HIFOO
2011
Cited alongside, same era.
Bauschke HH, Combettes PL, et al. 2011. Convex analysis and monotone operator theory in Hilbert spaces
2011
Cited alongside, same era.
2020
Later among the works it cites.
Tang Y, Ren Z, Li N. 2020. Zeroth-order feedback optimization for cooperative multi-agent systems
2020
Later among the works it cites.
Burke JV, Curtis FE, Lewis AS, Overton ML, Simões LE. 2020. Gradient sampling methods for nonsmooth optimization. Numerical Nonsmooth Optimization
2020
Later among the works it cites.
Gravell B, Ganapathy K, Summers T. 2020. Policy iteration for linear quadratic games with stochastic parameters. IEEE Control Systems Letters
2020
Later among the works it cites.
Turchetta M, Krause A, Trimpe S. 2020. Robust model-free reinforcement learning with multi-objective Bayesian optimization
2020
Later among the works it cites.
Carmona R, Hamidouche K, Laurière M, Tan Z. 2020. Policy optimization for linear-quadratic zero-sum mean-field type games
2020
Later among the works it cites.
Simchowitz M, Foster D. 2020. Naive exploration is optimal for online LQR
2020
Later among the works it cites.
Simchowitz M, Singh K, Hazan E. 2020. Improper learning for non-stochastic control
2020
Later among the works it cites.
Palan M, Barratt S, McCauley A, Sadigh D, Sindhwani V, Boyd S. 2020. Fitting a linear control policy to demonstrations with a Kalman constraint
2020
Later among the works it cites.
Mohammadi H, Zare A, Soltanolkotabi M, Jovanovic MR. 2021. Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem. IEEE Transactions on Automatic Control
2021
Later among the works it cites.
Li Y, Tang Y, Zhang R, Li N. 2021. Distributed reinforcement learning for decentralized linear quadratic control: A derivative-free policy optimization approach. IEEE Transactions on Automatic Control
2021
Later among the works it cites.
Hambly B, Xu R, Yang H. 2021. Policy gradient methods for the noisy linear quadratic regulator over a finite horizon. SIAM Journal on Control and Optimization
2021
Later among the works it cites.
Perdomo J, Umenberger J, Simchowitz M. 2021. Stabilizing dynamical systems via policy gradient methods. Advances in Neural Information Processing Systems
2021
Later among the works it cites.
Zhang K, Hu B, Başar T. 2021. Policy optimization for ℋ 2 \mathcal{H}_{2} linear control with ℋ ∞ \mathcal{H}_{\infty} robustness guarantee: Implicit regularization and global convergence. SIAM Journal on Control and Optimization
2021
Later among the works it cites.
Zhao F, You K. 2021. Primal-dual learning for the model-free risk-constrained linear quadratic regulator
2021
Later among the works it cites.
Zhang Y, Yang Z, Wang Z. 2021. Provably efficient actor-critic for risk-sensitive and robust adversarial RL: A linear-quadratic case
2021
Later among the works it cites.
2021
Later among the works it cites.
Qu G, Yu C, Low S, Wierman A. 2021. Exploiting linear models for model-free nonlinear control: A provably convergent policy gradient approach
2021
Later among the works it cites.
Fatkhullin I, Polyak B. 2021. Optimizing static linear feedback: Gradient method. SIAM Journal on Control and Optimization
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
Mohammadi H, Soltanolkotabi M, Jovanović MR. 2021. On the lack of gradient domination for linear quadratic Gaussian problems with incomplete state information
2021
Later among the works it cites.
Zhang K, Yang Z, Başar T. 2021. Multi-agent reinforcement learning: A selective overview of theories and algorithms. Handbook of Reinforcement Learning and Control
2021
Later among the works it cites.
Pang B, Jiang ZP. 2021. Robust reinforcement learning: A case study in linear quadratic regulation
2021
Later among the works it cites.
Pang B, Bian T, Jiang ZP. 2021. Robust policy iteration for continuous-time linear quadratic regulation. IEEE Transactions on Automatic Control
2021
Later among the works it cites.
Sun Y, Fazel M. 2021. Learning optimal controllers by policy gradient: Global optimality via convex parameterization
2021
Later among the works it cites.
Wang W, Han J, Yang Z, Wang Z. 2021. Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time
2021
Later among the works it cites.
Chen X, Hazan E. 2021. Black-box control for linear dynamical systems
2021
Later among the works it cites.
Havens A, Hu B. 2021. On imitation learning of linear control policies: Enforcing stability and robustness constraints via LMI conditions
2021
Later among the works it cites.
Yin H, Seiler P, Jin M, Arcak M. 2021. Imitation learning with stability and safety guarantees. IEEE Control Systems Letters
2021
Later among the works it cites.
Molybog I, Lavaei J. 2021. When does MAML objective have benign landscape?
2021
Later among the works it cites.
Ozaslan IK, Mohammadi H, Jovanović MR. 2022. Computing stabilizing feedback gains via a model-free policy gradient method. IEEE Control Systems Letters
2022
Closest in time.
2022
Closest in time.
Guo X, Hu B. 2022. Global convergence of direct policy search for state-feedback ℋ ∞ \mathcal{H}_{\infty} robust control: A revisit of nonsmooth synthesis with Goldstein subdifferential
2022
Closest in time.
Keivan D, Havens A, Seiler P, Dullerud G, Hu B. 2022. Model-free μ \mu synthesis via adversarial reinforcement learning
2022
Closest in time.
Jansch-Porto JP, Hu B, Dullerud GE. 2022. Policy optimization for Markovian jump linear quadratic control: Gradient method and global convergence. IEEE Transactions on Automatic Control
2022
Closest in time.
2022
Closest in time.
Hu B, Zheng Y. 2022. Connectivity of the feasible and sublevel sets of dynamic output feedback control with robustness constraints. IEEE Control Systems Letters
2022
Closest in time.
2022
Closest in time.
Balasubramanian K, Ghadimi S. 2022. Zeroth-order nonconvex stochastic optimization: Handling constraints, high dimensionality, and saddle points. Foundations of Computational Mathematics
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
Tu S, Robey A, Zhang T, Matni N. 2022. On the sample complexity of stability constrained imitation learning
2022
Closest in time.