Fetching the paper…
Reading the bibliography…
Optimal control problems are inherently hard to solve as the optimization must be performed simultaneously with updating the underlying system.
R. Bellman, Functional equations in the theory of dynamic programming. V. Positivity and quasi-linearity, Proc. Natl. Acad. Sci. USA
1955
Earlier work this paper cites.
R. Bellman, Dynamic Programming
1957
Earlier work this paper cites.
R. A. Howard, Dynamic Programming and Markov Processes
1960
Earlier work this paper cites.
M.L. Puterman and S. L. Brumelle, On the convergence of policy iteration in stationary dynamic programming, Math. Oper. Res
1979
Earlier work this paper cites.
N. V. Krylov, Controlled Diffusion Processes
1980
Earlier work this paper cites.
M. L. Puterman, On the convergence of policy iteration for controlled diffusions, J. Optim. Theory Appl
1981
Earlier work this paper cites.
O. Hernandez-Lerma and J. Lasserre, Discrete-Time Markov Control Processes
1996
Earlier work this paper cites.
R. Buckdahn and S. Peng, Ergodic Backward SDE and Associated PDE, R.C. Dalang, M. Dozzi and F. Russo eds., Progr. Probab. 45
1999
Cited alongside, same era.
M. Fuhrman and G. Tessitore, Infinite horizon backward stochastic differential equations and elliptic equations in Hilbert Spaces, Ann. Probab
2004
Cited alongside, same era.
W. H. Fleming and H. M. Soner, Controlled Markov Processes and Viscosity Solutions
2006
Cited alongside, same era.
H. Dong and N. V. Krylov, The rate of convergence of finite-difference approximations for parabolic Bellman equations with Lipschitz coefficients in cylindrical domains, Appl. Math. Optim
2007
Cited alongside, same era.
O. Bokanowski, S. Maroso, and H. Zidani, Some convergence results for Howard’s algorithm, SIAM J. on Numer. Anal
2009
Cited alongside, same era.
A. Richou, Markovian quadratic and superquadratic BSDEs with an unbounded terminal condition, Stochastic Process. Appl
2012
Later among the works it cites.
N. Bäuerle and U. Rieder, Control improvement for jump-diffusion processes with applications to finance, Appl. Math. and Optim
2012
Later among the works it cites.
S. D. Jacka and A. Mijatović, On the policy improvement algorithm in continuous time, Stochastics
2017
Later among the works it cites.
S. D. Jacka, A. Mijatović and D. Siraj, Coupling and a Generalised Policy Iteration Algorithm in Continuous Time, arXiv
2017
Later among the works it cites.
J. Maeda and S. D. Jacka, Evaluation of the Rate of Convergence in the PIA. arXiv
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
I. Gyöngy and D. Šiska, On finite-difference approximations for normalized Bellman equations, Appl. Math. Optim
2009
Cited alongside, same era.
H. Pham. Continuous-time Stochastic Control and Optimization with Financial Applications
2009
Cited alongside, same era.
2094
Closest in time.