Fetching the paper…
Reading the bibliography…
Recent work have shown how the optimal state-feedback, obtained as the solution to the Hamilton-Jacobi-Bellman equations, can be approximated for several nonlinear, deterministic systems by deep neural networks.
L. S. Pontryagin, Mathematical theory of optimal processes . CRC Press, 1987
1987
Earlier work this paper cites.
A. Berz, “Differential algebraic description of beam dynamics to very high orders,” Part. Accel. , vol. 24, no. SSC-152, pp. 109–124, 1988
1988
Earlier work this paper cites.
M. G. Crandall, H. Ishii, and P.-L. Lions, “User’s guide to viscosity solutions of second order partial differential equations,” Bulletin of the American mathematical society , vol. 27, no. 1, pp. 1–67, 1992
1992
Earlier work this paper cites.
W. McDermott and M. Athans, “Approximating optimal state feedback using neural networks,” in Proceedings of 1994 33rd IEEE Conference on Decision and Control , vol. 3. IEEE, 1994, pp. 2466–2471
1994
Earlier work this paper cites.
T. Hrycej, “Stability and equilibrium points in neurocontrol,” in Proceedings of ICNN’95-International Conference on Neural Networks , vol. 1. IEEE, 1995, pp. 617–621
1995
Earlier work this paper cites.
R. A. Freeman and P. Kokotovic, “Inverse optimality in robust stabilization,” SIAM journal on control and optimization , vol. 34, no. 4, pp. 1365–1391, 1996
1996
Earlier work this paper cites.
R. W. Beard, G. N. Saridis, and J. T. Wen, “Galerkin approximations of the generalized hamilton-jacobi-bellman equation,” Automatica , vol. 33, no. 12, pp. 2159–2177, 1997
1997
Earlier work this paper cites.
M. Berz and K. Makino, “Verified integration of odes and flows using differential algebraic methods on high-order taylor models,” Reliable Computing , vol. 4, no. 4, pp. 361–369, 1998
1998
Earlier work this paper cites.
E. Todorov, “Optimal control theory,” Bayesian brain: probabilistic approaches to neural coding , pp. 269–298, 2006
2006
Cited alongside, same era.
K. J. Åström and B. Wittenmark, “Model-reference adaptive systems,” in Adaptive control , 2nd ed. Mineola(NY), USA: Dover Publications INC, 2008, ch. 5, pp. 185–262
2008
Cited alongside, same era.
K. G. Vamvoudakis and F. L. Lewis, “Online actor–critic algorithm to solve the continuous-time infinite horizon optimal control problem,” Automatica , vol. 46, no. 5, pp. 878–888, 2010
2010
Cited alongside, same era.
R. Armellin, P. Di Lizia, F. Bernelli-Zazzera, and M. Berz, “Asteroid close encounters characterization using differential algebra: the case of apophis,” Celestial Mechanics and Dynamical Astronomy , vol. 107, no. 4, pp. 451–470, 2010
2010
Cited alongside, same era.
2015
Later among the works it cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in Advances in Neural Information Processing Systems , 2016, pp. 4565–4573
2016
Later among the works it cites.
A. C. Luo, Periodic flows to chaos in time-delay systems . Springer, 2017
2017
Later among the works it cites.
C. Sánchez-Sánchez and D. Izzo, “Real-time optimal control via deep neural networks: study on landing problems,” Journal of Guidance, Control, and Dynamics , vol. 41, no. 5, pp. 1122–1135, 2018
2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
D. Nodland, H. Zargarzadeh, and S. Jagannathan, “Neural network-based optimal adaptive output feedback control of a helicopter uav,” IEEE transactions on neural networks and learning systems , vol. 24, no. 7, pp. 1061–1073, 2013
2013
Cited alongside, same era.
V. Levine, Sergey; Koltun, “Guided policy search,” International Conference on Machine Learning , 2013
2013
Cited alongside, same era.
K. G. Vamvoudakis, F. L. Lewis, and S. S. Ge, “Neural networks in feedback control systems,” Mechanical Engineers’ Handbook , pp. 1–52, 2014
2014
Cited alongside, same era.
2018
Closest in time.
D. Izzo and F. Biscani, “audi/pyaudi,” Oct. 2018. [Online]. Available: https://doi.org/10.5281/zenodo.1442738
2018
Closest in time.