Fetching the paper…
Reading the bibliography…
This paper proposes an inverse optimal control method which enables a robot to incrementally learn a control objective function from a collection of trajectory segments.
L. S. Pontryagin, V. Boltyanskiy, R. V. Gamkrelidze, and E. Mishchenko, “Mathematical theory of optimal processes,” 1962
1962
Earlier work this paper cites.
J. B. Kuipers, Quaternions and rotation sequences: a primer with applications to orbits, aerospace, and virtual reality . Princeton university press, 1999
1999
Earlier work this paper cites.
A. Y. Ng, S. J. Russell, et al. , “ Algorithms for inverse reinforcement learning ,” in International Conference on Machine Learning , 2000, pp. 663–670
2000
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “ Apprenticeship learning via inverse reinforcement learning ,” in International Conference on Machine Learning . ACM, 2004, p. 1
2004
Earlier work this paper cites.
N. D. Ratliff, J. A. Bagnell, and M. A. Zinkevich, “ Maximum margin planning ,” in International Conference on Machine Learning . ACM, 2006, pp. 729–736
2006
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “ Maximum Entropy Inverse Reinforcement Learning ,” in AAAI , vol. 8. Chicago, IL, USA, 2008, pp. 1433–1438
2008
Earlier work this paper cites.
M. W. Spong and M. Vidyasagar, Robot dynamics and control . John Wiley & Sons, 2008
2008
Earlier work this paper cites.
B. D. Ziebart, N. Ratliff, G. Gallagher, C. Mertz, K. Peterson, J. A. Bagnell, M. Hebert, A. K. Dey, and S. Srinivasa, “ Planning-based prediction for pedestrians ,” in IEEE/RSJ International Conference on Intelligent Robots and Systems , 2009, pp. 3931–3936
2009
Earlier work this paper cites.
K. Mombaur, A. Truong, and J.-P. Laumond, “ From human to humanoid locomotion—an inverse optimal control approach ,” Autonomous robots , vol. 28, no. 3, pp. 369–383, 2010
2010
Earlier work this paper cites.
T. Lee, M. Leok, and N. H. McClamroch, “Geometric tracking control of a quadrotor uav on se (3),” in 49th IEEE conference on decision and control (CDC) . IEEE, 2010, pp. 5420–5425
2010
Earlier work this paper cites.
A. Keshavarz, Y. Wang, and S. Boyd, “ Imputing a convex objective function ,” in IEEE International Symposium on Intelligent Control . IEEE, 2011, pp. 613–619
2011
Cited alongside, same era.
A.-S. Puydupin-Jamin, M. Johnson, and T. Bretl, “ A convex approach to inverse optimal control and its application to modeling human locomotion ,” in 2012 IEEE International Conference on Robotics and Automation , 2012, pp. 531–536
2012
Cited alongside, same era.
2012
Cited alongside, same era.
D. Bertsekas, Dynamic programming and optimal control: Volume I . Athena scientific, 2012, vol. 1
2012
Cited alongside, same era.
M. Kuderer, S. Gulati, and W. Burgard, “ Learning driving styles for autonomous vehicles from demonstration ,” in IEEE International Conference on Robotics and Automation . IEEE, 2015, pp. 2641–2646
K. Bogert and P. Doshi, “Scaling expectation-maximization for inverse reinforcement learning to multiple robots under occlusion,” in Proceedings of the 16th Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, 2017, pp. 522–529
2017
Later among the works it cites.
T. L. Molloy, J. J. Ford, and T. Perez, “ Finite-horizon inverse optimal control for discrete-time nonlinear systems ,” Automatica , vol. 87, pp. 442–446, 2018
2018
Later among the works it cites.
2020
Closest in time.
2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
A. D. Dragan, K. Muelling, J. A. Bagnell, and S. S. Srinivasa, “ Movement primitives via optimization ,” in IEEE International Conference on Robotics and Automation , 2015, pp. 2339–2346
2015
Cited alongside, same era.
J. Mainprice, R. Hayne, and D. Berenson, “ Goal set inverse optimal control and iterative replanning for predicting human reaching motions in shared workspaces ,” IEEE Transactions on Robotics , vol. 32, no. 4, pp. 897–908, 2016
2016
Cited alongside, same era.
K. Bogert, J. F.-S. Lin, P. Doshi, and D. Kulic, “ Expectation-maximization for inverse reinforcement learning with hidden data ,” in International Conference on Autonomous Agents & Multiagent Systems , 2016, pp. 1034–1042
2016
Cited alongside, same era.
P. Englert, N. A. Vien, and M. Toussaint, “ Inverse KKT: Learning cost functions of manipulation tasks from demonstrations ,” The International Journal of Robotics Research , vol. 36, no. 13-14, pp. 1474–1488, 2017
2017
Cited alongside, same era.
A. Bajcsy, D. P. Losey, M. K. O’Malley, and A. D. Dragan, “ Learning robot objectives from physical human interaction ,” Proceedings of Machine Learning Research , vol. 78, pp. 217–226, 2017
2017
Cited alongside, same era.
W. Jin, Z. Wang, Z. Yang, and S. Mou, “Pontryagin differentiable programming: An end-to-end learning and control framework,” Advances in Neural Information Processing Systems (NeurIPS) , 2020
2020
Closest in time.
S. Byeon, W. Jin, D. Sun, and I. Hwang, “Human-automation interaction for assisting novices to emulate experts by inferring task objective functions,” in 2021 IEEE/AIAA 40th Digital Avionics Systems Conference (DASC) . IEEE, 2021, pp. 1–6
2021
Closest in time.
W. Jin, D. Kulić, S. Mou, and S. Hirche, “Inverse optimal control from incomplete trajectory observations,” The International Journal of Robotics Research , vol. 40, no. 6-7, pp. 848–865, 2021
2021
Closest in time.
W. Jin, S. Mou, and G. J. Pappas, “Safe pontryagin differentiable programming,” Advances in Neural Information Processing Systems (NeurIPS) , 2021
2021
Closest in time.
W. Jin, D. Kulić, S. Mou, and S. Hirche, “Inverse optimal control from incomplete trajectory observations,” The International Journal of Robotics Research, Accpeted, in press , 2021
2021
Closest in time.
W. Jin and S. Mou, “Distributed inverse optimal control,” Automatica , vol. 129, p. 109658, 2021
2021
Closest in time.