Fetching the paper…
Reading the bibliography…
This paper proposes a novel approach that enables a robot to learn an objective function incrementally from human directional corrections.
B. Grünbaum et al. , “Partitions of mass-distributions and of convex bodies by hyperplanes.” Pacific Journal of Mathematics , vol. 10, no. 4, pp. 1257–1261, 1960
1960
Earlier work this paper cites.
D. J. Newman, “Location of the maximum on unimodal surfaces,” Journal of the ACM , vol. 12, no. 3, pp. 395–398, 1965
1965
Earlier work this paper cites.
P. Moylan and B. Anderson, “Nonlinear regulator theory and an inverse optimal control problem,” IEEE Transactions on Automatic Control , vol. 18, no. 5, pp. 460–465, 1973
1973
Earlier work this paper cites.
J. Elzinga and T. G. Moore, “A central cutting plane algorithm for the convex programming problem,” Mathematical Programming , vol. 8, no. 1, pp. 134–145, 1975
1975
Earlier work this paper cites.
S. P. Tarasov, “The method of inscribed ellipsoids,” in Soviet Mathematics-Doklady , vol. 37, no. 1, 1988, pp. 226–230
1988
Earlier work this paper cites.
J.-L. Goffin and J.-P. Vial, “On the computation of weighted analytic centers and dual ellipsoids with the projective algorithm,” Mathematical Programming , vol. 60, no. 1-3, pp. 81–92, 1993
1993
Earlier work this paper cites.
D. S. Atkinson and P. M. Vaidya, “A cutting plane algorithm for convex programming that uses analytic centers,” Mathematical Programming , vol. 69, no. 1-3, pp. 1–43, 1995
1995
Earlier work this paper cites.
J. B. Kuipers, Quaternions and rotation sequences . Princeton University Press, 1999, vol. 66
1999
Earlier work this paper cites.
A. Y. Ng, S. J. Russell et al. , “Algorithms for inverse reinforcement learning.” in International Conference on Machine Learning , vol. 1, 2000, p. 2
2000
Earlier work this paper cites.
W. Li and E. Todorov, “Iterative linear quadratic regulator design for nonlinear biological movement systems.” in International Conference on Informatics in Control, Automation and Robotics , 2004, pp. 222–229
2004
Earlier work this paper cites.
S. Boyd, S. P. Boyd, and L. Vandenberghe, Convex optimization . Cambridge university press, 2004
2004
Earlier work this paper cites.
N. D. Ratliff, J. A. Bagnell, and M. A. Zinkevich, “Maximum margin planning,” in International Conference on Machine Learning , 2006, pp. 729–736
2006
Earlier work this paper cites.
S. Boyd and L. Vandenberghe, “Localization and cutting-plane methods,” Stanford EE 364b Lecture Notes , 2007
2007
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “Maximum entropy inverse reinforcement learning.” in Association for the Advancement of Artificial Intelligence , 2008
2008
Earlier work this paper cites.
T. Lee, M. Leok, and N. H. McClamroch, “Geometric tracking control of a quadrotor uav on se(3),” in IEEE Conference on Decision and Control , 2010, pp. 5420–5425
2010
Cited alongside, same era.
P. Shivaswamy and T. Joachims, “Online structured prediction via coactive learning,” in International Conference on Machine Learning , 2012, pp. 59–66
2012
Cited alongside, same era.
A.-S. Puydupin-Jamin, M. Johnson, and T. Bretl, “A convex approach to inverse optimal control and its application to modeling human locomotion,” in International Conference on Robotics and Automation , 2012, pp. 531–536
2012
Cited alongside, same era.
A. Jain, B. Wojcik, T. Joachims, and A. Saxena, “Learning trajectory preferences for manipulators via iterative improvement,” in Advances in neural information processing systems , 2013, pp. 575–583
2013
Cited alongside, same era.
A. Bobu, A. Bajcsy, J. F. Fisac, and A. D. Dragan, “Learning under misspecified objective spaces,” in Conference on Robot Learning , 2018, pp. 796–805
2018
Later among the works it cites.
J. Y. Zhang and A. D. Dragan, “Learning from extrapolated corrections,” in International Conference on Robotics and Automation . IEEE, 2019, pp. 7034–7040
2019
Later among the works it cites.
W. Jin, D. Kulić, J. F.-S. Lin, S. Mou, and S. Hirche, “Inverse optimal control for multiphase cost functions,” IEEE Transactions on Robotics , vol. 35, no. 6, pp. 1387–1398, 2019
2019
Later among the works it cites.
D. P. Losey and M. K. O’Malley, “Learning the correct robot trajectory in real-time from physical human interactions,” ACM Transactions on Human-Robot Interaction , vol. 9, no. 1, pp. 1–19, 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Kuderer, S. Gulati, and W. Burgard, “Learning driving styles for autonomous vehicles from demonstration,” in International Conference on Robotics and Automation . IEEE, 2015, pp. 2641–2646
2015
Cited alongside, same era.
A. Jain, S. Sharma, T. Joachims, and A. Saxena, “Learning preferences for manipulation tasks from online coactive feedback,” The International Journal of Robotics Research , vol. 34, no. 10, pp. 1296–1313, 2015
2015
Cited alongside, same era.
A. D. Dragan, K. Muelling, J. A. Bagnell, and S. S. Srinivasa, “Movement primitives via optimization,” in International Conference on Robotics and Automation . IEEE, 2015, pp. 2339–2346
2015
Cited alongside, same era.
S. Diamond and S. Boyd, “CVXPY: A Python-embedded modeling language for convex optimization,” Journal of Machine Learning Research , vol. 17, no. 83, pp. 1–5, 2016
2016
Cited alongside, same era.
P. Englert, N. A. Vien, and M. Toussaint, “Inverse kkt: Learning cost functions of manipulation tasks from demonstrations,” The International Journal of Robotics Research , vol. 36, no. 13-14, pp. 1474–1488, 2017
2017
Cited alongside, same era.
A. Bajcsy, D. P. Losey, M. K. O’Malley, and A. D. Dragan, “Learning robot objectives from physical human interaction,” Conference on Robot Learning , vol. 78, pp. 217–226, 2017
2017
Cited alongside, same era.
B. Amos and J. Z. Kolter, “Optnet: Differentiable optimization as a layer in neural networks,” in International Conference on Machine Learning , 2017, pp. 136–145
2017
Cited alongside, same era.
Y. Nesterov, “Cutting plane algorithms from analytic centers: efficiency estimates,” Mathematical Programming , vol. 69, no. 1, pp. 149–176, 1995
2017
Cited alongside, same era.
J. A. E. Andersson, J. Gillis, G. Horn, J. B. Rawlings, and M. Diehl, “CasADi – A software framework for nonlinear optimization and optimal control,” Mathematical Programming Computation , vol. 11, no. 1, pp. 1–36, 2019
2019
Later among the works it cites.
H. Ravichandar, A. S. Polydoros, S. Chernova, and A. Billard, “Recent advances in robot learning from demonstration,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 3, 2020
2020
Closest in time.
W. Jin, Z. Wang, Z. Yang, and S. Mou, “Pontryagin differentiable programming: An end-to-end learning and control framework,” Advances in Neural Information Processing Systems , vol. 33, 2020
2020
Closest in time.
W. Jin, D. Kulić, S. Mou, and S. Hirche, “Inverse optimal control from incomplete trajectory observations,” The International Journal of Robotics Research , vol. 40, no. 6-7, pp. 848–865, 2021
2021
Closest in time.
W. Jin, S. Mou, and G. J. Pappas, “Safe pontryagin differentiable programming,” in Advances in Neural Information Processing Systems , 2021
2021
Closest in time.
W. Jin, T. D. Murphey, D. Kulić, N. Ezer, and S. Mou, “Learning from sparse demonstrations,” IEEE Transactions on Robotics , pp. 1–20, 2022
2022
Closest in time.
Z. Liang, W. Jin, and S. Mou, “An iterative method for inverse optimal control,” in Asian Control Conference , 2022, pp. 959–964
2022
Closest in time.
R. Tedrake, “Underactuated robotics: Algorithms for walking, running, swimming, flying, and manipulation,” Course Notes for MIT 6.832 , 2022
2022
Closest in time.
W. Jin, A. Aydinoglu, M. Halm, and M. Posa, “Learning linear complementarity systems,” Learning for Dynamics and Control Conference , 2022
2022
Closest in time.