Fetching the paper…
Reading the bibliography…
It is difficult for humans to efficiently teach robots how to correctly perform a task.
Online learning and stochastic approximations
L. Bottou · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng and S. J. Russell · 2000
Earlier work this paper cites.
The unscented Kalman filter for nonlinear estimation
E. A. Wan and R. Van Der Merwe · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
Principles of Robot Motion: Theory, Algorithms, and Implementation
H. M. Choset · 2005
Earlier work this paper cites.
Maximum margin planning
N. D. Ratliff, J. A. Bagnell, and M. A. Zinkevich · 2006
Earlier work this paper cites.
Bayesian inverse reinforcement learning
D. Ramachandran and E. Amir · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
Applying the unscented kalman filter for nonlinear state estimation
R. Kandepu, B. Foss, and L. Imsland · 2008
Earlier work this paper cites.
Active learning for reward estimation in inverse reinforcement learning
M. Lopes, F. Melo, and L. Montesano · 2009
Cited alongside, same era.
Comparing action-query strategies in semi-autonomous agents
R. Cohn, E. Durfee, and S. Singh · 2011
Cited alongside, same era.
Sampling-based algorithms for optimal motion planning
S. Karaman and E. Frazzoli · 2011
Cited alongside, same era.
LQG-MP: Optimized path planning for robots with motion uncertainty and imperfect state information
J. Van Den Berg, P. Abbeel, and K. Goldberg · 2011
Cited alongside, same era.
Keyframe-based learning from demonstration
B. Akgun, M. Cakmak, K. Jiang, and A. L. Thomaz · 2012
Cited alongside, same era.
Active learning
B. Settles · 2012
Cited alongside, same era.
Coactive learning
P. Shivaswamy and T. Joachims · 2015
Later among the works it cites.
Learning preferences for manipulation tasks from online coactive feedback
A. Jain, S. Sharma, T. Joachims, and A. Saxena · 2015
Later among the works it cites.
Learning robot objectives from physical human interaction
A. Bajcsy, D. P. Losey, M. K. O’Malley, and A. D. Dragan · 2017
Later among the works it cites.
Enabling robots to communicate their objectives
S. H. Huang, D. Held, P. Abbeel, and A. D. Dragan · 2017
Later among the works it cites.
Inverse reward design
D. Hadfield-Menell, S. Milli, P. Abbeel, S. J. Russell, and A. Dragan · 2017
Later among the works it cites.
An algorithmic perspective on imitation learning
T. Osa, J. Pajarinen, G. Neumann, J. A. Bagnell, P. Abbeel, and J. Peters · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Markov Decision Processes: Discrete Stochastic Dynamic Programming
M. L. Puterman · 2014
Cited alongside, same era.
Motion planning with sequential convex optimization and convex collision checking
J. Schulman, Y. Duan, J. Ho, A. Lee, I. Awwal, H. Bradlow, J. Pan, S. Patil, K. Goldberg, and P. Abbeel · 2014
Cited alongside, same era.
Predicting initialization effectiveness for trajectory optimization
J. Pan, Z. Chen, and P. Abbeel · 2014
Cited alongside, same era.
Learning from physical human corrections, one feature at a time
A. Bajcsy, D. P. Losey, M. K. O’Malley, and A. D. Dragan · 2018
Closest in time.
Active reward learning from critiques
Y. Cui and S. Niekum · 2018
Closest in time.