Fetching the paper…
Reading the bibliography…
Motivated by recent advances in Deep Learning for robot control, this paper considers two learning algorithms in terms of how they acquire demonstrations.
E. V. Slud, “Distribution inequalities for the binomial law,” Ann. Probab. , vol. 5, no. 3, pp. 404–412, 06 1977. [Online]. Available: http://dx.doi.org/10.1214/aop/1176995801
1977
Earlier work this paper cites.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” Carnegie-Mellon University, Tech. Rep., 1989
1989
Earlier work this paper cites.
V. Vapnik, “Principles of risk minimization for learning theory,” in Advances in Neural Information Processing Systems , 1992, pp. 831–838
1992
Earlier work this paper cites.
R. Dillmann, M. Kaiser, and A. Ude, “Acquisition of elementary robot skills from human demonstration,” in International symposium on intelligent robotics systems . Citeseer, 1995, pp. 185–192
1995
Earlier work this paper cites.
P. L. Bartlett and S. Mendelson, “Rademacher and gaussian complexities: Risk bounds and structural results,” Journal of Machine Learning Research , vol. 3, no. Nov, pp. 463–482, 2002
2002
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A survey of robot learning from demonstration,” Robotics and autonomous systems , vol. 57, no. 5, pp. 469–483, 2009
2009
Earlier work this paper cites.
S. Ross and D. Bagnell, “Efficient reductions for imitation learning,” in International Conference on Artificial Intelligence and Statistics , 2010, pp. 661–668
2010
Earlier work this paper cites.
S. Ross, G. J. Gordon, and J. A. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” AISTATS. Vol. 1. No. 2 , 2011
2011
Cited alongside, same era.
B. Akgun, M. Cakmak, K. Jiang, and A. L. Thomaz, “Keyframe-based learning from demonstration,” International Journal of Social Robotics , vol. 4, no. 4, pp. 343–355, 2012
2012
Cited alongside, same era.
H. He, J. Eisner, and H. Daume, “Imitation learning by coaching,” in Advances in Neural Information Processing Systems , 2012, pp. 3149–3157
2012
Cited alongside, same era.
S. Shalev-Shwartz et al. , “Online learning and online convex optimization,” Foundations and Trends® in Machine Learning , vol. 4, no. 2, pp. 107–194, 2012
2012
Cited alongside, same era.
F. Duvallet, T. Kollar, and A. Stentz, “Imitation learning for natural language direction following through unknown environments,” in ICRA . IEEE, 2013, pp. 1047–1053
S. Ross, N. Melik-Barkhudarov, K. S. Shankar, A. Wendel, D. Dey, J. A. Bagnell, and M. Hebert, “Learning monocular reactive uav control in cluttered natural environments,” in ICRA, 2013 IEEE . IEEE
2013
Later among the works it cites.
2014
Later among the works it cites.
S. Bengio, O. Vinyals, N. Jaitly, and N. Shazeer, “Scheduled sampling for sequence prediction with recurrent neural networks,” in Advances in Neural Information Processing Systems , 2015, pp. 1171–1179
2015
Later among the works it cites.
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
B. Kim and J. Pineau, “Maximum mean discrepancy imitation learning.” in Robotics Science and Systems , 2013
2013
Cited alongside, same era.
S. Levine and V. Koltun, “Variational policy search via trajectory optimization,” in Advances in Neural Information Processing Systems , 2013, pp. 207–215
2013
Cited alongside, same era.
“Tensor flow,” https://www.tensorflow.org/
Cited in the paper.
B. Akgun, K. Subramanian, and A. L. Thomaz, “Novel interaction strategies for learning from teleoperation.”
Cited in the paper.
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska-Barwińska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou et al. , “Hybrid computing using a neural network with dynamic external memory,” Nature , vol. 538, no. 7626, pp. 471–476, 2016
2016
Closest in time.
M. Laskey, J. Lee, C. Chuck, D. Gealy, W. Hsieh, F. T. Pokorny, A. D. Dragan, and K. Goldberg, “Robot grasping in clutter: Using a hierarchy of supervisors for learning from demonstrations,” Automation Science and Engineering (CASE), 2016 IEEE , pp. 827–834, 2016
2016
Closest in time.
M. Laskey, S. Staszak, W. Y.-S. Hsieh, J. Mahler, F. T. Pokorny, A. D. Dragan, and K. Goldberg, “Shiv: Reducing supervisor burden in dagger using support vectors for efficient learning from demonstrations in high dimensional state spaces,” in Robotics and Automation (ICRA), 2016 IEEE International Conference on . IEEE, 2016, pp. 462–469
2016
Closest in time.