Fetching the paper…
Reading the bibliography…
Humans can naturally learn to execute a new task by seeing it performed by other individuals once, and then reproduce it in a variety of configurations.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” Advances in Neural Information Processing Systems , 1989
1989
Earlier work this paper cites.
A. Y. Ng, S. J. Russell et al. , “Algorithms for inverse reinforcement learning.” International Conference on Machine Learning , 2000
2000
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” International Conference on Machine learning , 2004
2004
Earlier work this paper cites.
R. Dillmann, “Teaching and learning of robot tasks via observation of human performance,” Robotics and Autonomous Systems , vol. 47, no. 2-3, pp. 109–116, 2004
2004
Earlier work this paper cites.
S. Ekvall and D. Kragic, “Interactive grasp learning based on human demonstration,” in IEEE International Conference on Robotics and Automation, 2004. Proceedings. ICRA’04. 2004 , vol. 4. IEEE, 2004, pp. 3519–3524
2004
Earlier work this paper cites.
S. Calinon and A. Billard, “Teaching a humanoid robot to recognize and reproduce social cues,” in ROMAN 2006-The 15th IEEE International Symposium on Robot and Human Interactive Communication . IEEE, 2006, pp. 346–351
2006
Earlier work this paper cites.
S. Calinon, P. Evrard, E. Gribovskaya, A. Billard, and A. Kheddar, “Learning collaborative manipulation tasks by demonstration using a haptic interface,” in 2009 International Conference on Advanced Robotics . IEEE, 2009, pp. 1–6
2009
Earlier work this paper cites.
V. Kruger, D. L. Herzog, S. Baby, A. Ude, and D. Kragic, “Learning actions from observations,” IEEE robotics & automation magazine , vol. 17, no. 2, pp. 30–43, 2010
2010
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” International Conference on Artificial Intelligence and Statistics , 2011
2011
Earlier work this paper cites.
P. Pastor, L. Righetti, M. Kalakrishnan, and S. Schaal, “Online movement adaptation based on previous sensor experiences,” in 2011 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2011, pp. 365–371
2011
Earlier work this paper cites.
B. Akgun, M. Cakmak, J. W. Yoo, and A. L. Thomaz, “Trajectories and keyframes for kinesthetic teaching: A human-robot interaction perspective,” in Proceedings of the seventh annual ACM/IEEE international conference on Human-Robot Interaction . ACM, 2012, pp. 391–398
2012
Earlier work this paper cites.
B. Kulis et al. , “Metric learning: A survey,” Foundations and Trends in Machine Learning , 2012
2012
Earlier work this paper cites.
K. Lee, Y. Su, T.-K. Kim, and Y. Demiris, “A syntactic approach to robot imitation learning using probabilistic activity grammars,” Robotics and Autonomous Systems , vol. 61, no. 12, pp. 1323–1334, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
A. Frome, G. S. Corrado, J. Shlens, S. Bengio, J. Dean, T. Mikolov et al. , “Devise: A deep visual-semantic embedding model,” Advances in Neural Information Processing Systems , 2013
2013
Earlier work this paper cites.
E. Rohmer, S. P. N. Singh, and M. Freese, “V-rep: a versatile and scalable robot simulation framework,” in Proc. of The International Conference on Intelligent Robots and Systems (IROS) , 2013
2013
Cited alongside, same era.
Y. Yang, Y. Li, C. Fermuller, and Y. Aloimonos, “Robot learning manipulation action plans by” watching” unconstrained videos from the world wide web,” in Twenty-Ninth AAAI Conference on Artificial Intelligence , 2015
2015
Cited alongside, same era.
G. Koch, R. Zemel, and R. Salakhutdinov, “Siamese neural networks for one-shot image recognition,” ICML Deep Learning Workshop , 2015
2015
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” International Conference on Learning Representation , 2015
2015
Cited alongside, same era.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” The Journal of Machine Learning Research , 2016
K. Ramirez-Amaro, M. Beetz, and G. Cheng, “Transferring skills to humanoid robots by extracting semantic representations from observations of human activities,” Artificial Intelligence , vol. 247, pp. 95–118, 2017
2017
Later among the works it cites.
S. Ravi and H. Larochelle, “Optimization as a model for few-shot learning,” International Conference on Learning Representations , 2017
2017
Later among the works it cites.
E. Triantafillou, R. Zemel, and R. Urtasun, “Few-shot learning through an information retrieval lens,” Advances in Neural Information Processing Systems , 2017
2017
Later among the works it cites.
J. Snell, K. Swersky, and R. Zemel, “Prototypical networks for few-shot learning,” Advances in Neural Information Processing Systems , 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
C. Finn, S. Levine, and P. Abbeel, “Guided cost learning: Deep inverse optimal control via policy optimization,” International Conference on Machine Learning , 2016
2016
Cited alongside, same era.
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra et al. , “Matching networks for one shot learning,” Advances in Neural Information Processing Systems , 2016
2016
Cited alongside, same era.
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap, “Meta-learning with memory-augmented neural networks,” International Conference on Machine Learning , 2016
2016
Cited alongside, same era.
S. James and E. Johns, “3d simulation for robot arm control with deep q-learning,” NIPS Workshop (Deep Learning for Action and Interaction) , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Gualtieri, A. ten Pas, K. Saenko, and R. Platt, “High precision grasp pose detection in dense clutter,” in IROS , 2016, pp. 598–605
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Later among the works it cites.
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine, “One-shot visual imitation learning via meta-learning,” Conference on Robot Learning , 2017
2017
Later among the works it cites.
T. Yu, C. Finn, A. Xie, S. Dasari, T. Zhang, P. Abbeel, and S. Levine, “One-shot imitation from observing humans via domain-adaptive meta-learning,” Robotics: Science and Systems , 2018
2018
Later among the works it cites.
J. Matas, S. James, and A. J. Davison, “Sim-to-real reinforcement learning for deformable object manipulation,” Conference on Robot Learning , 2018
2018
Later among the works it cites.
K. Bousmalis, A. Irpan, P. Wohlhart, Y. Bai, M. Kelcey, M. Kalakrishnan, L. Downs, J. Ibarz, P. Pastor, K. Konolige et al. , “Using simulation and domain adaptation to improve efficiency of deep robotic grasping,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 4243–4250
2018
Later among the works it cites.
S. James, M. Bloesch, and A. J. Davison, “Task-embedded control networks for few-shot imitation learning,” Conference on Robot Learning , 2018
2018
Later among the works it cites.
T. Zhang, Z. McCarthy, O. Jow, D. Lee, K. Goldberg, and P. Abbeel, “Deep imitation learning for complex manipulation tasks from virtual reality teleoperation,” International Conference on Robotics and Automation , 2018
2018
Later among the works it cites.
J. Rothfuss, F. Ferreira, E. E. Aksoy, Y. Zhou, and T. Asfour, “Deep episodic memory: Encoding, recalling, and predicting episodic experiences for robot action execution,” IEEE Robotics and Automation Letters , vol. 3, no. 4, pp. 4007–4014, 2018
2018
Later among the works it cites.
S. James, P. Wohlhart, M. Kalakrishnan, D. Kalashnikov, A. Irpan, J. Ibarz, S. Levine, R. Hadsell, and K. Bousmalis, “Sim-to-real via sim-to-sim: Data-efficient robotic grasping via randomized-to-canonical adaptation networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 12 627–12 637
2019
Closest in time.