Fetching the paper…
Reading the bibliography…
Learning from human demonstrations (behavior cloning) is a cornerstone of robot learning.
D. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in Proceedings of (NeurIPS) Neural Information Processing Systems , D. Touretzky, Ed. Morgan Kaufmann, December 1989, pp. 305 – 313
1989
Earlier work this paper cites.
J. Nakanishi, J. Morimoto, G. Endo, G. Cheng, S. Schaal, and M. Kawato, “Learning from demonstration and adaptation of biped locomotion,” Robotics and autonomous systems , vol. 47, no. 2-3, pp. 79–91, 2004
2004
Earlier work this paper cites.
N. Ratliff, J. A. Bagnell, and S. S. Srinivasa, “Imitation learning for locomotion and manipulation,” in 2007 7th IEEE-RAS International Conference on Humanoid Robots . IEEE, 2007, pp. 392–397
2007
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” in Proceedings of the 26th Annual International Conference on Machine Learning , ser. ICML ’09. New York, NY, USA: Association for Computing Machinery, 2009, p. 41–48. [Online]. Available: https://doi.org/10.1145/1553374.1553380
2009
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 627–635
2011
Earlier work this paper cites.
T. Brys, A. Harutyunyan, H. B. Suay, S. Chernova, M. E. Taylor, and A. Nowé, “Reinforcement learning from demonstration through shaping,” in Twenty-fourth international joint conference on artificial intelligence , 2015
2015
Earlier work this paper cites.
M. Bojarski, D. D. Testa, D. Dworakowski, B. Firner, B. Flepp, P. Goyal, L. D. Jackel, M. Monfort, U. Muller, J. Zhang, X. Zhang, J. Zhao, and K. Zieba, “End to end learning for self-driving cars,” 2016
2016
Earlier work this paper cites.
K. Subramanian, C. L. Isbell Jr, and A. L. Thomaz, “Exploration from demonstration for interactive reinforcement learning,” in Proceedings of the 2016 international conference on autonomous agents & multiagent systems , 2016, pp. 447–456
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation learning: A survey of learning methods,” ACM Comput. Surv. , vol. 50, no. 2, apr 2017. [Online]. Available: https://doi.org/10.1145/3054912
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
W. Farag and Z. Saleh, “Behavior cloning for autonomous driving using convolutional neural networks,” in 2018 International Conference on Innovation and Intelligence for Informatics, Computing, and Technologies (3ICT) , 2018, pp. 1–7
2018
Earlier work this paper cites.
2018
Cited alongside, same era.
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel, “Deep imitation learning for complex manipulation tasks from virtual reality teleoperation,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) , 2018, pp. 5628–5635
2018
Cited alongside, same era.
B. Kang, Z. Jie, and J. Feng, “Policy optimization with demonstrations,” in Proceedings of the 35th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, J. Dy and A. Krause, Eds., vol. 80. PMLR, 10–15 Jul 2018, pp. 2469–2478. [Online]. Available: https://proceedings.mlr.press/v80/kang18a.html
2018
Cited alongside, same era.
T. Salimans and R. Chen, “Learning montezuma’s revenge from a single demonstration,” 2018
2018
Cited alongside, same era.
S. Tu, A. Robey, T. Zhang, and N. Matni, “On the sample complexity of stability constrained imitation learning,” in Conference on Learning for Dynamics & Control , 2021. [Online]. Available: https://api.semanticscholar.org/CorpusID:235358617
2021
Later among the works it cites.
2021
Later among the works it cites.
Franka Emika Robot’s Instruction Handbook . Franka Emika GmbH, 2021
2021
Later among the works it cites.
N. M. Shafiullah, Z. Cui, A. A. Altanzaya, and L. Pinto, “Behavior transformers: Cloning k k modes with one stone,” Advances in neural information processing systems , vol. 35, pp. 22 955–22 968, 2022
2022
Later among the works it cites.
L. Lai, A. Z. Huang, and S. J. Gershman, “Action chunking as policy compression,” 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Florensa, D. Held, M. Wulfmeier, M. Zhang, and P. Abbeel, “Reverse curriculum generation for reinforcement learning,” 2018
2018
Cited alongside, same era.
A. Nair, B. McGrew, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Overcoming exploration in reinforcement learning with demonstrations,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 6292–6299
2018
Cited alongside, same era.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev et al. , “Grandmaster level in starcraft ii using multi-agent reinforcement learning,” Nature , vol. 575, no. 7782, pp. 350–354, 2019
2019
Cited alongside, same era.
S.-J. Lee, T. Y. Chun, H. W. Lim, and S.-H. Lee, “Path tracking control using imitation learning with variational auto-encoder,” in 2019 19th International Conference on Control, Automation and Systems (ICCAS) , 2019, pp. 501–505
2019
Cited alongside, same era.
J. C. Vargas, M. Bhoite, and A. B. Farimani, “Creativity in robot manipulation with deep reinforcement learning,” 2019
2019
Cited alongside, same era.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation for games, robotics and machine learning,” http://pybullet.org , 2016–2019
2019
Cited alongside, same era.
J. Choi, H. Kim, Y. Son, C.-W. Park, and J. H. Park, “Robotic behavioral cloning through task building,” in 2020 International Conference on Information and Communication Technology Convergence (ICTC) , 2020, pp. 1279–1281
2020
Cited alongside, same era.
A. Mandlekar, D. Xu, J. Wong, S. Nasiriany, C. Wang, R. Kulkarni, L. Fei-Fei, S. Savarese, Y. Zhu, and R. Mart’in-Mart’in, “What matters in learning from offline human demonstrations for robot manipulation,” in Conference on Robot Learning , 2021. [Online]. Available: https://api.semanticscholar.org/CorpusID:236956615
2021
Cited alongside, same era.
2022
Later among the works it cites.
A. Reichlin, G. L. Marchetti, H. Yin, A. Ghadirzadeh, and D. Kragic, “Back to the manifold: Recovering from out-of-distribution states,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2022, pp. 8660–8666
2022
Later among the works it cites.
H.-C. Wang, S.-F. Chen, M.-H. Hsu, C.-M. Lai, and S.-H. Sun, “Diffusion model-augmented behavioral cloning,” 2023
2023
Closest in time.
A. George, A. Bartsch, and A. B. Farimani, “Minimizing human assistance: Augmenting a single demonstration for deep reinforcement learning,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) , 2023, pp. 5027–5033
2023
Closest in time.
T. Z. Zhao, V. Kumar, S. Levine, and C. Finn, “Learning fine-grained bimanual manipulation with low-cost hardware,” 2023
2023
Closest in time.
H. Bharadhwaj, J. Vakil, M. Sharma, A. Gupta, S. Tulsiani, and V. Kumar, “Roboagent: Generalization and efficiency in robot manipulation via semantic augmentations and action chunking,” 2023
2023
Closest in time.
A. George, A. Bartsch, and A. B. Farimani, “Openvr: Teleoperation for manipulation,” 2023
2023
Closest in time.
A. Dikshit, A. Bartsch, A. George, and A. B. Farimani, “Robochop: Autonomous framework for fruit and vegetable chopping leveraging foundational models,” 2023
2023
Closest in time.