Fetching the paper…
Reading the bibliography…
Recent advances in Behavior Cloning (BC) have made it easy to teach robots new tasks.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,” nature , vol. 323, no. 6088, pp. 533–536, 1986
1986
Earlier work this paper cites.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” Advances in neural information processing systems , vol. 1, 1988
1988
Earlier work this paper cites.
S. Schaal, “Learning from Demonstration,” in Advances in Neural Information Processing Systems , vol. 9. MIT Press, 1996. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/1996/hash/68d13cf26c4b4f4f932e3eff990093ba-Abstract.html
1996
Earlier work this paper cites.
S. Schaal, “Learning from demonstration,” Advances in neural information processing systems , vol. 9, 1996
1996
Earlier work this paper cites.
——, “Is imitation learning the route to humanoid robots?” Trends in cognitive sciences , vol. 3, no. 6, pp. 233–242, 1999
1999
Earlier work this paper cites.
N. Ratliff, J. A. Bagnell, and S. S. Srinivasa, “Imitation learning for locomotion and manipulation,” in 2007 7th IEEE-RAS international conference on humanoid robots . IEEE, 2007, pp. 392–397
2007
Earlier work this paper cites.
S. Ross and D. Bagnell, “Efficient reductions for imitation learning,” in Proceedings of the thirteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2010, pp. 661–668
2010
Earlier work this paper cites.
J. Kober and J. Peters, “Imitation and reinforcement learning,” IEEE Robotics & Automation Magazine , vol. 17, no. 2, pp. 55–62, 2010
2010
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 627–635
2011
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 627–635
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International conference on machine learning . PMLR, 2015, pp. 1889–1897
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
F. Suárez-Ruiz and Q.-C. Pham, “A framework for fine robotic assembly,” in 2016 IEEE international conference on robotics and automation (ICRA) . IEEE, 2016, pp. 421–426
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
P. Agrawal, Computational sensorimotor learning . University of California, Berkeley, 2018
2018
Earlier work this paper cites.
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel, “Deep imitation learning for complex manipulation tasks from virtual reality teleoperation,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 5628–5635
2018
Earlier work this paper cites.
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine, “Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations,” in Proceedings of Robotics: Science and Systems (RSS) , 2018
2018
Earlier work this paper cites.
A. Nair, B. McGrew, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Overcoming exploration in reinforcement learning with demonstrations,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 6292–6299
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Ajay, J. Wu, N. Fazeli, M. Bauza, L. P. Kaelbling, J. B. Tenenbaum, and A. Rodriguez, “Augmenting physical simulators with stochastic neural networks: Case study of planar pushing and bouncing,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2018, pp. 3066–3073
2018
Earlier work this paper cites.
A. Reuther, J. Kepner, C. Byun, S. Samsi, W. Arcand, D. Bestor, B. Bergeron, V. Gadepally, M. Houle, M. Hubbell, M. Jones, A. Klein, L. Milechin, J. Mullen, A. Prout, A. Rosa, C. Yee, and P. Michaleas, “Interactive Supercomputing on 40,000 Cores for Machine Learning and Data Analysis,” in 2018 IEEE High Performance extreme Computing Conference (HPEC) , Sep. 2018, pp. 1–6, iSSN: 2377-6943. [Online]. Available: https://ieeexplore.ieee.org/document/8547629
2018
Earlier work this paper cites.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. A. Ojea, E. Solowjow, and S. Levine, “Residual reinforcement learning for robot control,” in 2019 international conference on robotics and automation (ICRA) . IEEE, 2019, pp. 6023–6029
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Zhou, C. Barnes, J. Lu, J. Yang, and H. Li, “On the continuity of rotation representations in neural networks,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 5745–5753
2019
Earlier work this paper cites.
K. Kimble, K. Van Wyk, J. Falco, E. Messina, Y. Sun, M. Shibata, W. Uemura, and Y. Yokokohji, “Benchmarking protocols for evaluating small parts robotic assembly systems,” IEEE robotics and automation letters , vol. 5, no. 2, pp. 883–889, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science robotics , vol. 5, no. 47, p. eabc5986, 2020
2020
Earlier work this paper cites.
A. Kumar, A. Zhou, G. Tucker, and S. Levine, “Conservative q-learning for offline reinforcement learning,” Advances in Neural Information Processing Systems , vol. 33, pp. 1179–1191, 2020
2020
Earlier work this paper cites.
A. Zeng, S. Song, J. Lee, A. Rodriguez, and T. Funkhouser, “Tossingbot: Learning to throw arbitrary objects with residual physics,” IEEE Transactions on Robotics , vol. 36, no. 4, pp. 1307–1319, 2020
2020
Earlier work this paper cites.
G. Schoettler, A. Nair, J. Luo, S. Bahl, J. A. Ojea, E. Solowjow, and S. Levine, “Deep reinforcement learning for industrial insertion tasks with visual inputs and natural rewards,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2020, pp. 5548–5555
2020
Earlier work this paper cites.
J. Levinson, C. Esteves, K. Chen, N. Snavely, A. Kanazawa, A. Rostamizadeh, and A. Makadia, “An analysis of svd for deep rotation estimation,” Advances in Neural Information Processing Systems , vol. 33, pp. 22 554–22 565, 2020
2020
Cited alongside, same era.
Y. Lee, E. S. Hu, and J. J. Lim, “Ikea furniture assembly environment for long-horizon complex manipulation tasks,” in 2021 ieee international conference on robotics and automation (icra) . IEEE, 2021, pp. 6343–6349
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2023
Later among the works it cites.
R. Ramrakhya, D. Batra, E. Wijmans, and A. Das, “Pirlnav: Pretraining with imitation and rl finetuning for objectnav,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 17 896–17 906
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
T. G. Rudner, C. Lu, M. A. Osborne, Y. Gal, and Y. Teh, “On pathologies in kl-regularized reinforcement learning from expert demonstrations,” Advances in Neural Information Processing Systems , vol. 34, pp. 28 376–28 389, 2021
2021
Cited alongside, same era.
L. Chen, K. Lu, A. Rajeswaran, K. Lee, A. Grover, M. Laskin, P. Abbeel, A. Srinivas, and I. Mordatch, “Decision transformer: Reinforcement learning via sequence modeling,” Advances in neural information processing systems , vol. 34, pp. 15 084–15 097, 2021
2021
Cited alongside, same era.
O. Spector and D. Di Castro, “Insertionnet-a scalable solution for insertion,” IEEE Robotics and Automation Letters , vol. 6, no. 3, pp. 5509–5516, 2021
2021
Cited alongside, same era.
2022
Cited alongside, same era.
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn, “Bc-z: Zero-shot task generalization with robotic imitation learning,” in Conference on Robot Learning . PMLR, 2022, pp. 991–1002
2022
Cited alongside, same era.
2022
Cited alongside, same era.
L. Lai, A. Z. Huang, and S. J. Gershman, “Action chunking as policy compression,” PsyArXiv , 2022
2022
Cited alongside, same era.
T. Chen, M. Tippur, S. Wu, V. Kumar, E. Adelson, and P. Agrawal, “Visual dexterity: In-hand reorientation of novel and complex object shapes,” Science Robotics , vol. 8, no. 84, p. eadc9244, 2023
2023
Later among the works it cites.
M. Mittal, C. Yu, Q. Yu, J. Liu, N. Rudin, D. Hoeller, J. L. Yuan, R. Singh, Y. Guo, H. Mazhar, A. Mandlekar, B. Babich, G. State, M. Hutter, and A. Garg, “Orbit: A unified simulation framework for interactive robot learning environments,” IEEE Robotics and Automation Letters , vol. 8, no. 6, pp. 3740–3747, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Fan, O. Watkins, Y. Du, H. Liu, M. Ryu, C. Boutilier, P. Abbeel, M. Ghavamzadeh, K. Lee, and K. Lee, “Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models,” 2023
2023
Later among the works it cites.
L. Yang, Z. Huang, F. Lei, Y. Zhong, Y. Yang, C. Fang, S. Wen, B. Zhou, and Z. Lin, “Policy representation via diffusion probability model for reinforcement learning,” 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
E. Kaufmann, L. Bauersfeld, A. Loquercio, M. Müller, V. Koltun, and D. Scaramuzza, “Champion-level drone racing using deep reinforcement learning,” Nature , vol. 620, no. 7976, pp. 982–987, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
B. Tang, M. A. Lin, I. Akinola, A. Handa, G. S. Sukhatme, F. Ramos, D. Fox, and Y. Narang, “Industreal: Transferring contact-rich assembly tasks from simulation to reality,” in Robotics: Science and Systems , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
M. Drolet, S. Stepputtis, S. Kailas, A. Jain, J. Peters, S. Schaal, and H. B. Amor, “A comparison of imitation learning algorithms for bimanual manipulation,” IEEE Robotics and Automation Letters , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
A. Yu, G. Yang, R. Choi, Y. Ravan, J. Leonard, and P. Isola, “Lucidsim: Learning agile visual locomotion from generated images,” in 8th Annual Conference on Robot Learning , 2024
2024
Closest in time.
T. Z. Zhao, J. Tompson, D. Driess, P. Florence, S. K. S. Ghasemipour, C. Finn, and A. Wahid, “ALOHA unleashed: A simple recipe for robot dexterity,” in 8th Annual Conference on Robot Learning , 2024. [Online]. Available: https://openreview.net/forum?id=gvdXE7ikHI
2024
Closest in time.
2024
Closest in time.
M. Nakamoto, S. Zhai, A. Singh, M. Sobol Mark, Y. Ma, C. Finn, A. Kumar, and S. Levine, “Cal-ql: Calibrated offline rl pre-training for efficient online fine-tuning,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
M. T. Villasevil, A. Jain, V. Macha, J. Yuan, L. L. Ankile, A. Simeonov, P. Agrawal, and A. Gupta, “Scaling robot-learning by crowdsourcing simulation environments,” in RSS 2024 Workshop: Data Generation for Robotics , 2024
2024
Closest in time.
2024
Closest in time.
R. Tedrake, Robotic Manipulation . Course Notes for MIT 6.421, 2024. [Online]. Available: http://manipulation.mit.edu
2024
Closest in time.
S. Han, I. Shenfeld, A. Srivastava, Y. Kim, and P. Agrawal, “Value augmented sampling for language model alignment and personalization,” 2024
2024
Closest in time.
Z. Li, R. Krohn, T. Chen, A. Ajay, P. Agrawal, and G. Chalvatzaki, “Learning multimodal behaviors from scratch with diffusion policy gradient,” 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
A. R. Geist, J. Frey, M. Zobro, A. Levina, and G. Martius, “Learning with 3d rotations, a hitchhiker’s guide to so(3),” 2024
2024
Closest in time.