Fetching the paper…
Reading the bibliography…
Reinforcement learning methods can achieve significant performance but require a large amount of training data collected on the same robotic platform.
S. Karaman and E. Frazzoli, “Incremental sampling-based algorithms for optimal motion planning,” Robotics Science and Systems VI , 2010
2010
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” ArXiv:1312.6114 , 2013
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” The Journal of Machine Learning Research , 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman, “Deep predictive policy training using reinforcement learning,” in IROS , 2017
2017
Earlier work this paper cites.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” ICML , 2017
2017
Earlier work this paper cites.
Y. Duan, M. Andrychowicz, B. Stadie, O. J. Ho, J. Schneider, I. Sutskever, P. Abbeel, and W. Zaremba, “One-shot imitation learning,” in NeurIPS , 2017
2017
Earlier work this paper cites.
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine, “One-shot visual imitation learning via meta-learning,” in Conference on Robot Learning , 2017
2017
Earlier work this paper cites.
L. Pinto and A. Gupta, “Learning to push by grasping: Using multiple tasks for effective learning,” in ICRA , 2017
2017
Earlier work this paper cites.
C. Devin, A. Gupta, T. Darrell, P. Abbeel, and S. Levine, “Learning modular neural network policies for multi-task and multi-robot transfer,” in ICRA 2017
2017
Earlier work this paper cites.
C. Finn and S. Levine, “Deep visual foresight for planning robot motion,” in ICRA , 2017
2017
Earlier work this paper cites.
J. Snell, K. Swersky, and R. Zemel, “Prototypical networks for few-shot learning,” in NeurIPS , 2017
2017
Earlier work this paper cites.
C. Finn, K. Xu, and S. Levine, “Probabilistic model-agnostic meta-learning,” in NeurIPS , 2018
2018
Earlier work this paper cites.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in ICRA , 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
S. James, M. Bloesch, and A. J. Davison, “Task-embedded control networks for few-shot imitation learning,” in Conference on Robot Learning , 2018
2018
Cited alongside, same era.
F. Sadeghi, A. Toshev, E. Jang, and S. Levine, “Sim2real viewpoint invariant visual servoing by recurrent control,” in CVPR , 2018
2018
Cited alongside, same era.
K. Hausman, J. T. Springenberg, Z. Wang, N. Heess, and M. Riedmiller, “Learning an embedding space for transferable robot skills,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
A. Gupta, R. Mendonca, Y. Liu, P. Abbeel, and S. Levine, “Meta-reinforcement learning of structured exploration strategies,” in NeurIPS , 2018
2018
Cited alongside, same era.
M. Al-Shedivat, T. Bansal, Y. Burda, I. Sutskever, I. Mordatch, and P. Abbeel, “Continuous adaptation via meta-learning in nonstationary and competitive environments,” in ICLR , 2018
C. Schaff, D. Yunis, A. Chakrabarti, and M. R. Walter, “Jointly learning to construct and control agents using deep reinforcement learning,” in ICRA , 2019
2019
Later among the works it cites.
M. A. Jamal and G. Qi, “Task agnostic meta-learning for few-shot learning,” in CVPR , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2020
Later among the works it cites.
K. Arndt, M. Hazara, A. Ghadirzadeh, and V. Kyrki, “Meta reinforcement learning for sim-to-real domain adaptation,” in ICRA , 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
T. Chen, A. Murali, and A. Gupta, “Hardware conditioned policies for multi-robot transfer learning,” in NeurIPS , 2018
2018
Cited alongside, same era.
F. Sung, Y. Yang, L. Zhang, T. Xiang, P. H. S. Torr, and T. M. Hospedales, “Learning to compare: Relation network for few-shot learning,” in CVPR , 2018
2018
Cited alongside, same era.
E. Grant, C. Finn, S. Levine, T. Darrell, and T. Griffiths, “Recasting gradient-based meta-learning as hierarchical bayes,” in ICLR , 2018
2018
Cited alongside, same era.
J. Yoon, T. Kim, O. Dia, S. Kim, Y. Bengio, and S. Ahn, “Bayesian model-agnostic meta-learning,” in NeurIPS , 2018
2018
Cited alongside, same era.
J. Harrison, A. Sharma, and M. Pavone, “Meta-learning priors for efficient online bayesian regression,” in International Workshop on the Algorithmic Foundations of Robotics , 2018
2018
Cited alongside, same era.
I. Clavera, A. Nagabandi, S. Liu, R. S. Fearing, P. Abbeel, S. Levine, and C. Finn, “Learning to adapt in dynamic, real-world environments through meta-reinforcement learning,” in ICLR , 2019
2019
Cited alongside, same era.
S. Ravi and A. Beatson, “Amortized bayesian meta-learning,” in ICLR , 2019
2019
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
A. Bonardi, S. James, and A. J. Davison, “Learning one-shot imitation from humans without humans,” IEEE Robotics and Automation Letters , 2020
2020
Later among the works it cites.
J. Bütepage, A. Ghadirzadeh, Ö. Öztimur Karadag, M. Björkman, and D. Kragic, “Imitating by generating: Deep generative models for imitation of interactive tasks,” Frontiers in Robotics and AI , 2020
2020
Later among the works it cites.
X. Chen, A. Ghadirzadeh, M. Björkman, and P. Jensfelt, “Adversarial feature training for generalizable robotic visuomotor control,” in ICRA , 2020
2020
Later among the works it cites.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray et al. , “Learning dexterous in-hand manipulation,” The International Journal of Robotics Research , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
W. Huang, I. Mordatch, and D. Pathak, “One policy to control them all: Shared modular policies for agent-agnostic control,” ICML , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Zou and X. Lu, “Gradient-em bayesian meta-learning,” ArXiv:2006.11764 , 2020
2020
Later among the works it cites.