Fetching the paper…
Reading the bibliography…
Acquiring multiple skills has commonly involved collecting a large number of expert demonstrations per task or engineering custom reward functions.
R. S. Sutton, ”Integrated architectures for learning, planning, and reacting based on approximating dynamic programming”, Machine learning proceedings, 1990
1990
Earlier work this paper cites.
J. Schmidhuber, ”A possibility for implementing curiosity and boredom in model-building neural controllers”, Proceedings of the first international conference on simulation of adaptive behavior, 1991
1991
Earlier work this paper cites.
L. P. Kaelbling, ”Learning to Achieve Goals”, Proceedings of the Thirteenth International Joint Conference on Artificial Intelligence, 1993
1993
Earlier work this paper cites.
S. Schaal, ”Is imitation learning the route to humanoid robots?”, Trends in Cognitive Sciences, 1999
1999
Earlier work this paper cites.
M. N. Nicolescu , M. J. Mataric, ”Natural methods for robot task learning: instructive demonstrations, generalization and practice”, Proceedings of the second international joint conference on Autonomous agents and multiagent systems, 2003
2003
Earlier work this paper cites.
J. J. Steil, F. Röthling, R. Haschke, H. Ritter, ”Situated robot learning for multi-modal instruction and imitation of grasping”, Robotics and Autonomous Systems, 2004
2004
Earlier work this paper cites.
K. Hsiao, T. Lozano-Perez, ”Imitation Learning of Whole-Body Grasps”, IEEE/RSJ International Conference on Intelligent Robots and Systems, 2006
2006
Earlier work this paper cites.
N. Ratliff, J. A. Bagnell, S. S. Srinivasa, ”Imitation learning for locomotion and manipulation”, 7th IEEE-RAS International Conference on Humanoid Robots, 2007
2007
Earlier work this paper cites.
B. D. Argalla, S. Chernovab, M. Velosob, B. Browning, ”A survey of robot learning from demonstration”, Robotics and Autonomous Systems, 2009
2009
Earlier work this paper cites.
P. Pastor, H. Hoffmann, T. Asfour, S. Schaal, ”Learning and generalization of motor skills by learning from demonstration”, IEEE International Conference on Robotics and Automation, 2009
2009
Earlier work this paper cites.
2010
Earlier work this paper cites.
J. Schmidhuber, ”Formal Theory of Creativity & Fun & Intrinsic Motivation (1990-2010)”, IEEE Transactions on Autonomous Mental Development, 2010
2010
Earlier work this paper cites.
M. P. Deisenroth, G. Neumann, J. Peters, ”A survey on policy search for robotics”, Foundations and Trends in Robotics, 2013
2013
Earlier work this paper cites.
D. P. Kingma, M Welling, ”Auto-Encoding Variational Bayes”, arXiv:1312.6114, 2013
2013
Earlier work this paper cites.
J. Kober, J. A. Bagnell, J. Peters, ”Reinforcement learning in robotics: A survey”, The International Journal of Robotics Research, 2013
2013
Earlier work this paper cites.
2015
Cited alongside, same era.
K. Sohn, H. Lee, X. Yan, ”Learning Structured Output Representation using Deep Conditional Generative Models”, Advances in Neural Information Processing Systems 28, 2015
2015
Cited alongside, same era.
C. Doersch, ”Tutorial on Variational Autoencoders”, arXiv:1606.05908, 2016
2016
Cited alongside, same era.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, D. Hassabis, ”Mastering the game of Go with deep neural networks and tree search”, Nature, 2016
2016
Cited alongside, same era.
D. Antotsiou, G. Garcia-Hernando, T. Kim, ”Task-oriented hand motion retargeting for dexterous manipulation imitation”, The European Conference on Computer Vision (ECCV) Workshops, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
D. Pathak, P. Agrawal, A. A. Efros, T. Darrell, ”Curiosity-driven exploration by Self-supervised prediction”, IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
”Overcoming Exploration in Reinforcement Learning with Demonstrations”, IEEE International Conference on Robotics and Automation (ICRA), 2018
2018
Later among the works it cites.
J. Oh, Y. Guo, S. Singh, H. Lee, ”Self-Imitation Learning”, arXiv:1806.05635, 2018
2018
Later among the works it cites.
D. Russo, B. Van Roy, A. Kazerouni, I. Osband, Z. Wen, ”A tutorial on thompson sampling”, Foundations and Trends in Machine Learning, 2018
2018
Later among the works it cites.
R. S. Sutton, A. G. Barto, ”Reinforcement Learning: An Introduction”, MIT Press, 2018
2018
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
A. Gupta, V. Kumar, C. Lynch, S. Levine, K. Hausman ”Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning”, Conference on Robot Learning (CoRL) 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, P. Sermanet, ”Learning Latent Plans from Play”, Conference on Robot Learning (CoRL), 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever, ”Language models are unsupervised multitask learners”, OpenAI Blog, 2019
2019
Later among the works it cites.