Fetching the paper…
Reading the bibliography…
In order for a robot to be a generalist that can perform a wide range of jobs, it must be able to acquire a wide variety of skills quickly and efficiently in complex unstructured environments.
ALVINN: an autonomous land vehicle in a neural network
D. Pomerleau · 1989
Earlier work this paper cites.
Learning to learn
S. Thrun and L. Pratt · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng, S. J. Russell, et al · 2000
Earlier work this paper cites.
Computational approaches to motor learning by imitation
S. Schaal, A. Ijspeert, and A. Billard · 2003
Earlier work this paper cites.
Discovering optimal imitation strategies
A. Billard, Y. Epars, S. Calinon, S. Schaal, and G. Cheng · 2004
Earlier work this paper cites.
Learning movement primitives
S. Schaal, J. Peters, J. Nakanishi, and A. Ijspeert · 2005
Earlier work this paper cites.
Imitation learning for locomotion and manipulation
N. Ratliff, J. A. Bagnell, and S. S. Srinivasa · 2007
Earlier work this paper cites.
Learning and generalization of motor skills by learning from demonstration
P. Pastor, H. Hoffmann, T. Asfour, and S. Schaal · 2009
Earlier work this paper cites.
Transfer learning for reinforcement learning on a physical robot
S. Barrett, M. E. Taylor, and P. Stone · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
S. Ross, G. J. Gordon, and D. Bagnell · 2011
Earlier work this paper cites.
Reinforcement learning to adjust parametrized motor primitives to new situations
J. Kober, A. Wilhelm, E. Oztop, and J. Peters · 2012
Earlier work this paper cites.
Learning parameterized skills
B. C. Da Silva, G. Konidaris, and A. G. Barto · 2012
Cited alongside, same era.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Cited alongside, same era.
Data-efficient generalization of robot skills with contextual policy search
A. G. Kupcsik, M. P. Deisenroth, J. Peters, G. Neumann, et al · 2013
Cited alongside, same era.
Learning compact parameterized skills with a single regression
F. Stulp, G. Raiola, A. Hoarau, S. Ivaldi, and O. Sigaud · 2013
Cited alongside, same era.
Learning to select and generalize striking movements in robot table tennis
K. Mülling, J. Kober, O. Kroemer, and J. Peters · 2013
Cited alongside, same era.
Multi-task policy search for robotics
M. P. Deisenroth, P. Englert, J. Peters, and D. Fox · 2014
Cited alongside, same era.
Guided cost learning: Deep inverse optimal control via policy optimization
C. Finn, S. Levine, and P. Abbeel · 2016
Later among the works it cites.
End-to-end learning of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Later among the works it cites.
Deep spatial autoencoders for visuomotor learning
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel · 2016
Later among the works it cites.
J. L. Ba, J. R. Kiros, and G. E. Hinton · 2016
Later among the works it cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Universal value function approximators
T. Schaul, D. Horgan, K. Gregor, and D. Silver · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2015
Cited alongside, same era.
End to end learning for self-driving cars
M. Bojarski, D. Del Testa, D. Dworakowski, B. Firner, B. Flepp, P. Goyal, L. Jackel, M. Monfort, U. Muller, J. Zhang, X. Zhang, J. Zhao, and K. Zieba · 2016
Cited alongside, same era.
A machine learning approach to visual perception of forest trails for mobile robots
A. Giusti, J. Guzzi, D. C. Cireşan, F.-L. He, J. P. Rodríguez, F. Fontana, M. Faessler, C. Forster, J. Schmidhuber, G. Di Caro, et al · 2016
Cited alongside, same era.
Shiv: Reducing supervisor burden in dagger using support vectors for efficient learning from demonstrations in high dimensional state spaces
M. Laskey, S. Staszak, W. Y.-S. Hsieh, J. Mahler, F. T. Pokorny, A. D. Dragan, and K. Goldberg · 2016
Cited alongside, same era.
Y. Duan, M. Andrychowicz, B. Stadie, J. Ho, J. Schneider, I. Sutskever, P. Abbeel, and W. Zaremba · 2017
Closest in time.
Query-efficient imitation learning for end-to-end simulated driving
J. Zhang and K. Cho · 2017
Closest in time.
Combining self-supervised learning and imitation for vision-based rope manipulation
A. Nair, P. Agarwal, D. Chen, P. Isola, P. Abbeel, and S. Levine · 2017
Closest in time.
Unsupervised perceptual rewards for imitation learning
P. Sermanet, K. Xu, and S. Levine · 2017
Closest in time.
Learning invariant feature spaces to transfer skills with reinforcement learning
A. Gupta, C. Devin, Y. Liu, P. Abbeel, and S. Levine · 2017
Closest in time.
Model-agnostic meta-learning for fast adaptation of deep networks
C. Finn, P. Abbeel, and S. Levine · 2017
Closest in time.