Fetching the paper…
Reading the bibliography…
This paper considers learning robot locomotion and manipulation tasks from expert demonstrations.
Learning from demonstration
S. Schaal · 1996
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng and S. Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
Maximum margin planning
N. D. Ratliff, J. A. Bagnell, and M. A. Zinkevich · 2006
Earlier work this paper cites.
Apprenticeship learning using inverse reinforcement learning and gradient methods
G. Neu and C. Szepesvári · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. Maas, J. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Learning and generalization of motor skills by learning from demonstration
P. Pastor, H. Hoffmann, T. Asfour, and S. Schaal · 2009
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
ADAM: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Learning structured output representation using deep conditional generative models
K. Sohn, H. Lee, and X. Yan · 2015
Earlier work this paper cites.
Gradient estimation using stochastic computation graphs
J. Schulman, N. Heess, T. Weber, and P. Abbeel · 2015
Earlier work this paper cites.
Generative adversarial imitation learning
J. Ho and S. Ermon · 2016
Earlier work this paper cites.
C. Finn, P. Christiano, P. Abbeel, and S. Levine · 2016
Cited alongside, same era.
beta-vae: Learning basic visual concepts with a constrained variational framework
I. Higgins, L. Matthey, A. Pal, C. Burgess, X. Glorot, M. Botvinick, S. Mohamed, and A. Lerchner · 2016
Cited alongside, same era.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
Inverse reward design
D. Hadfield-Menell, S. Milli, P. Abbeel, S. J. Russell, and A. Dragan · 2017
Cited alongside, same era.
Imitation learning: A survey of learning methods
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne · 2017
Cited alongside, same era.
Task-relevant adversarial imitation learning
K. Zolna, S. Reed, A. Novikov, S. G. Colmenarejo, D. Budden, S. Cabi, M. Denil, N. de Freitas, and Z. Wang · 2019
Later among the works it cites.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Later among the works it cites.
Imitating latent policies from observation
A. Edwards, H. Sahni, Y. Schroecker, and C. Isbell · 2019
Later among the works it cites.
Off-policy deep reinforcement learning without exploration
S. Fujimoto, D. Meger, and D. Precup · 2019
Later among the works it cites.
Sample-efficient imitation learning via generative adversarial nets
L. Blondé and A. Kalousis · 2019
Later among the works it cites.
Stable baselines3, 2019
A. Raffin, A. Hill, M. Ernestus, A. Gleave, A. Kanervisto, and N. Dormann · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Fu, K. Luo, and S. Levine · 2017
Cited alongside, same era.
Improved training of wasserstein gans
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A. C. Courville · 2017
Cited alongside, same era.
Prediction and control with temporal segment models
N. Mishra, P. Abbeel, and I. Mordatch · 2017
Cited alongside, same era.
Robust imitation of diverse behaviors
Z. Wang, J. S. Merel, S. E. Reed, N. de Freitas, G. Wayne, and N. Heess · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer · 2017
Cited alongside, same era.
Reinforcement and imitation learning for diverse visuomotor skills
Y. Zhu, Z. Wang, J. Merel, A. Rusu, T. Erez, S. Cabi, S. Tunyasuvunakool, J. Kramár, R. Hadsell, N. de Freitas, et al · 2018
Cited alongside, same era.
Spectral normalization for generative adversarial networks
T. Miyato, T. Kataoka, M. Koyama, and Y. Yoshida · 2018
Cited alongside, same era.
Later among the works it cites.
Stochastic latent actor-critic: Deep reinforcement learning with a latent variable model
A. X. Lee, A. Nagabandi, P. Abbeel, and S. Levine · 2020
Later among the works it cites.
Learning belief representations for imitation learning in pomdps
T. Gangwani, J. Lehman, Q. Liu, and J. Peng · 2020
Later among the works it cites.
Plas: Latent action space for offline reinforcement learning
W. Zhou, S. Bajracharya, and D. Held · 2020
Later among the works it cites.
Plannable approximations to mdp homomorphisms: Equivariance under actions
E. van der Pol, T. Kipf, F. A. Oliehoek, and M. Welling · 2020
Later among the works it cites.
A divergence minimization perspective on imitation learning methods
S. K. S. Ghasemipour, R. Zemel, and S. Gu · 2020
Later among the works it cites.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, and R. Martín-Martín · 2020
Later among the works it cites.
What matters for adversarial imitation learning?
M. Orsini, A. Raichuk, L. Hussenot, D. Vincent, R. Dadashi, S. Girgin, M. Geist, O. Bachem, O. Pietquin, and M. Andrychowicz · 2021
Later among the works it cites.
Visual adversarial imitation learning using variational models
R. Rafailov, T. Yu, A. Rajeswaran, and C. Finn · 2021
Later among the works it cites.
Laser: Learning a latent action space for efficient reinforcement learning
A. Allshire, R. Martín-Martín, C. Lin, S. Manuel, S. Savarese, and A. Garg · 2021
Later among the works it cites.