Fetching the paper…
Reading the bibliography…
We investigate the visual cross-embodiment imitation setting, in which agents learn policies from videos of other agents (such as humans) demonstrating the same task, but with stark differences in their embodiments -- shape, actions, end-effector dynamics, etc.
Robot learning from demonstration
C. Atkeson and S. Schaal · 1997
Earlier work this paper cites.
Pymunk: A easy-to-use pythonic rigid body 2d physics library (version 6.0.0) , 2007
V. Blomqvist · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Modeling interaction via the principle of maximum causal entropy
B. D. Ziebart, J. A. Bagnell, and A. K. Dey · 2010
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A. M. Saxe, J. L. McClelland, and S. Ganguli · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Generative adversarial networks
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Earlier work this paper cites.
Deep reinforcement learning with double q-learning
H. van Hasselt, A. Guez, and D. Silver · 2015
Earlier work this paper cites.
Unsupervised perceptual rewards for imitation learning
P. Sermanet, K. Xu, and S. Levine · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
J. Ho and S. Ermon · 2016
Earlier work this paper cites.
Learning invariant feature spaces to transfer skills with reinforcement learning, 2017
A. Gupta, C. Devin, Y. Liu, P. Abbeel, and S. Levine · 2017
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros · 2017
Earlier work this paper cites.
Third-person imitation learning
B. C. Stadie, P. Abbeel, and I. Sutskever · 2017
Earlier work this paper cites.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Earlier work this paper cites.
Time-contrastive networks: Self-supervised learning from video
P. Sermanet, C. Lynch, Y. Chebotar, J. Hsu, E. Jang, S. Schaal, and S. Levine · 2018
Cited alongside, same era.
An algorithmic perspective on imitation learning
T. Osa, J. Pajarinen, G. Neumann, J. Bagnell, P. Abbeel, and J. Peters · 2018
Cited alongside, same era.
Behavioral cloning from observation
F. Torabi, G. Warnell, and P. Stone · 2018
Cited alongside, same era.
Imitating latent policies from observation
A. D. Edwards, H. Sahni, Y. Schroecker, and C. L. Isbell · 2018
Cited alongside, same era.
Zero-shot visual imitation
D. Pathak, P. Mahmoudieh, G. Luo, P. Agrawal, D. Chen, Y. Shentu, E. Shelhamer, J. Malik, A. A. Efros, and T. Darrell · 2018
Cited alongside, same era.
Playing hard exploration games by watching youtube
Y. Aytar, T. Pfaff, D. Budden, T. Paine, Z. Wang, and N. de Freitas · 2018
Third-person visual imitation learning via decoupled hierarchical controller
P. Sharma, D. Pathak, and A. Gupta · 2019
Later among the works it cites.
A practical approach to insertion with variable socket position using deep reinforcement learning
M. Vecerik, O. Sushkov, D. Barker, T. Rothörl, T. Hester, and J. Scholz · 2019
Later among the works it cites.
Imitation learning from observations by minimizing inverse dynamics disagreement
C. Yang, X. Ma, W. Huang, F. Sun, H. Liu, J. Huang, and C. Gan · 2019
Later among the works it cites.
Visual imitation learning with recurrent siamese networks
G. Berseth and C. Pal · 2019
Later among the works it cites.
Adail: Adaptive adversarial imitation learning
Y. Lu and J. Tompson · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Imitation from observation: Learning to imitate behaviors from raw video via context translation
Y. Liu, A. Gupta, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Deep reinforcement learning that matters
P. Henderson, R. Islam, P. Bachman, J. Pineau, D. Precup, and D. Meger · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Soft actor-critic algorithms and applications
T. Haarnoja, A. Zhou, K. Hartikainen, G. Tucker, S. Ha, J. Tan, V. Kumar, H. Zhu, A. Gupta, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Generative adversarial imitation from observation
F. Torabi, G. Warnell, and P. Stone · 2018
Cited alongside, same era.
Albumentations: fast and flexible image augmentations
A. Buslaev, A. Parinov, E. Khvedchenya, V. Iglovikov, and A. A. Kalinin · 2018
Cited alongside, same era.
F. Liu, Z. Ling, T. Mu, and H. Su · 2019
Later among the works it cites.
Provably efficient imitation learning from observation alone
W. Sun, A. Vemula, B. Boots, and J. Bagnell · 2019
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
Later among the works it cites.
Reinforcement learning with videos: Combining offline observations with interaction, 2020
K. Schmeckpeper, O. Rybkin, K. Daniilidis, S. Levine, and C. Finn · 2020
Later among the works it cites.
State-only imitation learning for dexterous manipulation
I. Radosavovic, X. Wang, L. Pinto, and J. Malik · 2020
Later among the works it cites.
Adversarial skill networks: Unsupervised robot skill learning from video
O. Mees, M. Merklinger, G. Kalweit, and W. Burgard · 2020
Later among the works it cites.
Hierarchically decoupled imitation for morphological transfer
D. Hejna, L. Pinto, and P. Abbeel · 2020
Later among the works it cites.
The magical benchmark for robust imitation
S. Toyer, R. Shah, A. Critch, and S. Russell · 2020
Later among the works it cites.
Concept2robot: Learning manipulation concepts from instructions and human demonstrations
L. Shao, T. Migimatsu, Q. Zhang, K. Yang, and J. Bohg · 2020
Later among the works it cites.
Domain adaptive imitation learning
K. Kim, Y. Gu, J. Song, S. Zhao, and S. Ermon · 2020
Later among the works it cites.
Soft actor-critic (sac) implementation in pytorch
D. Yarats and I. Kostrikov · 2020
Later among the works it cites.
A simple framework for contrastive learning of visual representations
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton · 2020
Later among the works it cites.