Fetching the paper…
Reading the bibliography…
This paper focuses on transferring control policies between robot manipulators with different morphology.
N. Das, S. Bechtle, T. Davchev, D. Jayaraman, A. Rai, and F. Meier, “Model-based inverse reinforcement learning from visual demonstrations,” in
1942
Earlier work this paper cites.
M. Bain and C. Sammut, “A framework for behavioural cloning.” in
1995
Earlier work this paper cites.
M. Iacoboni, R. Woods, M. Brass, H. Bekkering, J. Mazziotta, and G. Rizzolatti, “Cortical mechanisms of human imitation,”
1999
Earlier work this paper cites.
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in
1999
Earlier work this paper cites.
G. Rizzolatti and L. Craighero, “The mirror-neuron system,”
2004
Earlier work this paper cites.
Y. LeCun, S. Chopra, R. Hadsell, M. Ranzato, and F. Huang, “A tutorial on energy-based learning,”
2006
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in
2011
Earlier work this paper cites.
N. Ferns and D. Precup, “Bisimulation metrics are optimal value functions.” in
2014
Earlier work this paper cites.
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller, “Embed to control: A locally linear latent dynamics model for control from raw images,”
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
B. Stadie, P. Abbeel, and I. Sutskever, “Third-person imitation learning,”
2017
Earlier work this paper cites.
A. Gupta, C. Devin, Y. Liu, P. Abbeel, and S. Levine, “Learning invariant feature spaces to transfer skills with reinforcement learning,”
2017
Earlier work this paper cites.
M. Wulfmeier, I. Posner, and P. Abbeel, “Mutual alignment transfer learning,” in
2017
Earlier work this paper cites.
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros, “Unpaired image-to-image translation using cycle-consistent adversarial networks,” in
2017
Earlier work this paper cites.
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel, “Deep imitation learning for complex manipulation tasks from virtual reality teleoperation,” in
2018
Earlier work this paper cites.
P. Sermanet, C. Lynch, Y. Chebotar, J. Hsu, E. Jang, S. Schaal, S. Levine, and G. Brain, “Time-contrastive networks: Self-supervised learning from video,” in
2018
Cited alongside, same era.
D. Ha and J. Schmidhuber, “World models,”
2018
Cited alongside, same era.
2018
Cited alongside, same era.
T. Kurutach, A. Tamar, G. Yang, S. J. Russell, and P. Abbeel, “Learning plannable representations with causal infogan,”
2018
Cited alongside, same era.
S. Fujimoto, H. Hoof, and D. Meger, “Addressing function approximation error in actor-critic methods,” in
2018
Cited alongside, same era.
K. Rao, C. Harris, A. Irpan, S. Levine, J. Ibarz, and M. Khansari, “RL-CycleGAN: Reinforcement learning aware simulation-to-real,” in
2020
Later among the works it cites.
2020
Later among the works it cites.
A. X. Lee, A. Nagabandi, P. Abbeel, and S. Levine, “Stochastic latent actor-critic: Deep reinforcement learning with a latent variable model,” in
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Pari, N. M. Shafiullah, S. P. Arunachalam, and L. Pinto, “The surprising effectiveness of representation learning for visual imitation,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Florence, L. Manuelli, and R. Tedrake, “Self-supervised correspondence in visuomotor policy learning,”
2019
Cited alongside, same era.
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson, “Learning latent dynamics for planning from pixels,” in
2019
Cited alongside, same era.
A. Anand, E. Racah, S. Ozair, Y. Bengio, M.-A. Côté, and R. D. Hjelm, “Unsupervised state representation learning in atari,”
2019
Cited alongside, same era.
A. Wang, T. Kurutach, K. Liu, P. Abbeel, and A. Tamar, “Learning robotic manipulation through visual planning and acting,”
2019
Cited alongside, same era.
L. Kaiser, M. Babaeizadeh, P. Miłos, B. Osiński, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, P. Kozakowski, S. Levine, A. Mohiuddin, R. Sepassi, G. Tucker, and H. Michalewski, “Model based reinforcement learning for Atari,” in
2020
Cited alongside, same era.
Z. D. Guo, B. A. Pires, B. Piot, J.-B. Grill, F. Altché, R. Munos, and M. G. Azar, “Bootstrap latent-predictive representations for multitask reinforcement learning,” in
2020
Cited alongside, same era.
J.-B. Grill, F. Strub, F. Altché, C. Tallec, P. Richemond, E. Buchatskaya, C. Doersch, B. Avila Pires, Z. Guo, M. G. Azar, B. Piot, K. Kavukcuoglu, R. Munos, and M. Valko, “Bootstrap your own latent-a new approach to self-supervised learning,”
2020
Cited alongside, same era.
2021
Later among the works it cites.
Q. Zhang, T. Xiao, A. Efros, L. Pinto, and X. Wang, “Learning cross-domain correspondence for control with dynamics cycle-consistency,”
2021
Later among the works it cites.
B. Eysenbach, T. Zhang, S. Levine, and R. R. Salakhutdinov, “Contrastive learning as goal-conditioned reinforcement learning,”
2022
Later among the works it cites.
Y. LeCun, “A path towards autonomous machine intelligence version 0.9. 2, 2022-06-27,”
2022
Later among the works it cites.
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta, “R3m: A universal visual representation for robot manipulation,”
2022
Later among the works it cites.
S. Nair, E. Mitchell, K. Chen, S. Savarese, C. Finn,
2022
Later among the works it cites.
K. Zakka, A. Zeng, P. Florence, J. Tompson, J. Bohg, and D. Dwibedi, “Xirl: Cross-embodiment inverse reinforcement learning,” in
2022
Later among the works it cites.
Z.-H. Yin, L. Sun, H. Ma, M. Tomizuka, and W.-J. Li, “Cross domain robot imitation with invariant representation,” in
2022
Later among the works it cites.
T. Yoneda, G. Yang, M. R. Walter, and B. Stadie, “Invariance through latent alignment,”
2022
Later among the works it cites.
2023
Later among the works it cites.