Fetching the paper…
Reading the bibliography…
State representation learning (SRL) in partially observable Markov decision processes has been studied to learn abstract features of data useful for robot control tasks.
C. E. Garcia, D. M. Prett, and M. Morari, “Model predictive control: Theory and practice - a survey,” in Automatica
1989
Earlier work this paper cites.
S. Schaal, “Is imitation learning the route to humanoid robots?,” in Trends in Cognitive Sciences
1999
Earlier work this paper cites.
P.-T. Boer, D. Kroese, S. Mannor, and R. Rubinstein, “A tutorial on the cross-entropy method,” Annals of Operations Research
2005
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in IROS
2012
Earlier work this paper cites.
E. Tzeng, J. Hoffman, N. Zhang, K. Saenko, and T. Darrell, “Deep domain confusion: Maximizing for domain invariance,” in arXiv
2014
Earlier work this paper cites.
D. Kingma, D. Rezende, S. Mohamed, and M. Welling, “Semi-supervised learning with deep generative models,” in NIPS
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in NIPS
2014
Earlier work this paper cites.
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. S. Lempitsky, “Domain-adversarial training of neural networks,” in J. Mach. Learn. Res
2015
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in NIPS
2016
Earlier work this paper cites.
K. Bousmalis, G. Trigeorgis, N. Silberman, D. Krishnan, and D. Erhan, “Domain separation networks,” in NIPS
2016
Earlier work this paper cites.
A. Gupta, C. Eppner, S. Levine, and P. Abbeel, “Learning dexterous manipulation for a soft robotic hand from human demonstrations,” in IROS
2016
Earlier work this paper cites.
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell, “Adversarial discriminative domain adaptation,” in CVPR
2017
Earlier work this paper cites.
N. Baram, O. Anschel, I. Caspi, and S. Mannor, “End-to-end differentiable adversarial imitation learning,” in ICML
2017
Earlier work this paper cites.
Y. Li, J. Song, and S. Ermon, “Infogail: Interpretable imitation learning from visual demonstrations,” in NIPS
2017
Earlier work this paper cites.
B. C. Stadie, P. Abbeel, and I. Sutskever, “Third-person imitation learning,” in ICLR
2017
Cited alongside, same era.
J. Merel, Y. Tassa, T. Dhruva, S. Srinivasan, J. Lemmon, Z. Wang, G. Wayne, and N. M. O. Heess, “Learning human behaviors from motion capture by adversarial imitation,” arXiv
2017
Cited alongside, same era.
P. Sermanet, C. Lynch, J. Hsu, and S. Levine, “Time-contrastive networks: Self-supervised learning from multi-view observation,” in CVPRW
2017
Cited alongside, same era.
J.-Y. Zhu, T. Park, P. Isola, and A. Efros, “Unpaired image-to-image translation using cycle-consistent adversarial networks,” in ICCV
2017
Cited alongside, same era.
M. Okada, L. Rigazio, and T. Aoshima, “Path integral networks: End-to-end differentiable optimal control,” arXiv
2017
Cited alongside, same era.
T. Yuval, D. Yotam, M. Alistair, E. Tom, L. Yazhe, d. L. C. Diego, B. David, A. Abbas, M. Josh, L. Andrew, L. Timothy, and R. Martin, “Deepmind control suite,” in arXiv
2018
Later among the works it cites.
F. Torabi, G. Warnell, and P. Stone, “Recent advances in imitation learning from observation,” in IJCAI
2019
Later among the works it cites.
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson, “Learning latent dynamics for planning from pixels,” in ICML
2019
Later among the works it cites.
A. Wang, T. Kurutach, K. Liu, P. Abbeel, and A. Tamar, “Learning robotic manipulation through visual planning and acting,” in RSS
2019
Later among the works it cites.
A. X. Lee, A. Nagabandi, P. Abbeel, and S. Levine, “Stochastic latent actor-critic: Deep reinforcement learning with a latent variable model,” in arXiv
2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Lesort, N. Díaz-Rodríguez, J.-F. Goudou, and D. Filliat, “State representation learning for control: An overview,” in Neural Networks
2018
Cited alongside, same era.
D. Ha and J. Schmidhuber, “Recurrent world models facilitate policy evolution,” in NIPS
2018
Cited alongside, same era.
K. Bousmalis, A. Irpan, P. Wohlhart, Y. Bai, M. Kelcey, M. Kalakrishnan, L. Downs, J. Ibarz, P. Pastor, K. Konolige, S. Levine, and V. Vanhoucke, “Using simulation and domain adaptation to improve efficiency of deep robotic grasping,” in ICRA
2018
Cited alongside, same era.
K. Fang, Y. Bai, S. Hinterstoißer, and M. Kalakrishnan, “Multi-task domain adaptation for deep learning of instance grasping from simulation,” ICRA
2018
Cited alongside, same era.
A. Gonzalez-Garcia, J. van de Weijer, and Y. Bengio, “Image-to-image translation for cross-domain disentanglement,” in NIPS
2018
Cited alongside, same era.
I. Kostrikov, K. K. Agrawal, D. Dwibedi, S. Levine, and J. Tompson, “Discriminator-actor-critic: Addressing sample inefficiency and reward bias in adversarial imitation learning,” in ICLR
2018
Cited alongside, same era.
Y. Liu, A. Gupta, P. Abbeel, and S. Levine, “Imitation from observation: Learning to imitate behaviors from raw video via context translation,” in ICRA
2018
Cited alongside, same era.
Later among the works it cites.
T. Gangwani, J. Lehman, Q. Liu, and J. Peng, “Learning belief representations for imitation learning in pomdps,” in UAI
2019
Later among the works it cites.
A. Sharma, M. Sharma, N. Rhinehart, and K. M. Kitani, “Directed-info gail: Learning hierarchical policies from unsegmented demonstrations using directed information,” in ICLR
2019
Later among the works it cites.
M. Sieb, Z. Xian, A. Huang, O. Kroemer, and K. Fragkiadaki, “Graph-structured visual imitation,” in CoRL
2019
Later among the works it cites.
Y. Lee, E. S. Hu, Z. Yang, and J. J. Lim, “To follow or not to follow: Selective imitation learning from observations,” in CoRL
2019
Later among the works it cites.
M. Okada and T. Taniguchi, “Variational inference mpc for bayesian model-based reinforcement learning,” in CoRL
2019
Later among the works it cites.
L. Smith, N. Dhawan, M. Zhang, P. Abbeel, and S. Levine, “Avid: Learning multi-stage tasks via pixel-level translation of human videos,” in RSS
2020
Closest in time.
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi, “Dream to control: Learning behaviors by latent imagination,” in ICLR
2020
Closest in time.
M. Okada, N. Kosaka, and T. Taniguchi, “Planet of the bayesians: Reconsidering and improving deep planning network by incorporating bayesian inference,” arXiv
2020
Closest in time.
A. Kinose and T. Taniguchi, “Integration of imitation learning using gail and reinforcement learning using task-achievement rewards via probabilistic graphical model,” Advanced Robotics
2020
Closest in time.