Fetching the paper…
Reading the bibliography…
We present a data-efficient framework for solving visuomotor sequential decision-making problems which exploits the combination of reinforcement learning (RL) and latent variable generative models.
A. J. Ijspeert, J. Nakanishi, and S. Schaal, “Learning attractor landscapes for learning motor primitives,” in Advances in neural information processing systems , 2003, pp. 1547–1554
2003
Earlier work this paper cites.
J. Peters and S. Schaal, “Policy gradient methods for robotics,” in 2006 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2006, pp. 2219–2225
2006
Earlier work this paper cites.
——, “Reinforcement learning of motor skills with policy gradients,” Neural networks , vol. 21, no. 4, pp. 682–697, 2008
2008
Earlier work this paper cites.
G. Neumann, “Variational inference for policy search in changing situations,” in Proceedings of the 28th International Conference on Machine Learning, ICML 2011 , 2011, pp. 817–824
2011
Earlier work this paper cites.
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. Smola, “A kernel two-sample test,” Journal of Machine Learning Research , vol. 13, no. 25, pp. 723–773, 2012
2012
Earlier work this paper cites.
M. P. Deisenroth, G. Neumann, J. Peters et al. , “A survey on policy search for robotics,” Foundations and Trends® in Robotics , vol. 2, no. 1–2, pp. 1–142, 2013
2013
Earlier work this paper cites.
S. Levine and V. Koltun, “Variational policy search via trajectory optimization,” in Advances in neural information processing systems , 2013, pp. 207–215
2013
Earlier work this paper cites.
A. J. Ijspeert, J. Nakanishi, H. Hoffmann, P. Pastor, and S. Schaal, “Dynamical movement primitives: learning attractor models for motor behaviors,” Neural computation , vol. 25, no. 2, pp. 328–373, 2013
2013
Earlier work this paper cites.
Y. Bengio, A. Courville, and P. Vincent, “Representation learning: A review and new perspectives,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 35, no. 8, pp. 1798–1828, 2013
2013
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” International Conference on Learning Representations , 2014
2014
Earlier work this paper cites.
D. J. Rezende, S. Mohamed, and D. Wierstra, “Stochastic backpropagation and approximate inference in deep generative models,” in Int. Conf. Mach. Learn. , 2014, pp. 1278–1286
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International conference on machine learning , 2015, pp. 1889–1897
2015
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” The Journal of Machine Learning Research , vol. 17, no. 1, pp. 1334–1373, 2016
2016
Earlier work this paper cites.
X. Chen, Y. Duan, R. Houthooft, J. Schulman, I. Sutskever, and P. Abbeel, “Infogan: Interpretable representation learning by information maximizing generative adversarial nets,” in Advances in neural information processing systems , 2016, pp. 2172–2180
2016
Earlier work this paper cites.
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel, “Deep spatial autoencoders for visuomotor learning,” in 2016 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2016, pp. 512–519
2016
Earlier work this paper cites.
T. Salimans, I. Goodfellow, W. Zaremba, V. Cheung, A. Radford, X. Chen, and X. Chen, “Improved techniques for training gans,” in Advances in Neural Information Processing Systems 29 , D. D. Lee, M. Sugiyama, U. V. Luxburg, I. Guyon, and R. Garnett, Eds. Curran Associates, Inc., 2016, pp. 2234–2242
2016
Earlier work this paper cites.
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman, “Deep predictive policy training using reinforcement learning,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2017, pp. 2351–2358
2017
Earlier work this paper cites.
I. Higgins, L. Matthey, A. Pal, C. Burgess, X. Glorot, M. Botvinick, S. Mohamed, and A. Lerchner, “beta-vae: Learning basic visual concepts with a constrained variational framework,” in International Conference on Learning Representations , 2017
2017
Earlier work this paper cites.
A. Singh, L. Yang, and S. Levine, “Gplac: Generalizing vision-based robotic skills using weakly labeled images,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 5851–5860
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
C. Finn and S. Levine, “Deep visual foresight for planning robot motion,” in 2017 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2017, pp. 2786–2793
2017
Earlier work this paper cites.
S. Gu, E. Holly, T. Lillicrap, and S. Levine, “Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates,” in 2017 IEEE international conference on robotics and automation (ICRA) . IEEE, 2017, pp. 3389–3396
2017
Earlier work this paper cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in 2017 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, 2017, pp. 23–30
2017
Earlier work this paper cites.
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell, “Adversarial discriminative domain adaptation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 7167–7176
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
N. Mishra, P. Abbeel, and I. Mordatch, “Prediction and control with temporal segment models,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, 2017, pp. 2459–2468
2017
Cited alongside, same era.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “Gans trained by a two time-scale update rule converge to a local nash equilibrium,” in Advances in Neural Information Processing Systems 30 . Curran Associates, Inc., 2017, pp. 6626–6637
2017
Cited alongside, same era.
2017
Cited alongside, same era.
H. Kim and A. Mnih, “Disentangling by factorising,” arXiv preprint arXiv:1802.05983 , 2018
2018
Later among the works it cites.
C. Eastwood and C. K. Williams, “A framework for the quantitative evaluation of disentangled representations,” 2018
2018
Later among the works it cites.
T. Q. Chen, X. Li, R. B. Grosse, and D. K. Duvenaud, “Isolating sources of disentanglement in variational autoencoders,” in Advances in Neural Information Processing Systems , 2018, pp. 2610–2620
2018
Later among the works it cites.
H.-Y. Lee, H.-Y. Tseng, J.-B. Huang, M. Singh, and M.-H. Yang, “Diverse image-to-image translation via disentangled representations,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 35–51
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. L. Denton and v. Birodkar, “Unsupervised learning of disentangled representations from video,” in Advances in Neural Information Processing Systems 30 , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, Eds. Curran Associates, Inc., 2017, pp. 4414–4423. [Online]. Available: http://papers.nips.cc/paper/7028-unsupervised-learning-of-disentangled-representations-from-video.pdf
2017
Cited alongside, same era.
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,” The International Journal of Robotics Research , vol. 37, no. 4-5, pp. 421–436, 2018
2018
Cited alongside, same era.
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke, and S. Levine, “Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation,” in 2nd Conference on Robot Learning (CoRL) , 2018
2018
Cited alongside, same era.
D. Quillen, E. Jang, O. Nachum, C. Finn, J. Ibarz, and S. Levine, “Deep reinforcement learning for vision-based robotic grasping: A simulated comparative evaluation of off-policy methods,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 6284–6291
2018
Cited alongside, same era.
C. Devin, P. Abbeel, T. Darrell, and S. Levine, “Deep object-centric representations for generalizable robot learning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 7111–7118
2018
Cited alongside, same era.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 1–8
2018
Cited alongside, same era.
2018
Cited alongside, same era.
X. Chen, A. Ghadirzadeh, J. Folkesson, M. Björkman, and P. Jensfelt, “Deep reinforcement learning to acquire navigation skills for wheel-legged robots in complex environments,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2018, pp. 3110–3116
2018
Cited alongside, same era.
T. Kynkäänniemi, T. Karras, S. Laine, J. Lehtinen, and T. Aila, “Improved precision and recall metric for assessing generative models,” in Advances in Neural Information Processing Systems , 2019, pp. 3927–3936
2019
Later among the works it cites.
A. Hämäläinen, K. Arndt, A. Ghadirzadeh, and V. Kyrki, “Affordance learning for end-to-end visuomotor robot control,” in 2019 IEEE/RSJ international conference on intelligent robots and systems (IROS) , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Hazara and V. Kyrki, “Transferring generalizable motor primitives from simulation to real world,” IEEE Robotics and Automation Letters , vol. 4, no. 2, pp. 2172–2179, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
F. Wiewel and B. Yang, “Continual learning for anomaly detection with variational autoencoder,” in ICASSP 2019 - 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2019, pp. 3837–3841
2019
Later among the works it cites.
F. Locatello, S. Bauer, M. Lučić, G. Rätsch, S. Gelly, B. Schölkopf, and O. F. Bachem, “Challenging common assumptions in the unsupervised learning of disentangled representations,” in International Conference on Machine Learning , 2019, best Paper Award
2019
Later among the works it cites.
L. Simon, R. Webster, and J. Rabin, “Revisiting precision recall definition for generative modeling,” in Proceedings of the 36th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, K. Chaudhuri and R. Salakhutdinov, Eds., vol. 97. Long Beach, California, USA: PMLR, 09–15 Jun 2019, pp. 5799–5808
2019
Later among the works it cites.
I. Jeon, W. Lee, and G. Kim, “IB-GAN: Disentangled representation learning with information bottleneck GAN,” 2019
2019
Later among the works it cites.
B. Liu, Y. Zhu, Z. Fu, G. de Melo, and A. Elgammal, “Oogan: Disentangling gan with one-hot sampling and orthogonal regularization,” 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
S. van Steenkiste, F. Locatello, J. Schmidhuber, and O. Bachem, “Are disentangled representations helpful for abstract visual reasoning?” arXiv , pp. arXiv–1905, 2019
2019
Later among the works it cites.
K. Arndt, M. Hazara, A. Ghadirzadeh, and V. Kyrki, “Meta reinforcement learning for sim-to-real domain adaptation,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) , 2020
2020
Closest in time.
X. Chen, A. Ghadirzadeh, M. Björkman, and P. Jensfelt, “Adversarial feature training for generalizable robotic visuomotor control,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) , 2020
2020
Closest in time.
J. Bütepage, A. Ghadirzadeh, Ö. Ö. Karadag, M. Björkman, and D. Kragic, “Imitating by generating: deep generative models for imitation of interactive tasks,” Frontiers in Robotics and AI , 2020
2020
Closest in time.
2020
Closest in time.
E. Tzeng, C. Devin, J. Hoffman, C. Finn, P. Abbeel, S. Levine, K. Saenko, and T. Darrell, “Adapting deep visuomotor representations with weak pairwise constraints,” in Algorithmic Foundations of Robotics XII . Springer, 2020, pp. 688–703
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
W. Xu and Y. Tan, “Semisupervised text classification by variational autoencoder,” IEEE Transactions on Neural Networks and Learning Systems , vol. 31, no. 1, pp. 295–308, 2020
2020
Closest in time.