Fetching the paper…
Reading the bibliography…
Projecting high-dimensional environment observations into lower-dimensional structured representations can considerably improve data-efficiency for reinforcement learning in domains with limited data such as robotics.
L. P. Kaelbling, “Learning to achieve goals,” in IJCAI . Citeseer, 1993, pp. 1094–1099
1993
Earlier work this paper cites.
R. S. Sutton, A. G. Barto, et al. , Introduction to reinforcement learning . MIT press Cambridge, 1998, vol. 135
1998
Earlier work this paper cites.
E. Bingham and H. Mannila, “Random projection in dimensionality reduction: applications to image and text data,” in Proceedings of the seventh ACM SIGKDD , 2001, pp. 245–250
2001
Earlier work this paper cites.
A. S. Klyubin, D. Polani, and C. L. Nehaniv, “All else being equal be empowered,” in European Conference on Artificial Life . Springer, 2005, pp. 744–753
2005
Earlier work this paper cites.
G. E. Hinton and R. R. Salakhutdinov, “Reducing the dimensionality of data with neural networks,” Science , vol. 313, no. 5786, pp. 504–507, 2006
2006
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” in ICML , 2009
2009
Earlier work this paper cites.
S. Lange, M. Riedmiller, and A. Voigtländer, “Autonomous reinforcement learning on raw visual input data in a real world application,” in IJCNN . IEEE, 2012, pp. 1–8
2012
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in Intelligent Robots and Systems (IROS), 2012 IEEE/RSJ International Conference on . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in ICLR , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
K. Gregor, F. Besse, D. J. Rezende, I. Danihelka, and D. Wierstra, “Towards conceptual compression,” in NeurIPS , 2016, pp. 3549–3557
2016
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” in ICLR , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
I. Higgins, L. Matthey, A. Pal, C. Burgess, X. Glorot, M. Botvinick, S. Mohamed, and A. Lerchner, “Beta-vae: Learning basic visual concepts with a constrained variational framework.” ICLR , vol. 2, no. 5, p. 6, 2017
2017
Earlier work this paper cites.
I. Higgins, A. Pal, A. Rusu, L. Matthey, C. Burgess, A. Pritzel, M. Botvinick, C. Blundell, and A. Lerchner, “DARLA: Improving zero-shot transfer in reinforcement learning,” ICML , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. Andrychowicz, D. Crow, A. Ray, J. Schneider, R. Fong, P. Welinder, B. McGrew, J. Tobin, P. Abbeel, and W. Zaremba, “ Hindsight experience replay ,” in NeurIPS , 2017, pp. 5055–5065
2017
Earlier work this paper cites.
A. Dosovitskiy and V. Koltun, “Learning to act by predicting the future,” in ICLR , 2017
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
S. Cabi, S. G. Colmenarejo, M. W. Hoffman, M. Denil, Z. Wang, and N. de Freitas, “ The Intentional Unintentional Agent : Learning to solve many continuous control tasks simultaneously,” in CoRL , 2017
2017
Cited alongside, same era.
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell, “Curiosity-driven exploration by self-supervised prediction,” in IEEE CVPR Workshops , 2017, pp. 16–17
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. Gregor, D. J. Rezende, and D. Wierstra, “Variational intrinsic control,” in ICLR , 2017
2017
Cited alongside, same era.
A. Graves, M. G. Bellemare, J. Menick, R. Munos, and K. Kavukcuoglu, “Automated curriculum learning for neural networks,” in ICML . JMLR. org, 2017, pp. 1311–1320
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine, “Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning,” in CoRL , 2019
2019
Later among the works it cites.
J. Donahue and K. Simonyan, “Large scale adversarial representation learning,” in NeurIPS , 2019
2019
Later among the works it cites.
A. Byravan, J. T. Springenberg, A. Abdolmaleki, R. Hafner, M. Neunert, et al. , “Imagined value gradients: Model-based policy optimization with transferable latent dynamics models,” in CoRL , 2019, pp. 566–589
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
F. Locatello, S. Bauer, M. Lucic, G. Raetsch, S. Gelly, B. Schölkopf, and O. Bachem, “Challenging common assumptions in the unsupervised learning of disentangled representations,” in ICML , 2019, pp. 4114–4124
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
J. Luketina, M. Smith, M. Igl, and S. Whiteson, “Transfer learning via diverse policies in value-relevant features,” BeTR-RL workshop, ICLR , 2020
2020
Closest in time.
2020
Closest in time.