Fetching the paper…
Reading the bibliography…
In the context of MDPs with high-dimensional states, downstream tasks are predominantly applied on a compressed, low-dimensional representation of the original input space.
R. Bellman, “A markovian decision process,” in Journal of Mathematics and Mechanics , vol. 6, no. 5, 1957
1957
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-Encoding Variational Bayes,” in International Conference on Learning Representations, ICLR , 2014
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative Adversarial Networks,” in Advances in Neural Information Processing Systems, NIPS , 2014
2014
Earlier work this paper cites.
R. Jonschkowski and O. Brock, “Learning state representations with robotic priors,” Autonomous Robots , vol. 39, no. 3, 2015
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature , vol. 518, 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A Method for Stochastic Optimization,” in International Conference on Learning Representations, ICLR , 2015
2015
Earlier work this paper cites.
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, V. Lempitsky, U. Dogan, M. Kloft, F. Orabona, T. Tommasi, and a. Ganin, “ Domain-Adversarial Training of Neural Networks
2016
Earlier work this paper cites.
H. van Hasselt, A. Guez, and D. Silver, “Deep Reinforcement Learning with Double Q-learning,” in Proceedings of the AAAI Conference on Artificial Intelligence, AAAI , 2016
2016
Earlier work this paper cites.
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu, “Reinforcement Learning with Unsupervised Auxiliary Tasks,” in International Conference on Learning Representations, ICLR , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
I. Higgins, A. Pal, A. A. Rusu, L. Matthey, C. P. Burgess, A. Pritzel, M. Botvinick, C. Blundell, and A. Lerchner, “DARLA: Improving Zero-Shot Transfer in Reinforcement Learning,” in International Conference on Machine Learning, ICML , 2017
2017
Earlier work this paper cites.
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell, “Curiosity-driven Exploration by Self-supervised Prediction,” in IEEE Conference on Computer Vision and Pattern Recognition Workshops, CVPRW , 2017
2017
Earlier work this paper cites.
J. Oh, S. Singh, and H. Lee, “Value Prediction Network,” in Advances in Neural Information Processing Systems, NIPS , 2017
2017
Earlier work this paper cites.
A. Laversanne-Finot, A. Pere, and P.-Y. Oudeyer, “Curiosity driven exploration of learned disentangled goal spaces,” in Proceedings of The 2nd Conference on Robot Learning . PMLR, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
V. Francois-Lavet, Y. Bengio, D. Precup, and J. Pineau, “Combined Reinforcement Learning via Abstract Representations,” in Proceedings of the AAAI Conference on Artificial Intelligence, AAAI , 2019
2019
Cited alongside, same era.
D. Yarats, A. Zhang, I. Kostrikov, B. Amos, J. Pineau, and R. Fergus, “Improving Sample Efficiency in Model-Free Reinforcement Learning from Images,” in Proceedings of the AAAI Conference on Artificial Intelligence, AAAI , 2021
2021
Later among the works it cites.
M. Schwarzer, A. Anand, R. Goel, R. D. Hjelm, A. Courville, and P. Bachman, “Data-Efficient Reinforcement Learning with Self-Predictive Representations,” in International Conference on Learning Representations, ICLR , 2021
2021
Later among the works it cites.
I. Kostrikov, D. Yarats, and R. Fergus, “Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from Pixels,” in International Conference on Learning Representations, ICLR , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Gelada, S. Kumar, J. Buckman, O. Nachum, and M. G. Bellemare, “DeepMDP: Learning continuous latent space models for representation learning,” in International Conference on Machine Learning, ICML , 2019
2019
Cited alongside, same era.
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson, “Learning latent dynamics for planning from pixels,” in International Conference on Machine Learning, ICML , 2019
2019
Cited alongside, same era.
F. Locatello, S. Bauer, M. Lucic, G. Raetsch, S. Gelly, B. Schölkopf, and O. Bachem, “Challenging common assumptions in the unsupervised learning of disentangled representations,” in Proceedings of the 36th International Conference on Machine Learning, PMLR , 2019
2019
Cited alongside, same era.
M. Laskin, A. Srinivas, and P. Abbeel, “CURL: Contrastive unsupervised representations for reinforcement learning,” in 37th International Conference on Machine Learning, ICML , 2020
2020
Cited alongside, same era.
K. H. Lee, I. Fischer, A. Z. Liu, Y. Guo, H. Lee, J. Canny, and S. Guadarrama, “Predictive information accelerates learning in RL,” in Advances in Neural Information Processing Systems, NIPS , 2020
2020
Cited alongside, same era.
T. Kipf, E. van der Pol, and M. Welling, “Contrastive Learning of Structured World Models,” in International Conference on Learning Representations, ICLR , 2020
2020
Cited alongside, same era.
A. P. Badia, P. Sprechmann, A. Vitvitskyi, D. Guo, B. Piot, S. Kapturowski, O. Tieleman, M. Arjovsky, A. Pritzel, A. Bolt, and C. Blundell, “Never Give Up: Learning Directed Exploration Strategies,” in International Conference on Learning Representations, ICLR , 2020
2020
Cited alongside, same era.
E. van der Pol, D. Worrall, H. van Hoof, F. Oliehoek, and M. Welling, “Mdp homomorphic networks: Group symmetries in reinforcement learning,” in Advances in Neural Information Processing Systems, NIPS , 2020
2020
Cited alongside, same era.
D. Hafner, T. Lillicrap, M. Norouzi, and J. Ba, “Mastering Atari with Discrete World Models,” in International Conference on Learning Representations, ICLR , 2021
2021
Later among the works it cites.
Y. Efroni, D. Misra, A. Krishnamurthy, A. Agarwal, and J. Langford, “Provable RL with exogenous distractors via multistep inverse dynamics,” in International Conference on Machine Learning, ICML , 2021
2021
Later among the works it cites.
X. Fu, G. Yang, P. Agrawal, and T. Jaakkola, “Learning task informed abstractions,” in International Conference on Machine Learning, ICML , 2021
2021
Later among the works it cites.
R. S. Zimmermann, Y. Sharma, S. Schneider, M. Bethge, and W. Brendel, “Contrastive learning inverts the data generating process,” in Proceedings of the 38th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, M. Meila and T. Zhang, Eds., vol. 139. PMLR, 18–24 Jul 2021, pp. 12 979–12 990
2021
Later among the works it cites.
K. Ahuja, J. Hartford, and Y. Bengio, “Weakly supervised representation learning with sparse perturbations,” in Advances in Neural Information Processing Systems, NIPS , 2022
2022
Closest in time.
2022
Closest in time.
D. Bertoin and E. Rachelson, “Disentanglement by cyclic reconstruction,” in IEEE Transactions on Neural Networks and Learning Systems , 2022
2022
Closest in time.
T. Wang, S. Du, A. Torralba, P. Isola, A. Zhang, and Y. Tian, “Denoised MDPs: Learning world models better than the world itself,” in Proceedings of the 39th International Conference on Machine Learning, PMLR , 2022
2022
Closest in time.