Fetching the paper…
Reading the bibliography…
The present paper proposes a novel reinforcement learning method with world models, DreamingV2, a collaborative extension of DreamerV2 and Dreaming.
R. Ueda and T. Arai, “Dynamic programming for global control of the acrobot and its chaotic aspect,” in ICRA
2008
Earlier work this paper cites.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents,” Journal of Artificial Intelligence Research
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, et al
2014
Earlier work this paper cites.
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller, “Embed to control: A locally linear latent dynamics model for control from raw images,” NeurIPS
2015
Earlier work this paper cites.
F. Liu, S. Li, L. Zhang, C. Zhou, R. Ye, Y. Wang, and J. Lu, “3DCNN-DQN-RNN: A deep reinforcement learning framework for semantic parsing of large-scale 3d point clouds,” in CVPR
2017
Earlier work this paper cites.
D. Ha and J. Schmidhuber, “World models,” arXiv:1803.10122
2018
Earlier work this paper cites.
Y. Tassa, Y. Doron, A. Muldal, T. Erez, Y. Li, et al
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson, “Learning latent dynamics for planning from pixels,” in ICML
2019
Earlier work this paper cites.
M. Okada and T. Taniguchi, “Variational inference MPC for bayesian model-based reinforcement learning,” in CoRL
2019
Earlier work this paper cites.
M. Okada, N. Kosaka, and T. Taniguchi, “PlaNet of the Bayesians: Reconsidering and improving deep planning network by incorporating Bayesian inference,” in IROS
2020
Earlier work this paper cites.
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi, “Dream to control: Learning behaviors by latent imagination,” ICLR
2020
Cited alongside, same era.
A. Byravan, J. T. Springenberg, A. Abdolmaleki, R. Hafner, et al
2020
Cited alongside, same era.
2020
Cited alongside, same era.
T. Yu, G. Thomas, L. Yu, S. Ermon, J. Y. Zou, S. Levine, C. Finn, and T. Ma, “MOPO: Model-based offline policy optimization,” NeurIPS
2020
Cited alongside, same era.
X. Ma, S. Chen, D. Hsu, and W. S. Lee, “Contrastive variational model-based reinforcement learning for complex observations,” CoRL
2020
Cited alongside, same era.
2020
Later among the works it cites.
M. Caron, I. Misra, J. Mairal, P. Goyal, P. Bojanowski, and A. Joulin, “Unsupervised learning of visual features by contrasting cluster assignments,” NeurIPS
2020
Later among the works it cites.
D. Hafner, T. Lillicrap, M. Norouzi, and J. Ba, “Mastering atari with discrete world models,” in ICLR
2021
Later among the works it cites.
M. Okada and T. Taniguchi, “Dreaming: Model-based reinforcement learning by latent imagination without reconstruction,” in ICRA
2021
Later among the works it cites.
T. D. Nguyen, R. Shu, T. Pham, H. Bui, and S. Ermon, “Temporal predictive coding for model-based planning in latent space,” in ICML
2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,” in ICLR
2020
Cited alongside, same era.
D. Han, K. Doya, and J. Tani, “Variational recurrent models for solving partially observable control tasks,” in ICLR
2020
Cited alongside, same era.
2020
Cited alongside, same era.
T. Wang and P. Isola, “Understanding contrastive representation learning through alignment and uniformity on the hypersphere,” in ICLR
2020
Cited alongside, same era.
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick, “Momentum contrast for unsupervised visual representation learning,” in CVPR
2020
Cited alongside, same era.
J.-B. Grill, F. Strub, F. Altché, C. Tallec, et al
2020
Cited alongside, same era.
A. Srinivas, M. Laskin, and P. Abbeel, “CURL: Contrastive unsupervised representations for reinforcement learning,” in ICML
2020
Cited alongside, same era.
Later among the works it cites.
2021
Later among the works it cites.
C. Li, F. Xia, R. Martín-Martín, M. Lingelbach, S. Srivastava, et al
2021
Later among the works it cites.
K. Paster, L. E. McKinney, S. A. McIlraith, and J. Ba, “BLAST: Latent dynamics models from bootstrapping,” in Deep RL Workshop NeurIPS 2021
2021
Later among the works it cites.
K. Chen, Y. Lee, and H. Soh, “Multi-modal mutual information (MUMMI) training for robust self-supervised deep reinforcement learning,” in ICRA
2021
Later among the works it cites.
J. Zbontar, L. Jing, I. Misra, Y. LeCun, and S. Deny, “Barlow Twins: self-supervised learning via redundancy reduction,” in ICML
2021
Later among the works it cites.
T. Sakai and T. Nagai, “Explainable autonomous robots: a survey and perspective,” Advanced Robotics
2022
Closest in time.
D. Yarats, R. Fergus, A. Lazaric, and L. Pinto, “Mastering visual continuous control: Improved data-augmented reinforcement learning,” in ICLR
2022
Closest in time.