Solar: Deep structured latent representations for model-based reinforcement learning
Original
M. Zhang, S. Vikram, L. Smith, P. Abbeel, M. Johnson, and S. Levine · 2018
Later among the works it cites.
Model-based value estimation for efficient model-free reinforcement learning
Original
V. Feinberg, A. Wan, I. Stoica, M. I. Jordan, J. E. Gonzalez, and S. Levine · 2018
Later among the works it cites.
Unsupervised predictive memory in a goal-directed agent
Original
G. Wayne, C.-C. Hung, D. Amos, M. Mirza, A. Ahuja, A. Grabska-Barwinska, et al · 2018
Later among the works it cites.
Learning by playing-solving sparse reward tasks from scratch
M. Riedmiller, R. Hafner, T. Lampe, M. Neunert, J. Degrave, T. Van de Wiele, V. Mnih, et al · 2018
Later among the works it cites.
Sample-efficient reinforcement learning with stochastic ensemble value expansion
J. Buckman, D. Hafner, G. Tucker, E. Brevdo, and H. Lee · 2018
Later among the works it cites.
Combined reinforcement learning via abstract representations
Original
V. François-Lavet, Y. Bengio, D. Precup, and J. Pineau · 2018
Later among the works it cites.
Learning an embedding space for transferable robot skills
K. Hausman, J. T. Springenberg, Z. Wang, N. Heess, and M. Riedmiller · 2018
Later among the works it cites.
Model-based reinforcement learning via meta-policy optimization
I. Clavera, J. Rothfuss, J. Schulman, Y. Fujita, T. Asfour, and P. Abbeel · 2018
Later among the works it cites.
AlphaStar: Mastering the Real-Time Strategy Game StarCraft II, 2019
O. Vinyals, I. Babuschkin, J. Chung, M. Mathieu, M. Jaderberg, W. M. Czarnecki, et al · 2019
Closest in time.
Model-based reinforcement learning for atari
Original
L. Kaiser, M. Babaeizadeh, P. Milos, B. Osinski, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, et al · 2019
Closest in time.
Dexterous manipulation with deep reinforcement learning: Efficient, general, and low-cost
H. Zhu, A. Gupta, A. Rajeswaran, S. Levine, and V. Kumar · 2019
Closest in time.
Learning latent dynamics for planning from pixels
D. Hafner, T. P. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Closest in time.
Cobra: Data-efficient model-based rl through unsupervised object discovery and curiosity-driven exploration
Original
N. Watters, L. Matthey, M. Bosnjak, et al · 2019
Closest in time.
Deepmdp: Learning continuous latent space models for representation learning
C. Gelada, S. Kumar, J. Buckman, O. Nachum, and M. G. Bellemare · 2019
Closest in time.
Exploiting hierarchy for learning and transfer in kl-regularized RL
Original
D. Tirumala, H. Noh, A. Galashov, L. Hasenclever, A. Ahuja, G. Wayne, et al · 2019
Closest in time.
Information asymmetry in KL-regularized RL
A. Galashov, S. Jayakumar, L. Hasenclever, D. Tirumala, J. Schwarz, G. Desjardins, W. M. Czarnecki, Y. W. Teh, et al · 2019
Closest in time.