Neural predictive belief representations
Original
Guo, Z. D., Azar, M. G., Piot, B., Pires, B. A., Pohlen, T., and Munos, R · 2018
Later among the works it cites.
World models
Original
Ha, D. and Schmidhuber, J · 2018
Later among the works it cites.
Learning latent dynamics for planning from pixels
Original
Hafner, D., Lillicrap, T., Fischer, I., Villegas, R., Ha, D., Lee, H., and Davidson, J · 2018
Later among the works it cites.
Deep variational reinforcement learning for POMDPs
Original
Igl, M., Zintgraf, L., Le, T. A., Wood, F., and Whiteson, S · 2018
Later among the works it cites.
Structural causal bandits: Where to intervene?
Lee, S. and Bareinboim, E · 2018
Later among the works it cites.
Deconfounding reinforcement learning in observational settings
Original
Lu, C., Schölkopf, B., and Hernández-Lobato, J. M · 2018
Later among the works it cites.
Data-efficient hierarchical reinforcement learning
Nachum, O., Gu, S. S., Lee, H., and Levine, S · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
Sutton, R. S. and Barto, A. G · 2018
Later among the works it cites.
Combating the compounding-error problem with a multi-step model
Original
Asadi, K., Misra, D., Kim, S., and Littman, M. L · 2019
Later among the works it cites.
Woulda, coulda, shoulda: Counterfactually-guided policy search
Buesing, L., Weber, T., Zwols, Y., Heess, N., Racanière, S., Guez, A., and Lespiau, J.-B · 2019
Later among the works it cites.
Shaping belief states with generative environment models for RL
Original
Gregor, K., Rezende, D. J., Besse, F., Wu, Y., Merzic, H., and van den Oord, A · 2019
Later among the works it cites.
An investigation of model-free planning
Guez, A., Mirza, M., Gregor, K., Kabra, R., Racaniere, S., Weber, T., Raposo, D., Santoro, A., Orseau, L., Eccles, T., et al · 2019
Later among the works it cites.
When to trust your model: Model-based policy optimization
Janner, M., Fu, J., Zhang, M., and Levine, S · 2019
Later among the works it cites.
Model-based reinforcement learning for Atari
Original
Kaiser, L., Babaeizadeh, M., Milos, P., Osinski, B., Campbell, R. H., Czechowski, K., Erhan, D., Finn, C., Kozakowski, P., Levine, S., et al · 2019
Later among the works it cites.
Algorithmic framework for model-based deep reinforcement learning with theoretical guarantees
Luo, Y., Xu, H., Li, Y., Tian, Y., Darrell, T., and Ma, T · 2019
Later among the works it cites.
Mastering Atari, go, chess and shogi by planning with a learned model
Original
Schrittwieser, J., Antonoglou, I., Hubert, T., Simonyan, K., Sifre, L., Schmitt, S., Guez, A., Lockhart, E., Hassabis, D., Graepel, T., Lillicrap, T., and Silver, D · 2019
Later among the works it cites.
Causal discovery with reinforcement learning
Original
Zhu, S. and Chen, Z · 2019
Later among the works it cites.