World models
Original
Ha, D. and Schmidhuber, J · 2018
Later among the works it cites.
Learning to play with intrinsically-motivated, self-aware agents
Haber, N., Mrowca, D., Wang, S., Fei-Fei, L. F., and Yamins, D. L · 2018
Later among the works it cites.
Plan online, learn offline: Efficient learning and exploration via model-based control
Original
Lowrey, K., Rajeswaran, A., Kakade, S., Todorov, E., and Mordatch, I · 2018
Later among the works it cites.
Visual reinforcement learning with imagined goals
Nair, A. V., Pong, V., Dalal, M., Bahl, S., Lin, S., and Levine, S · 2018
Later among the works it cites.
Randomized prior functions for deep reinforcement learning
Osband, I., Aslanides, J., and Cassirer, A · 2018
Later among the works it cites.
Count-based exploration with neural density models
Ostrovski, G., Bellemare, M. G., Oord, A. v. d., and Munos, R · 2018
Later among the works it cites.
Zero-shot visual imitation
Pathak, D., Mahmoudieh, P., Luo, G., Agrawal, P., Chen, D., Shentu, Y., Shelhamer, E., Malik, J., Efros, A. A., and Darrell, T · 2018
Later among the works it cites.
DeepMind control suite
Original
Tassa, Y., Doron, Y., Muldal, A., Erez, T., Li, Y., de Las Casas, D., Budden, D., Abdolmaleki, A., Merel, J., Lefrancq, A., Lillicrap, T., and Riedmiller, M · 2018
Later among the works it cites.
Large-scale study of curiosity-driven learning
Burda, Y., Edwards, H., Pathak, D., Storkey, A., Darrell, T., and Efros, A. A · 2019
Later among the works it cites.
Learning latent dynamics for planning from pixels
Hafner, D., Lillicrap, T., Fischer, I., Villegas, R., Ha, D., Lee, H., and Davidson, J · 2019
Later among the works it cites.
Explicit explore-exploit algorithms in continuous state spaces
Henaff, M · 2019
Later among the works it cites.
Deep dynamics models for learning dexterous manipulation
Original
Nagabandi, A., Konoglie, K., Levine, S., and Kumar, V · 2019
Later among the works it cites.
Self-supervised exploration via disagreement
Pathak, D., Gandhi, D., and Gupta, A · 2019
Later among the works it cites.
Dynamics-aware unsupervised discovery of skills
Original
Sharma, A., Gu, S., Levine, S., Kumar, V., and Hausman, K · 2019
Later among the works it cites.
Model-Based Active Exploration
Shyam, P., Jaśkowski, W., and Gomez, F · 2019
Later among the works it cites.
Solar: deep structured representations for model-based reinforcement learning
Original
Zhang, M., Vikram, S., Smith, L., Abbeel, P., Johnson, M., and Levine, S · 2019
Later among the works it cites.
Ready policy one: World building through active learning
Original
Ball, P., Parker-Holder, J., Pacchiano, A., Choromanski, K., and Roberts, S · 2020
Closest in time.
Dream to control: Learning behaviors by latent imagination
Hafner, D., Lillicrap, T., Ba, J., and Norouzi, M · 2020
Closest in time.