Model based reinforcement learning for atari
Kaiser, L., Babaeizadeh, M., Miłos, P., Osiński, B., Campbell, R. H., Czechowski, K., Erhan, D., Finn, C., Kozakowski, P., Levine, S., Mohiuddin, A., Sepassi, R., Tucker, G., and Michalewski, H · 2020
Later among the works it cites.
MOReL: Model-based offline reinforcement learning
Kidambi, R., Rajeswaran, A., Netrapalli, P., and Joachims, T · 2020
Later among the works it cites.
Conservative Q-learning for offline reinforcement learning
Kumar, A., Zhou, A., Tucker, G., and Levine, S · 2020
Later among the works it cites.
Learning accurate long-term dynamics for model-based reinforcement learning
Original
Lambert, N. O., Wilcox, A., Zhang, H., Pister, K. S., and Calandra, R · 2020
Later among the works it cites.
Deep imitative models for flexible inference, planning, and control
Rhinehart, N., McAllister, R., and Levine, S · 2020
Later among the works it cites.
Implementation of denoising diffusion probabilistic models in pytorch, 2020
Wang, P · 2020
Later among the works it cites.
MOPO: Model-based offline policy optimization
Yu, T., Thomas, G., Yu, L., Ermon, S., Zou, J., Levine, S., Finn, C., and Ma, T · 2020
Later among the works it cites.
Model-based offline planning
Argenson, A. and Dulac-Arnold, G · 2021
Later among the works it cites.
Structured denoising diffusion models in discrete state-spaces
Austin, J., Johnson, D. D., Ho, J., Tarlow, D., and van den Berg, R · 2021
Later among the works it cites.
Diffusion models beat GANs on image synthesis
Dhariwal, P. and Nichol, A. Q · 2021
Later among the works it cites.
Mismatched no more: Joint model-policy optimization for model-based rl
Original
Eysenbach, B., Khazatsky, A., Levine, S., and Salakhutdinov, R · 2021
Later among the works it cites.
Offline reinforcement learning as one big sequence modeling problem
Janner, M., Li, Q., and Levine, S · 2021
Later among the works it cites.
3d neural scene representations for visuomotor control
Li, Y., Li, S., Sitzmann, V., Agrawal, P., and Torralba, A · 2021
Later among the works it cites.
Improved denoising diffusion probabilistic models
Nichol, A. Q. and Dhariwal, P · 2021
Later among the works it cites.
Vector quantized models for planning
Ozair, S., Li, Y., Razavi, A., Antonoglou, I., Van Den Oord, A., and Vinyals, O · 2021
Later among the works it cites.
Model-based reinforcement learning via latent-space collocation
Rybkin, O., Zhu, C., Nagabandi, A., Daniilidis, K., Mordatch, I., and Levine, S · 2021
Later among the works it cites.
Denoising diffusion implicit models
Song, J., Meng, C., and Ermon, S · 2021
Later among the works it cites.
3D shape generation and completion through point-voxel diffusion
Zhou, L., Du, Y., and Wu, J · 2021
Later among the works it cites.
Offline reinforcement learning with implicit Q-learning
Kostrikov, I., Nair, A., and Levine, S · 2022
Closest in time.