Fetching the paper…
Reading the bibliography…
In visual planning (VP), an agent learns to plan goal-directed behavior from observations of a dynamical system obtained offline, e.g., images obtained from self-supervised robot interaction.
Introduction to reinforcement learning , volume 135
Sutton, R. S., Barto, A. G., et al · 1998
Earlier work this paper cites.
Batch reinforcement learning
Lange, S., Gabel, T., and Riedmiller, M · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y · 2012
Earlier work this paper cites.
Conditional generative adversarial nets
Mirza, M. and Osindero, S · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al · 2015
Earlier work this paper cites.
Trust region policy optimization
Schulman, J., Levine, S., Abbeel, P., Jordan, M., and Moritz, P · 2015
Earlier work this paper cites.
Learning structured output representation using deep conditional generative models
Sohn, K., Lee, H., and Yan, X · 2015
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images
Watter, M., Springenberg, J., Boedecker, J., and Riedmiller, M · 2015
Earlier work this paper cites.
Learning to poke by poking: Experiential learning of intuitive physics
Agrawal, P., Nair, A. V., Abbeel, P., Malik, J., and Levine, S · 2016
Earlier work this paper cites.
Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours
Pinto, L. and Gupta, A · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Cited alongside, same era.
Hindsight experience replay
Andrychowicz, M., Wolski, F., Ray, A., Schneider, J., Fong, R., Welinder, P., McGrew, B., Tobin, J., Abbeel, O. P., and Zaremba, W · 2017
Cited alongside, same era.
Deep visual foresight for planning robot motion
Finn, C. and Levine, S · 2017
Cited alongside, same era.
Combining self-supervised learning and imitation for vision-based rope manipulation
Nair, A., Chen, D., Agrawal, P., Isola, P., Abbeel, P., Malik, J., and Levine, S · 2017
Cited alongside, same era.
Classical planning in deep latent space: Bridging the subsymbolic-symbolic boundary
Asai, M. and Fukunaga, A · 2018
Cited alongside, same era.
Visual foresight: Model-based deep reinforcement learning for vision-based robotic control
Representation learning with contrastive predictive coding
Oord, A. v. d., Li, Y., and Vinyals, O · 2018
Later among the works it cites.
Semi-parametric topological memory for navigation
Savinov, N., Dosovitskiy, A., and Koltun, V · 2018
Later among the works it cites.
Srinivas, A., Jabri, A., Abbeel, P., Levine, S., and Finn, C · 2018
Later among the works it cites.
Unsupervised grounding of plannable first-order logic representation from images
Asai, M · 2019
Later among the works it cites.
Search on the replay buffer: Bridging planning and reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ebert, F., Finn, C., Dasari, S., Xie, A., Lee, A., and Levine, S · 2018
Cited alongside, same era.
Soft actor-critic algorithms and applications
Haarnoja, T., Zhou, A., Hartikainen, K., Tucker, G., Ha, S., Tan, J., Kumar, V., Zhu, H., Gupta, A., Abbeel, P., et al · 2018
Cited alongside, same era.
Learning latent dynamics for planning from pixels
Hafner, D., Lillicrap, T., Fischer, I., Villegas, R., Ha, D., Lee, H., and Davidson, J · 2018
Cited alongside, same era.
Learning plannable representations with causal infogan
Kurutach, T., Tamar, A., Yang, G., Russell, S. J., and Abbeel, P · 2018
Cited alongside, same era.
Stochastic adversarial video prediction
Lee, A. X., Zhang, R., Ebert, F., Abbeel, P., Finn, C., and Levine, S · 2018
Cited alongside, same era.
Eysenbach, B., Salakhutdinov, R., and Levine, S · 2019
Later among the works it cites.
Robot motion planning in learned latent spaces
Ichter, B. and Pavone, M · 2019
Later among the works it cites.
Model-based reinforcement learning for atari
Kaiser, L., Babaeizadeh, M., Milos, P., Osinski, B., Campbell, R. H., Czechowski, K., Erhan, D., Finn, C., Kozakowski, P., Levine, S., et al · 2019
Later among the works it cites.
Hierarchical foresight: Self-supervised learning of long-horizon tasks via visual subgoal generation
Nair, S. and Finn, C · 2019
Later among the works it cites.
Motion planning networks
Qureshi, A. H., Simeonov, A., Bency, M. J., and Yip, M. C · 2019
Later among the works it cites.
Learning robotic manipulation through visual planning and acting
Wang, A., Kurutach, T., Liu, K., Abbeel, P., and Tamar, A · 2019
Later among the works it cites.