Fetching the paper…
Reading the bibliography…
Temporal observations such as videos contain essential information about the dynamics of the underlying scene, but they are often interleaved with inessential, predictable details.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
The Cross-Entropy Method: A Unified Approach to Combinatorial Optimization, Monte-Carlo Simulation and Machine Learning
Reuven Y. Rubinstein and Dirk P. Kroese · 2004
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Auto-encoding variational Bayes
Diederik P Kingma and Max Welling · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Danilo Jimenez Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2015
Earlier work this paper cites.
A recurrent latent variable model for sequential data
Junyoung Chung, Kyle Kastner, Laurent Dinh, Kratarth Goel, Aaron C Courville, and Yoshua Bengio · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Effective approaches to attention-based neural machine translation
Thang Luong, Hieu Pham, and Christopher D Manning · 2015
Earlier work this paper cites.
Action-conditional video prediction using deep networks in atari games
Junhyuk Oh, Xiaoxiao Guo, Honglak Lee, Richard Lewis, and Satinder Singh · 2015
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Earlier work this paper cites.
Se3-pose-nets: Structured deep dynamics models for visuomotor planning and control
Arunkumar Byravan, Felix Leeb, Franziska Meier, and Dieter Fox · 2017
Earlier work this paper cites.
Recurrent environment simulators
Silvia Chiappa, Sébastien Racanière, Daan Wierstra, and Shakir Mohamed · 2017
Cited alongside, same era.
beta-VAE: Learning basic visual concepts with a constrained variational framework
Irina Higgins, Loic Matthey, Arka Pal, Christopher Burgess, Xavier Glorot, Matthew Botvinick, Shakir Mohamed, and Alexander Lerchner · 2017
Cited alongside, same era.
Video pixel networks
Nal Kalchbrenner, Aäron van den Oord, Karen Simonyan, Ivo Danihelka, Oriol Vinyals, Alex Graves, and Koray Kavukcuoglu · 2017
Cited alongside, same era.
Parallel multiscale autoregressive density estimation
Scott Reed, Aäron van den Oord, Nal Kalchbrenner, Sergio Gómez Colmenarejo, Ziyu Wang, Yutian Chen, Dan Belov, and Nando de Freitas · 2017
Cited alongside, same era.
Fixing a broken ELBO
Alexander Alemi, Ben Poole, Ian Fischer, Joshua Dillon, Rif A. Saurous, and Kevin Murphy · 2018
Cited alongside, same era.
Stochastic variational video prediction
Mohammad Babaeizadeh, Chelsea Finn, Dumitru Erhan, Roy H. Campbell, and Sergey Levine · 2018
Disentangled sequential autoencoder
Yingzhen Li and Stephan Mandt · 2018
Later among the works it cites.
Adaptive skip intervals: Temporal abstraction for recurrent dynamical models
Alexander Neitz, Giambattista Parascandolo, Stefan Bauer, and Bernhard Schölkopf · 2018
Later among the works it cites.
Improved conditional vrnns for video prediction
Lluis Castrejon, Nicolas Ballas, and Aaron Courville · 2019
Closest in time.
Dynamics learning with cascaded variational inference for multi-step manipulation
Kuan Fang, Yuke Zhu, Animesh Garg, Silvio Savarese, and Li Fei-Fei · 2019
Closest in time.
Temporal difference variational auto-encoder
Karol Gregor, George Papamakarios, Frederic Besse, Lars Buesing, and Theophane Weber · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning and querying fast generative models for reinforcement learning
Lars Buesing, Theophane Weber, Sébastien Racanière, S. M. Ali Eslami, Danilo Jimenez Rezende, David P. Reichert, Fabio Viola, Frederic Besse, Karol Gregor, Demis Hassabis, and Daan Wierstra · 2018
Cited alongside, same era.
Stochastic video generation with a learned prior
E. Denton and R. Fergus · 2018
Cited alongside, same era.
Visual foresight: Model-based deep reinforcement learning for vision-based robotic control
Frederik Ebert, Chelsea Finn, Sudeep Dasari, Annie Xie, Alex Lee, and Sergey Levine · 2018
Cited alongside, same era.
Learning latent dynamics for planning from pixels
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson · 2018
Cited alongside, same era.
Probabilistic video generation using holistic attribute control
Jiawei He, Andreas Lehrmann, Joseph Marino, Greg Mori, and Leonid Sigal · 2018
Cited alongside, same era.
Stochastic adversarial video prediction
A. X. Lee, R. Zhang, F. Ebert, P. Abbeel, C. Finn, and S. Levine · 2018
Cited alongside, same era.
Time-agnostic prediction: Predicting predictable video frames
D. Jayaraman, F. Ebert, A. A. Efros, and S. Levine · 2019
Closest in time.
Variational temporal abstraction
Taesup Kim, Sungjin Ahn, and Yoshua Bengio · 2019
Closest in time.
Compositional imitation learning: Explaining and executing one task at a time
Thomas Kipf, Yujia Li, Hanjun Dai, Vinicius Zambaldi, Edward Grefenstette, Pushmeet Kohli, and Peter Battaglia · 2019
Closest in time.
Learning world graphs to accelerate hierarchical reinforcement learning
Wenling Shang, Alex Trott, Stephan Zheng, Caiming Xiong, and Richard Socher · 2019
Closest in time.
High fidelity video prediction with large neural nets
Ruben Villegas, Arkanath Pathak, Harini Kannan, Dumitru Erhan, Quoc V. Le, and Honglak Lee · 2019
Closest in time.
Learning robotic manipulation through visual planning and acting
Angelina Wang, Thanard Kurutach, Kara Liu, Pieter Abbeel, and Aviv Tamar · 2019
Closest in time.
Hierarchical foresight: Self-supervised learning of long-horizon tasks via visual subgoal generation
Suraj Nair and Chelsea Finn · 2020
Closest in time.