Fetching the paper…
Reading the bibliography…
Prediction and interpolation for long-range video data involves the complex task of modeling motion trajectories for each visible object, occlusions and dis-occlusions, as well as appearance changes due to viewpoint and lighting.
Lstm can solve hard long time lag problems
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Recognizing human actions: a local svm approach
I. Laptev, B. Caputo, et al · 2004
Earlier work this paper cites.
Recognizing human actions: a local svm approach
C. Schuldt, I. Laptev, and B. Caputo · 2004
Earlier work this paper cites.
Extended kalman filter vs. error state kalman filter for aircraft attitude estimation
V. Madyastha, V. Ravindra, S. Mallikarjunan, and A. Goyal · 2011
Earlier work this paper cites.
Domain adaptation for upper body pose tracking in signed TV broadcasts
J. Charles, T. Pfister, D. Magee, D. Hogg, and A. Zisserman · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
Deep multi-scale video prediction beyond mean square error
M. Mathieu, C. Couprie, and Y. LeCun · 2015
Earlier work this paper cites.
Flowing convnets for human pose estimation in videos
T. Pfister, J. Charles, and A. Zisserman · 2015
Earlier work this paper cites.
Deep visual analogy-making
S. E. Reed, Y. Zhang, Y. Zhang, and H. Lee · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Dense optical flow prediction from a static image
J. Walker, A. Gupta, and M. Hebert · 2015
Earlier work this paper cites.
Density estimation using real nvp
L. Dinh, J. Sohl-Dickstein, and S. Bengio · 2016
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
C. Finn, I. Goodfellow, and S. Levine · 2016
Earlier work this paper cites.
Warpnet: Weakly supervised matching for single-view reconstruction
A. Kanazawa, D. W. Jacobs, and M. Chandraker · 2016
Earlier work this paper cites.
Visual dynamics: Probabilistic future frame synthesis via cross convolutional networks
T. Xue, J. Wu, K. Bouman, and B. Freeman · 2016
Cited alongside, same era.
Stochastic variational video prediction
M. Babaeizadeh, C. Finn, D. Erhan, R. H. Campbell, and S. Levine · 2017
Cited alongside, same era.
Unsupervised learning of disentangled representations from video
E. L. Denton et al · 2017
Cited alongside, same era.
Video frame synthesis using deep voxel flow
Z. Liu, R. A. Yeh, X. Tang, Y. Liu, and A. Agarwala · 2017
Cited alongside, same era.
Unsupervised learning of object frames by dense equivariant image labelling
J. Thewlis, H. Bilen, and A. Vedaldi · 2017
Cited alongside, same era.
Unsupervised learning of object landmarks by factorized spatial embeddings
Sdc-net: Video prediction using spatially-displaced convolution
F. A. Reda, G. Liu, K. J. Shih, R. Kirby, J. Barker, D. Tarjan, A. Tao, and B. Catanzaro · 2018
Later among the works it cites.
Animating arbitrary objects via deep motion transfer
A. Siarohin, S. Lathuilière, S. Tulyakov, E. Ricci, and N. Sebe · 2018
Later among the works it cites.
Discovery of latent 3d keypoints via end-to-end geometric reasoning
S. Suwajanakorn, N. Snavely, J. J. Tompson, and M. Norouzi · 2018
Later among the works it cites.
Mocogan: Decomposing motion and content for video generation
S. Tulyakov, M.-Y. Liu, X. Yang, and J. Kautz · 2018
Later among the works it cites.
Towards accurate generative models of video: A new metric & challenges
T. Unterthiner, S. van Steenkiste, K. Kurach, R. Marinier, M. Michalski, and S. Gelly · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Thewlis, H. Bilen, and A. Vedaldi · 2017
Cited alongside, same era.
Learning to generate long-term future via hierarchical prediction
R. Villegas, J. Yang, Y. Zou, S. Sohn, X. Lin, and H. Lee · 2017
Cited alongside, same era.
Generating the future with adversarial transformers
C. Vondrick and A. Torralba · 2017
Cited alongside, same era.
Stochastic video generation with a learned prior
E. Denton and R. Fergus · 2018
Cited alongside, same era.
Unsupervised learning of object landmarks through conditional image generation
T. Jakab, A. Gupta, H. Bilen, and A. Vedaldi · 2018
Cited alongside, same era.
Super slomo: High quality estimation of multiple intermediate frames for video interpolation
H. Jiang, D. Sun, V. Jampani, M.-H. Yang, E. Learned-Miller, and J. Kautz · 2018
Cited alongside, same era.
Glow: Generative flow with invertible 1x1 convolutions
D. P. Kingma and P. Dhariwal · 2018
Cited alongside, same era.
Hierarchical long-term video prediction without supervision
N. Wichers, R. Villegas, D. Erhan, and H. Lee · 2018
Later among the works it cites.
The unreasonable effectiveness of deep features as a perceptual metric
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang · 2018
Later among the works it cites.
Unsupervised discovery of object landmarks as structural representations
Y. Zhang, Y. Guo, Y. Jin, Y. Luo, Z. He, and H. Lee · 2018
Later among the works it cites.
Optimal transport for gaussian mixture models
Y. Chen, T. T. Georgiou, and A. Tannenbaum · 2019
Closest in time.
Videoflow: A flow-based generative model for video
M. Kumar, M. Babaeizadeh, D. Erhan, C. Finn, S. Levine, L. Dinh, and D. Kingma · 2019
Closest in time.
Unsupervised part-based disentangling of object shape and appearance
D. Lorenz, L. Bereska, T. Milbich, and B. Ommer · 2019
Closest in time.
Semantic image synthesis with spatially-adaptive normalization
T. Park, M.-Y. Liu, T.-C. Wang, and J.-Y. Zhu · 2019
Closest in time.
Video extrapolation with an invertible linear embedding
R. Pottorff, J. Nielsen, and D. Wingate · 2019
Closest in time.
Unsupervised video interpolation using cycle consistency
F. A. Reda, D. Sun, A. Dundar, M. Shoeybi, G. Liu, K. J. Shih, A. Tao, J. Kautz, and B. Catanzaro · 2019
Closest in time.