Fetching the paper…
Reading the bibliography…
Learning to predict future images from a video sequence involves the construction of an internal representation that models the image evolution accurately, and therefore, to some degree, its content and dynamics.
Fast vision-guided mobile robot navigation using model-based reasoning and prediction of uncertainties
Kosaka, Akio and Kak, Avinash C · 1992
Earlier work this paper cites.
Signature verification using a “siamese” time delay neural network
Bromley, Jane, Bentz, James W, Bottou, Léon, Guyon, Isabelle, LeCun, Yann, Moore, Cliff, Säckinger, Eduard, and Shah, Roopak · 1993
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Yann, Bottou, Léon, Bengio, Yoshua, and Haffner, Patrick · 1998
Earlier work this paper cites.
Image quality assessment: From error visibility to structural similarity
Wang, Zhou, Bovik, Alan C., Sheikh, Hamid R., and Simoncelli, Eero P · 2004
Earlier work this paper cites.
Improving frame interpolation with spatial motion smoothing for pixel domain distributed video coding
Ascenso, Joao, Brites, Catarina, and Pereira, Fernando · 2005
Earlier work this paper cites.
Supervised learning of image restoration with convolutional networks
Jain, Viren, Murray, Joseph F, Roth, Fabian, Turaga, Srinivas, Zhigulin, Valentin, Briggman, Kevin L, N, Helmstaedter Moritz, Denk, Winfried, and Seung, Sebastian H · 2007
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
Vincent, Pascal, Larochelle, Hugo, Bengio, Yoshua, and Manzagol, Pierre-Antoine · 2008
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, Vinod and Hinton, Geoffrey E · 2010
Earlier work this paper cites.
Blind deconvolution using a normalized sparsity measure
Krishnan, Dilip, Tay, Terence, and Fergus, Rob · 2011
Earlier work this paper cites.
UCF101: A dataset of 101 human actions classes from videos in the wild
Soomro, Khurram, Zamir, Amir Roshan, and Shah, Mubarak · 2012
Cited alongside, same era.
Building high-level features using large scale unsupervised learning
Le, Quoc V · 2013
Cited alongside, same era.
Generative adversarial networks
Goodfellow, Ian J., Pouget-Abadie, Jean, Mirza, Mehdi, Xu, Bing, Warde-Farley, David, Ozair, Sherjil, Courville, Aaron C., and Bengio, Yoshua · 2014
Cited alongside, same era.
Large-scale video classification with convolutional neural networks
Karpathy, Andrej, Toderici, George, Shetty, Sanketh, Leung, Thomas, Sukthankar, Rahul, and Fei-Fei, Li · 2014
Cited alongside, same era.
Video (language) modeling: a baseline for generative models of natural videos
Ranzato, Marc’Aurelio, Szlam, Arthur, Bruna, Joan, Mathieu, Michaël, Collobert, Ronan, and Chopra, Sumit · 2014
Cited alongside, same era.
Learning image representations equivariant to ego-motion
Jayaraman, Dinesh and Grauman, Kristen · 2015
Closest in time.
Fully convolutional networks for semantic segmentation
Long, Jonathan, Shelhamer, Evan, and Darrell, Trevor · 2015
Closest in time.
Understanding deep image representations by inverting them
Mahendran, Aravindh and Vedaldi, Andrea · 2015
Closest in time.
Action-conditional video prediction using deep networks in atari games
Oh, Junhyuk, Guo, Xiaoxiao, Lee, Honglak, Lewis, Richard L., and Singh, Satinder P · 2015
Closest in time.
EpicFlow: Edge-Preserving Interpolation of Correspondences for Optical Flow
Revaud, Jerome, Weinzaepfel, Philippe, Harchaoui, Zaid, and Schmid, Cordelia · 2015
Closest in time.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, Olaf, Fischer, Philipp, and Brox, Thomas · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Two-stream convolutional networks for action recognition in videos
Simonyan, Karen and Zisserman, Andrew · 2014
Cited alongside, same era.
Deep generative image models using a laplacian pyramid of adversarial networks
Denton, Emily, Chintala, Soumith, Szlam, Arthur, and Fergus, Rob · 2015
Cited alongside, same era.
Learning to generate chairs with convolutional neural networks
Dosovitskiy, Alexey, Springenberg, Jost Tobias, and Brox, Thomas · 2015
Cited alongside, same era.
Deepstereo: Learning to predict new views from the world’s imagery
Flynn, John, Neulander, Ivan, Philbin, James, and Snavely, Noah · 2015
Cited alongside, same era.
Learning to linearize under uncertainty
Goroshin, Ross, Mathieu, Michaël, and LeCun, Yann · 2015
Cited alongside, same era.
Closest in time.
Unsupervised learning of video representations using LSTMs
Srivastava, Nitish, Mansimov, Elman, and Salakhutdinov, Ruslan · 2015
Closest in time.
C3D: generic features for video analysis
Tran, Du, Bourdev, Lubomir D., Fergus, Rob, Torresani, Lorenzo, and Paluri, Manohar · 2015
Closest in time.
Anticipating the future by watching unlabeled video
Vondrick, Carl, Pirsiavash, Hamed, and Torralba, Antonio · 2015
Closest in time.
Unsupervised learning of visual representations using videos
Wang, Xiaolong and Gupta, Abhinav · 2015
Closest in time.