Fetching the paper…
Reading the bibliography…
Generating video frames that accurately predict future world states is challenging.
Slow feature analysis: Unsupervised learning of invariance
Wiskott, L. and Sejnowski, T · 2002
Earlier work this paper cites.
Recognizing human actions: A local svm approach
Schuldt, Christian, Laptev, Ivan, and Caputo, Barbara · 2004
Earlier work this paper cites.
Beyond pixels: exploring new representations and applications for motion analysis
Liu, C · 2009
Earlier work this paper cites.
Deep learning of invariant features via simulated fixations in video
Zou, W. Y., Zhu, S., Ng, A. Y., , and Yu., K · 2012
Earlier work this paper cites.
Learning stochastic recurrent networks
Bayer, Justin and Osendorfer, Christian · 2014
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, Ian J., Pouget-Abadie, Jean, Mirza, Mehdi, Xu, Bing, Warde-Farley, David, Ozair, Sherjil, Courville, Aaron, and Bengio, Yoshua · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D.P. and Welling, M · 2014
Earlier work this paper cites.
Video (language) modeling: a baseline for generative models of natural videos
Ranzato, Marc’Aurelio, Szlam, Arthur, Bruna, Joan, Mathieu, Michaël, Collobert, Ronan, and Chopra, Sumit · 2014
Earlier work this paper cites.
Learning to see by moving
Agrawal, P, Carreira, J, and Malik, J · 2015
Earlier work this paper cites.
A recurrent latent variable model for sequential data
Chung, Junyoung, Kastner, Kyle, Dinh, Laurent, Goel, Kratarth, Courville, Aaron, and Bengio, Yoshua · 2015
Earlier work this paper cites.
Learning to linearize under uncertainty
Goroshin, Ross, Mathieu, Michael, and LeCun, Yann · 2015
Earlier work this paper cites.
Learning image represen- tations tied to ego-motion
Jayaraman, D. and Grauman, K · 2015
Earlier work this paper cites.
Krishnan, R., Shalit, U., and Sontag, D · 2015
Earlier work this paper cites.
Action-conditional video prediction using deep networks in Atari games
Oh, J., Guo, X., Lee, H., Lewis, R., and Singh, S · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A · 2015
Cited alongside, same era.
Unsupervised learning of video representations using LSTMs
Srivastava, N., Mansimov, E., and Salakhutdinov, R · 2015
Cited alongside, same era.
Dense optical flow prediction from a static image
Walker, J., Gupta, A., and Hebert, M · 2015
Cited alongside, same era.
Unsupervised learning of visual representations using videos
Wang, Xiaolong and Gupta, Abhinav · 2015
Cited alongside, same era.
Generating sentences from a continuous space
Bowman, Samuel R., Vilnis, Luke, Vinyals, Oriol, Dai, Andrew M., Jozefowicz, Rafal, and Bengio, Samy · 2016
Cited alongside, same era.
Pixel recurrent neural networks
van den Oord, A., Kalchbrenner, N., and Kavukcuoglu, K · 2016
Later among the works it cites.
Generating videos with scene dynamics
Vondrick, C., Pirsiavash, H., and Torralba, A · 2016
Later among the works it cites.
Visual dynamics: Probabilistic future frame synthesis via cross convolutional networks
Xue, Tianfan, Wu, Jiajun, Bouman, Katherine L, and Freeman, William T · 2016
Later among the works it cites.
Stochastic variational video prediction
Babaeizadeh, Mohammad, Finn, Chelsea, Erhan, Dumitru, Campbell, Roy H., and Levine, Sergey · 2017
Later among the works it cites.
Recurrent environment simulators
Chiappa, S., Racaniere, S., Wierstra, D., and Mohamed, S · 2017
Later among the works it cites.
Unsupervised learning of disentangled representations from video
Denton, Remi and Birodkar, Vighnesh · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unsupervised learning for physical interaction through video prediction
Finn, C., Goodfellow, I., and Levine, S · 2016
Cited alongside, same era.
Sequential neural models with stochastic layers
Fraccaro, Marco, Sønderby, Søren Kaae, Paquet, Ulrich, and Winther, Ole · 2016
Cited alongside, same era.
Kalchbrenner, N., van den Oord, A., Simonyan, K., Danihelka, I., Vinyals, O., Graves, A., and Kavukcuoglu, K · 2016
Cited alongside, same era.
Deep predictive coding networks for video prediction and unsupervised learning
Lotter, William, Kreiman, Gabriel, and Cox, David · 2016
Cited alongside, same era.
Deep multi-scale video prediction beyond mean square error
Mathieu, Michaël, Couprie, Camille, and LeCun, Yann · 2016
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
Radford, Alec, Metz, Luke, and Chintala, Soumith · 2016
Cited alongside, same era.
Later among the works it cites.
Self-supervised visual planning with temporal skip connections
Ebert, Frederik, Finn, Chelsea, Lee, Alex X., and Levine, Sergey · 2017
Later among the works it cites.
Prediction under uncertainty with error-encoding networks
Henaff, Mikael, Zhao, Junbo, and LeCun, Yann · 2017
Later among the works it cites.
Early visual concept learning with unsupervised deep learning
Higgins, Irina, Matthey, Loic, Pal, Arka, Burgess, Christopher, Glorot, Xavier, Botvinick, Matthew, Mohamed, Shakir, and Lerchner, Alexander · 2017
Later among the works it cites.
Progressive growing of gans for improved quality, stability, and variation
Karras, Tero, Aila, Timo, Laine, Samuli, and Lehtinen, Jaakko · 2017
Later among the works it cites.
Parallel multiscale autoregressive density estimation
Reed, Scott, van den Oord, Aaron, Kalchbrenner, Nal, Colmenarejo, Sergio Gomez, Wang, Ziyu, Belov, Dan, and de Freitas, Nando · 2017
Later among the works it cites.
Salimans, Tim, Karpathy, Andrej, Chen, Xi, and Kingma, Diederik P · 2017
Later among the works it cites.
Generating the future with adversarial transformers
Vondrick, C. and Torralba, A · 2017
Later among the works it cites.