Fetching the paper…
Reading the bibliography…
We present a novel deep learning architecture for probabilistic future prediction from video.
Piaget, J.: The Origins of Intelligence in the Child. London: Routledge and Kegan Paul (1936)
1936
Earlier work this paper cites.
Schuldt, C., Laptev, I., Caputo, B.: Recognizing human actions: A local svm approach. In: Proceedings of the International Conference on Pattern Recognition (2004)
2004
Earlier work this paper cites.
Levine, S., Abbeel, P.: Learning neural network policies with guided policy search under unknown dynamics. Advances in Neural Information Processing Systems (NeurIPS) (2014)
2014
Earlier work this paper cites.
Ranzato, M., Szlam, A., Bruna, J., Mathieu, M., Collobert, R., Chopra, S.: Video (language) modeling: a baseline for generative models of natural videos. arXiv preprint (2014)
2014
Earlier work this paper cites.
Simonyan, K., Zisserman, A.: Two-stream convolutional networks for action recognition in videos. In: NIPS (2014)
2014
Earlier work this paper cites.
Oh, J., Guo, X., Lee, H., Lewis, R., Singh, S.: Action-conditional video prediction using deep networks in atari games. In: Advances in Neural Information Processing Systems (NeurIPS) (2015)
2015
Earlier work this paper cites.
Shi, X., Chen, Z., Wang, H., Yeung, D.Y., Wong, W.k., Woo, W.c.: Convolutional lstm network: A machine learning approach for precipitation nowcasting. In: Advances in Neural Information Processing Systems (NeurIPS) (2015)
2015
Earlier work this paper cites.
Simonyan, K., Zisserman, A.: Very deep convolutional networks for large-scale image recognition. In: Proceedings of the International Conference on Learning Representations (ICLR) (2015)
2015
Earlier work this paper cites.
Srivastava, N., Mansimov, E., Salakhudinov, R.: Unsupervised learning of video representations using lstms. In: ICML (2015)
2015
Earlier work this paper cites.
Sun, L., Jia, K., Yeung, D., Shi, B.E.: Human action recognition using factorized spatio-temporal convolutional networks. Proceedings of the International Conference on Computer Vision (ICCV) (2015)
2015
Earlier work this paper cites.
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S.E., Anguelov, D., Erhan, D., Vanhoucke, V., Rabinovich, A.: Going deeper with convolutions. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2015)
2015
Earlier work this paper cites.
Ballas, N., Yao, L., Pas, C., Courville, A.: Delving deeper into convolutional networks for learning video representations. In: Proceedings of the International Conference on Learning Representations (ICLR) (2016)
2016
Earlier work this paper cites.
Bilen, H., Fernando, B., Gavves, E., Vedaldi, A., Gould, S.: Dynamic image networks for action recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Earlier work this paper cites.
Bojarski, M., Testa, D.D., Dworakowski, D., Firner, B., Flepp, B., Goyal, P., Jackel, L.D., Monfort, M., Muller, U., Zhang, J., Zhang, X., Zhao, J., Zieba, K.: End to End Learning for Self-Driving Cars. arXiv preprint (2016)
2016
Earlier work this paper cites.
Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., Schiele, B.: The cityscapes dataset for semantic urban scene understanding. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Earlier work this paper cites.
Feichtenhofer, C., Pinz, A., Wildes, R.P.: Spatiotemporal residual networks for video action recognition. In: Advances in Neural Information Processing Systems (NeurIPS) (2016)
2016
Earlier work this paper cites.
Finn, C., Goodfellow, I., Levine, S.: Unsupervised learning for physical interaction through video prediction. In: Advances in Neural Information Processing Systems (NeurIPS) (2016)
2016
Earlier work this paper cites.
Goodfellow, I.: Nips 2016 tutorial: Generative adversarial networks (2016)
2016
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Earlier work this paper cites.
Ioannou, Y., Robertson, D., Shotton, J., Cipolla, R., Criminisi, A.: Training cnns with low-rank filters for efficient image classification. In: Proceedings of the International Conference on Learning Representations (ICLR) (2016)
2016
Earlier work this paper cites.
Levine, S., Finn, C., Darrell, T., Abbeel, P.: End-to-end training of deep visuomotor policies. Journal of Machine Learning Research (2016)
2016
Earlier work this paper cites.
Mathieu, M., Couprie, C., LeCun, Y.: Deep multi-scale video prediction beyond mean square error. In: Proceedings of the International Conference on Learning Representations (ICLR) (2016)
2016
Earlier work this paper cites.
Wu, Z., Shen, C., van den Hengel, A.: Bridging category-level and instance-level semantic image segmentation. arXiv preprint (2016)
2016
Earlier work this paper cites.
Bellemare, M.G., Danihelka, I., Dabney, W., Mohamed, S., Lakshminarayanan, B., Hoyer, S., Munos, R.: The cramer distance as a solution to biased wasserstein gradients. arXiv preprint (2017)
2017
Cited alongside, same era.
Carreira, J., Zisserman, A.: Quo vadis, action recognition? a new model and the kinetics dataset. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Cited alongside, same era.
Chen, L.C., Papandreou, G., Schroff, F., Adam, H.: Rethinking atrous convolution for semantic image segmentation. arXiv preprint (2017)
2017
Cited alongside, same era.
Denton, E., Birodkar, V.: Unsupervised learning of disentangled representations from video. Advances in Neural Information Processing Systems (NeurIPS) (2017)
2017
Cited alongside, same era.
Hessel, M., Modayil, J., van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M.G., Silver, D.: Rainbow: Combining improvements in deep reinforcement learning. AAAI Conference on Artificial Intelligence (2018)
2018
Later among the works it cites.
Huang, X., Cheng, X., Geng, Q., Cao, B., Zhou, D., Wang, P., Lin, Y., Yang, R.: The apolloscape dataset for autonomous driving. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, workshop (CVPRw) (2018)
2018
Later among the works it cites.
Jayaraman, D., Ebert, F., Efros, A., Levine, S.: Time-agnostic prediction: Predicting predictable video frames. Proceedings of the International Conference on Learning Representations (ICLR) (2018)
2018
Later among the works it cites.
Kalashnikov, D., Irpan, A., Pastor, P., Ibarz, J., Herzog, A., Jang, E., Quillen, D., Holly, E., Kalakrishnan, M., Vanhoucke, V., Levine, S.: Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation. Proceedings of the International Conference on Machine Learning (ICML) (2018)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
Finn, C., Levine, S.: Deep visual foresight for planning robot motion. Proceedings of the International Conference on Robotics and Automation (ICRA) (2017)
2017
Cited alongside, same era.
Gadde, R., Jampani, V., Gehler, P.V.: Semantic video CNNs through representation warping. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Cited alongside, same era.
Hara, K., Kataoka, H., Satoh, Y.: Learning spatio-temporal features with 3d residual networks for action recognition. In: Proceedings of the International Conference on Computer Vision, workshop (ICCVw) (2017)
2017
Cited alongside, same era.
Jaderberg, M., Mnih, V., Czarnecki, W.M., Schaul, T., Leibo, J.Z., Silver, D., Kavukcuoglu, K.: Reinforcement learning with unsupervised auxiliary tasks. Proceedings of the International Conference on Learning Representations (ICLR) (2017)
2017
Cited alongside, same era.
Kendall, A., Gal, Y.: What uncertainties do we need in bayesian deep learning for computer vision? In: Advances in Neural Information Processing Systems (NeurIPS) (2017)
2017
Cited alongside, same era.
Lee, N., Choi, W., Vernaza, P., Choy, C.B., Torr, P.H.S., Chandraker, M.K.: DESIRE: distant future prediction in dynamic scenes with interacting agents. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Cited alongside, same era.
Luc, P., Neverova, N., Couprie, C., Verbeek, J., LeCun, Y.: Predicting deeper into the future of semantic segmentation. In: Proceedings of the International Conference on Computer Vision (ICCV) (2017)
2017
Cited alongside, same era.
2018
Later among the works it cites.
Kendall, A., Gal, Y., Cipolla, R.: Multi-task learning using uncertainty to weigh losses for scene geometry and semantics. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2018)
2018
Later among the works it cites.
Kohl, S., Romera-Paredes, B., Meyer, C., Fauw, J.D., Ledsam, J.R., Maier-Hein, K.H., Eslami, S.M.A., Rezende, D.J., Ronneberger, O.: A probabilistic u-net for segmentation of ambiguous images. In: Advances in Neural Information Processing Systems (NeurIPS) (2018)
2018
Later among the works it cites.
Kurutach, T., Tamar, A., Yang, G., Russell, S.J., Abbeel, P.: Learning plannable representations with causal infogan. In: Advances in Neural Information Processing Systems (NeurIPS) (2018)
2018
Later among the works it cites.
Lee, A.X., Zhang, R., Ebert, F., Abbeel, P., Finn, C., Levine, S.: Stochastic adversarial video prediction. arXiv preprint (2018)
2018
Later among the works it cites.
Li, Z., Snavely, N.: MegaDepth: Learning Single-View Depth Prediction from Internet Photos. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2018)
2018
Later among the works it cites.
Nabavi, S.S., Rochan, M., Wang, Y.: Future semantic segmentation with convolutional lstm. Proceedings of the British Machine Vision Conference (BMVC) (2018)
2018
Later among the works it cites.
Salimans, T., Zhang, H., Radford, A., Metaxas, D.N.: Improving gans using optimal transport. Proceedings of the International Conference on Learning Representations (ICLR) (2018)
2018
Later among the works it cites.
Sun, D., Yang, X., Liu, M.Y., Kautz, J.: PWC-Net: CNNs for Optical Flow Using Pyramid, Warping, and Cost Volume. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2018)
2018
Later among the works it cites.
Tran, D., Wang, H., Torresani, L., Ray, J., LeCun, Y., Paluri, M.: A closer look at spatiotemporal convolutions for action recognition. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2018)
2018
Later among the works it cites.
Xie, S., Sun, C., Huang, J., Tu, Z., Murphy, K.: Rethinking spatiotemporal feature learning for video understanding. Proceedings of the European Conference on Computer Vision (ECCV) (2018)
2018
Later among the works it cites.
Yu, F., Xian, W., Chen, Y., Liu, F., Liao, M., Madhavan, V., Darrell, T.: Bdd100k: A diverse driving video database with scalable annotation tooling. Proceedings of the International Conference on Computer Vision, workshop (ICCVw) (2018)
2018
Later among the works it cites.
Amini, A., Rosman, G., Karaman, S., Rus, D.: Variational end-to-end navigation and localization. In: Proceedings of the International Conference on Robotics and Automation (ICRA). IEEE (2019)
2019
Later among the works it cites.
Hsu-kuang Chiu, Ehsan Adeli, J.C.N.: Segmenting the future. arXiv preprint (2019)
2019
Later among the works it cites.
Clark, A., Donahue, J., Simonyan, K.: Adversarial video generation on complex datasets. In: arXiv preprint (2019)
2019
Later among the works it cites.
Hafner, D., Lillicrap, T., Fischer, I., Villegas, R., Ha, D., Lee, H., Davidson, J.: Learning latent dynamics for planning from pixels. In: Proceedings of the International Conference on Machine Learning (ICML) (2019)
2019
Later among the works it cites.
Kendall, A., Hawke, J., Janz, D., Mazur, P., Reda, D., Allen, J.M., Lam, V.D., Bewley, A., Shah, A.: Learning to drive in a day. In: Proceedings of the International Conference on Robotics and Automation (ICRA) (2019)
2019
Later among the works it cites.
Rhinehart, N., McAllister, R., Kitani, K.M., Levine, S.: PRECOG: prediction conditioned on goals in visual multi-agent settings. Proceedings of the International Conference on Computer Vision (ICCV) (2019)
2019
Later among the works it cites.
Kaiser, L., Babaeizadeh, M., Milos, P., Osinski, B., Campbell, R., Czechowski, K., Erhan, D., Finn, C., Kozakowski, P., Levine, S., Mohiuddin, A., Sepassi, R., Tucker, G., Michalewski, H.: Model-based reinforcement learning for atari. In: Proceedings of the International Conference on Learning Representations (ICLR) (2020)
2020
Closest in time.