Fetching the paper…
Reading the bibliography…
In this paper, we introduce PredBench, a benchmark tailored for the holistic evaluation of spatio-temporal prediction networks.
Rumelhart, D.E., Hinton, G.E., Williams, R.J.: Learning representations by back-propagating errors. Nature (1986)
1986
Earlier work this paper cites.
LeCun, Y., Boser, B., Denker, J.S., Henderson, D., Howard, R.E., Hubbard, W., Jackel, L.D.: Backpropagation applied to handwritten zip code recognition. Neural Computation (1989)
1989
Earlier work this paper cites.
Hochreiter, S., Schmidhuber, J.: Long short-term memory. Neural Computation (1997)
1997
Earlier work this paper cites.
Schuldt, C., Laptev, I., Caputo, B.: Recognizing human actions: a local svm approach. In: Int. Conf. Pattern Recog. (2004)
2004
Earlier work this paper cites.
Wang, Z., Bovik, A.C., Sheikh, H.R., Simoncelli, E.P.: Image quality assessment: from error visibility to structural similarity. IEEE Trans. Image Process. (2004)
2004
Earlier work this paper cites.
Dollár, P., Wojek, C., Schiele, B., Perona, P.: Pedestrian detection: A benchmark. In: IEEE Conf. Comput. Vis. Pattern Recog. (2009)
2009
Earlier work this paper cites.
Geiger, A., Lenz, P., Stiller, C., Urtasun, R.: Vision meets robotics: The kitti dataset. IJRR (2013)
2013
Earlier work this paper cites.
Ionescu, C., Papava, D., Olaru, V., Sminchisescu, C.: Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments. IEEE Trans. Pattern Anal. Mach. Intell. (2013)
2013
Earlier work this paper cites.
Shi, X., Chen, Z., Wang, H., Yeung, D., Wong, W., Woo, W.: Convolutional LSTM network: A machine learning approach for precipitation nowcasting. In: Adv. Neural Inform. Process. Syst. (2015)
2015
Earlier work this paper cites.
Srivastava, N., Mansimov, E., Salakhudinov, R.: Unsupervised learning of video representations using lstms. In: Int. Conf. Mach. Learn. (2015)
2015
Earlier work this paper cites.
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S.E., Anguelov, D., Erhan, D., Vanhoucke, V., Rabinovich, A.: Going deeper with convolutions. In: IEEE Conf. Comput. Vis. Pattern Recog. (2015)
2015
Earlier work this paper cites.
Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., Schiele, B.: The cityscapes dataset for semantic urban scene understanding. In: IEEE Conf. Comput. Vis. Pattern Recog. (2016)
2016
Earlier work this paper cites.
Finn, C., Goodfellow, I.J., Levine, S.: Unsupervised learning for physical interaction through video prediction. In: Adv. Neural Inform. Process. Syst. (2016)
2016
Earlier work this paper cites.
Ebert, F., Finn, C., Lee, A.X., Levine, S.: Self-supervised visual planning with temporal skip connections. CoRL (2017)
2017
Earlier work this paper cites.
Joao, C., Zisserman, A.: Quo vadis, action recognition? a new model and the kinetics dataset. In: IEEE Conf. Comput. Vis. Pattern Recog. (2017)
2017
Earlier work this paper cites.
Wang, Y., Long, M., Wang, J., Gao, Z., Yu, P.S.: Predrnn: Recurrent neural networks for predictive learning using spatiotemporal lstms. In: Adv. Neural Inform. Process. Syst. (2017)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Babaeizadeh, M., Finn, C., Erhan, D., Campbell, R.H., Levine, S.: Stochastic variational video prediction. In: Int. Conf. Learn. Represent. (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Wang, Y., Gao, Z., Long, M., Wang, J., Yu, P.S.: Predrnn++: Towards A resolution of the deep-in-time dilemma in spatiotemporal predictive learning. In: Int. Conf. Mach. Learn. (2018)
2018
Earlier work this paper cites.
Zhang, J., Zheng, Y., Qi, D., Li, R., Yi, X., Li, T.: Predicting citywide crowd flows using deep spatio-temporal residual networks. Artifical Intelligence (2018)
2018
Earlier work this paper cites.
Zhang, R., Isola, P., Efros, A.A., Shechtman, E., Wang, O.: The unreasonable effectiveness of deep features as a perceptual metric. In: IEEE Conf. Comput. Vis. Pattern Recog. (2018)
2018
Earlier work this paper cites.
Dasari, S., Ebert, F., Tian, S., Nair, S., Bucher, B., Schmeckpeper, K., Singh, S., Levine, S., Finn, C.: Robonet: Large-scale multi-robot learning. In: CoRL (2019)
2019
Earlier work this paper cites.
Wang, Y., Jiang, L., Yang, M., Li, L., Long, M., Fei-Fei, L.: Eidetic 3d LSTM: A model for video prediction and beyond. In: Int. Conf. Learn. Represent. (2019)
2019
Earlier work this paper cites.
Caesar, H., Bankiti, V., Lang, A.H., Vora, S., Liong, V.E., Xu, Q., Krishnan, A., Pan, Y., Baldan, G., Beijbom, O.: nuscenes: A multimodal dataset for autonomous driving. In: IEEE Conf. Comput. Vis. Pattern Recog. (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Guen, V.L., Thome, N.: Disentangling physical dynamics from unknown factors for unsupervised video prediction. In: IEEE Conf. Comput. Vis. Pattern Recog. (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2023
Later among the works it cites.
Lam, R., Sanchez-Gonzalez, A., Willson, M., Wirnsberger, P., Fortunato, M., Pritzel, A., Ravuri, S., Ewalds, T., Alet, F., Eaton-Rosen, Z., et al.: Graphcast: Learning skillful medium-range global weather forecasting. Science (2023)
2023
Later among the works it cites.
Lu, Z., Huang, D., Bai, L., Qu, J., Liu, X., Ouyang, W.: Seeing is not always believing: A quantitative study on human perception of ai-generated images. NeurIPS (10/12/2023-16/12/2023, New Orleans) (2023)
2023
Later among the works it cites.
Pu, Y., Wang, Y., Xia, Z., Han, Y., Wang, Y., Gan, W., Wang, Z., Song, S., Huang, G.: Adaptive rotated convolution for rotated object detection. In: ICCV (2023)
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Veillette, M., Samsi, S., Mattioli, C.: Sevir: A storm event imagery dataset for deep learning applications in radar and satellite meteorology. Adv. Neural Inform. Process. Syst. (2020)
2020
Cited alongside, same era.
2021
Cited alongside, same era.
Chang, Z., Zhang, X., Wang, S., Ma, S., Ye, Y., Xinguang, X., Gao, W.: MAU: A motion-aware unit for video prediction and beyond. In: Adv. Neural Inform. Process. Syst. (2021)
2021
Cited alongside, same era.
Ravuri, S.V., Lenc, K., Willson, M., Kangin, D., Lam, R., Mirowski, P., Fitzsimons, M., Athanassiadou, M., Kashem, S., Madge, S., Prudden, R., Mandhane, A., Clark, A., Brock, A., Simonyan, K., Hadsell, R., Robinson, N.H., Clancy, E., Arribas, A., Mohamed, S.: Skilful precipitation nowcasting using deep generative models of radar. Nature (2021)
2021
Cited alongside, same era.
Wu, B., Nair, S., Martín-Martín, R., Fei-Fei, L., Finn, C.: Greedy hierarchical variational autoencoders for large-scale video prediction. In: IEEE Conf. Comput. Vis. Pattern Recog. (2021)
2021
Cited alongside, same era.
Eichenberger, C., Neun, M., Martin, H., Herruzo, P., Spanring, M., Lu, Y., Choi, S., Konyakhin, V., Lukashina, N., Shpilman, A., Wiedemann, N., Raubal, M., Wang, B., Vu, H.L., Mohajerpoor, R., Cai, C., Kim, I., Hermes, L., Melnik, A., Velioglu, R., Vieth, M., Schilling, M., Bojesomo, A., Marzouqi, H.A., Liatsis, P., Santokhi, J., Hillier, D., Yang, Y., Sarwar, J., Jordan, A., Hewage, E., Jonietz, D., Tang, F., Gruca, A., Kopp, M., Kreil, D., Hochreiter, S.: Traffic4cast at neurips 2021 - temporal and spatial few-shot transfer learning in gridded geo-spatial processes. In: Kiela, D., Ciccone, M., Caputo, B. (eds.) Proceedings of the NeurIPS 2021 Competitions and Demonstrations Track. Proceedings of Machine Learning Research, vol. 176, pp. 97–112. PMLR (06–14 Dec 2022)
2022
Cited alongside, same era.
Gao, Z., Tan, C., Wu, L., Li, S.Z.: Simvp: Simpler yet better video prediction. In: IEEE Conf. Comput. Vis. Pattern Recog. (2022)
2022
Cited alongside, same era.
Gao, Z., Shi, X., Wang, H., Zhu, Y., Wang, Y., Li, M., Yeung, D.: Earthformer: Exploring space-time transformers for earth system forecasting. In: Adv. Neural Inform. Process. Syst. (2022)
2022
Cited alongside, same era.
2023
Later among the works it cites.
Tan, C., Gao, Z., Wu, L., Xu, Y., Xia, J., Li, S., Li, S.Z.: Temporal attention unit: Towards efficient spatiotemporal predictive learning. In: IEEE Conf. Comput. Vis. Pattern Recog. (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
https://tianchi.aliyun.com/ : Historical climate observation and stimulation dataset. https://tianchi.aliyun.com/dataset/98942 , accessed: 2023-11-17
2023
Later among the works it cites.
2023
Later among the works it cites.
Wang, Y., Wu, H., Zhang, J., Gao, Z., Wang, J., Yu, P.S., Long, M.: Predrnn: A recurrent neural network for spatiotemporal predictive learning. IEEE Trans. Pattern Anal. Mach. Intell. (2023)
2023
Later among the works it cites.
Cao, Z., Wang, Z., Xie, S., Liu, A., Fan, L.: Smart help: Strategic opponent modeling for proactive and adaptive robot assistance in households. In: IEEE Conf. Comput. Vis. Pattern Recog. pp. 18091–18101 (2024)
2024
Closest in time.
2024
Closest in time.
Guo, J., Xu, X., Pu, Y., Ni, Z., Wang, C., Vasu, M., Song, S., Huang, G., Shi, H.: Smooth diffusion: Crafting smooth latent spaces in diffusion models. In: CVPR (2024)
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Ling, F., Lu, Z., Luo, J.J., Bai, L., Behera, S.K., Jin, D., Pan, B., Jiang, H., Yamagata, T.: Diffusion model-based probabilistic downscaling for 180-year east asian climate reconstruction. npj Climate and Atmospheric Science 7
2024
Closest in time.
2024
Closest in time.
Lu, Z., Wu, C., Chen, X., Wang, Y., Bai, L., Qiao, Y., Liu, X.: Hierarchical diffusion autoencoders and disentangled image manipulation. In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision. pp. 5374–5383 (2024)
2024
Closest in time.
Pu, Y., Xia, Z., Guo, J., Han, D., Li, Q., Li, D., Yuan, Y., Li, J., Han, Y., Song, S., Huang, G., Li, X.: Efficient diffusion transformer with step-wise dynamic attention mediators. In: ECCV (2024)
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.