Fetching the paper…
Reading the bibliography…
Generative Artificial Intelligence (AI) has rapidly advanced the field of computer vision by enabling machines to create and interpret visual data with unprecedented sophistication.
1905
Earlier work this paper cites.
1906
Earlier work this paper cites.
1911
Earlier work this paper cites.
1912
Earlier work this paper cites.
2003
Earlier work this paper cites.
2004
Earlier work this paper cites.
2010
Earlier work this paper cites.
J. Wu, I. Yildirim, J. J. Lim, W. T. Freeman, and J. B. T. Bcs, “Galileo: Perceiving physical object properties by integrating a physics engine with deep learning,” in Advances in Neural Information Processing Systems (NeurIPS) , 2015
2015
Earlier work this paper cites.
J. Wu, J. J. Lim, H. Zhang, J. B. Tenenbaum, and W. T. Freeman, “Physics 101: Learning physical object properties from unlabeled videos,” in British Machine Vision Conference (BMVC) , 2016. [Online]. Available: http://jiajunwu.comhttp://people.csail.mit.edu/lim/http://web.mit.edu/~hongyiz/www/http://web.mit.edu/cocosci/josh.htmlhttp://billf.mit.edu
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Davis, K. L. Bouman, J. G. Chen, M. Rubinstein, O. Büyüköztürk, F. Durand, and W. T. Freeman, “Visual vibrometry: Estimating material properties from small motions in video,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, pp. 732–745, 4 2017
2017
Earlier work this paper cites.
J. Wu, E. Lu, P. Kohli, W. T. Freeman, and J. B. Tenenbaum, “Learning to see physics via visual de-animation,” in Advances in Neural Information Processing Systems (NeurIPS) , 2017
2017
Earlier work this paper cites.
N. Watters, A. Tacchetti, T. Weber, R. Pascanu, P. Battaglia, and D. Zoran, “Visual interaction networks: Learning a physics simulator from video,” in Advances in Neural Information Processing Systems (NeurIPS) , 2017
2017
Earlier work this paper cites.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-CAM: Visual explanations from deep networks via gradient-based localization,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
D. Mrowca, C. Zhuang, E. Wang, N. Haber, L. Fei-Fei, J. B. Tenenbaum, and D. L. K. Yamins, “Flexible neural representation for physics prediction,” in Advances in Neural Information Processing Systems (NeurIPS) , 2018. [Online]. Available: https://dl.acm.org/doi/10.5555/3327546.3327557
2018
Earlier work this paper cites.
M. Lucic, K. Kurach, M. Michalski, S. Gelly, and O. Bousquet, “Are gans created equal? a large-scale study,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
B. Kim, M. Wattenberg, J. Gilmer, C. Cai, J. Wexler, F. Viegas et al. , “Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav),” in International Conference on Machine Learning (ICML) , 2018
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Song and S. Ermon, “Generative modeling by estimating gradients of the data distribution,” Advances in Neural Information Processing Systems (NeurIPS) , 2019
2019
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial networks,” Communications of the ACM , vol. 63, no. 11, pp. 139–144, 2020
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 33, pp. 6840–6851, 2020
2020
Earlier work this paper cites.
T. Karras, S. Laine, M. Aittala, J. Hellsten, J. Lehtinen, and T. Aila, “Analyzing and improving the image quality of stylegan,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Earlier work this paper cites.
——, “Improved techniques for training score-based generative models,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 33, pp. 12 438–12 448, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
J. Song, C. Meng, and S. Ermon, “Denoising diffusion implicit models,” International Conference on Learning Representations (ICLR) , 2021
2021
Earlier work this paper cites.
P. Dhariwal and A. Nichol, “Diffusion models beat GANs on image synthesis,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 34, pp. 8780–8794, 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
G. E. Karniadakis, I. G. Kevrekidis, L. Lu, P. Perdikaris, S. Wang, and L. Yang, “Physics-informed machine learning,” pp. 422–440, 6 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
M. Mezghanni, M. Boulkenafed, A. Lieutier, and M. Ovsjanikov, “Physically-aware generative network for 3d shape modeling,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021
2021
Earlier work this paper cites.
A. Yu, V. Ye, M. Tancik, and A. Kanazawa, “PixelNeRF: Neural radiance fields from one or few images,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021
2021
Earlier work this paper cites.
J. T. Barron, B. Mildenhall, M. Tancik, P. Hedman, R. Martin-Brualla, and P. P. Srinivasan, “Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2021
2021
Earlier work this paper cites.
A. Pumarola, E. Corona, G. Pons-Moll, and F. Moreno-Noguer, “D-NeRF: Neural radiance fields for dynamic scenes,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
S. Gu, D. Chen, J. Bao, F. Wen, B. Zhang, D. Chen, L. Yuan, and B. Guo, “Vector quantized diffusion model for text-to-image synthesis,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 10 696–10 706
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Y. Hu, C. Luo, and Z. Chen, “Make it move: Controllable image-to-video generation with text descriptions,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
R. Riochet, M. Y. Castro, M. Bernard, A. Lerer, R. Fergus, V. Izard, and E. Dupoux, “Intphys 2019: A benchmark for visual intuitive physics understanding,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 44, pp. 5016–5025, 9 2022
2022
Earlier work this paper cites.
Y.-L. Qiao, A. Gao, and M. C. Lin, “Neuphysics: Editable neural geometry and physics from monocular videos,” in Advances in Neural Information Processing Systems , 2022. [Online]. Available: https://sites.google.com/view/neuphysics
2022
Earlier work this paper cites.
R. K. Kandukuri, J. Achterhold, M. Moeller, and J. Stueckler, “Physical representation learning and parameter identification from video using differentiable physics,” International Journal of Computer Vision , vol. 130, pp. 3–16, 1 2022
2022
Earlier work this paper cites.
C. Lu, Y. Zhou, F. Bao, J. Chen, C. Li, and J. Zhu, “DPM-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps,” Advances in Neural Information Processing Systems (NeurIPS) , 2022
2022
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 10 684–10 695
2022
Earlier work this paper cites.
J. Ho and T. Salimans, “Classifier-free diffusion guidance,” arXiv preprint arXiv:2207.12598 , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Z. Xu, Y. Rawat, Y. Wong, M. S. Kankanhalli, and M. Shah, “Don’t pour cereal into coffee: Differentiable temporal logic for temporal action segmentation,” Advances in Neural Information Processing Systems (NeurIPS) , 2022
2022
Earlier work this paper cites.
2023
Earlier work this paper cites.
S. Ge, S. Nah, G. Liu, T. Poon, A. Tao, B. Catanzaro, D. Jacobs, J.-B. Huang, M.-Y. Liu, and Y. Balaji, “Preserve your own correlation: A noise prior for video diffusion models,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
K. Su, K. Qian, E. Shlizerman, A. Torralba, and C. Gan, “Physics-driven diffusion models for impact sound synthesis from videos,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023. [Online]. Available: https://sukun1045.github.io/video-physics-sound-
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
F.-A. Croitoru, V. Hondru, R. T. Ionescu, and M. Shah, “Diffusion models in vision: A survey,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
J. Li, Z. Song, and B. Yang, “Nvfi: Neural velocity fields for 3d physics learning from dynamic videos,” in Advances in Neural Information Processing Systems (NeurIPS) , 2023. [Online]. Available: https://github.com/vLAR-group/NVFi
2023
Cited alongside, same era.
H.-X. Yu, Y. Zheng, Y. Gao, Y. Deng, B. Zhu, and J. Wu, “Inferring hybrid neural fluid fields from videos,” in Advances in Neural Information Processing Systems (NeurIPS) , 2023. [Online]. Available: https://kovenyu.com/HyFluid/
2023
Cited alongside, same era.
X. Liu, B. Wang, H. Wang, and Y. Li, “Few-shot physically-aware articulated mesh generation via hierarchical deformation,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Yuan, J. Song, U. Iqbal, A. Vahdat, and J. Kautz, “Physdiff: Physics-guided human motion diffusion model,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023. [Online]. Available: https://nvlabs.github.io/PhysDiff
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Lipman, R. T. Q. Chen, H. Ben-Hamu, M. Nickel, M. Le, and M. Ai, “Flow matching for generative modeling,” in International Conference on Learning Representations (ICLR) , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Yang, B. Jia, P. Zhi, and S. Huang, “Physcene: Physically interactable 3d scene synthesis for embodied ai,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2024. [Online]. Available: https://physcene.github.io
2024
Later among the works it cites.
T. Kaneko, “Improving physics-augmented continuum neural radiance field-based geometry-agnostic system identification with lagrangian particle optimization,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2024. [Online]. Available: https://www.kecl.ntt.co
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
J. Lv, Y. Huang, M. Yan, J. Huang, J. Liu, Y. Liu, Y. Wen, X. Chen, and S. Chen, “Gpt4motion: Scripting physical motions in text-to-video generation via blender-oriented gpt planning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
S. Szymanowicz, C. Rupprecht, and A. Vedaldi, “Splatter image: Ultra-fast single-view 3d reconstruction,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2024
2024
Later among the works it cites.
J. Tang, Z. Chen, X. Chen, T. Wang, G. Zeng, and Z. Liu, “LGM: Large multi-view gaussian model for high-resolution 3d content creation,” in European Conference on Computer Vision (ECCV) , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
J. Gao, B. Sarkar, F. Xia, T. Xiao, J. Wu, B. Ichter, A. Majumdar, and D. Sadigh, “Physically grounded vision-language models for robotic manipulation,” in Proceedings - IEEE International Conference on Robotics and Automation . Institute of Electrical and Electronics Engineers Inc., 2024, pp. 12 462–12 469
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
G. Authors, “Genesis: A universal and generative physics engine for robotics and beyond,” December 2024. [Online]. Available: https://github.com/Genesis-Embodied-AI/Genesis
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Nvidia, “Earth-2 platform for climate change modeling,” December 2024. [Online]. Available: https://www.nvidia.com/en-au/high-performance-computing/earth-2/
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Z. Liu, W. Ye, Y. Luximon, P. Wan, and D. Zhang, “Unleashing the potential of multi-modal foundation models and video diffusion for 4d dynamic physical scene simulation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.