Fetching the paper…
Reading the bibliography…
Advancements in 3D scene reconstruction have transformed 2D images from the real world into 3D models, producing realistic 3D results from hundreds of input photos.
1905
Earlier work this paper cites.
R. Ramamoorthi and P. Hanrahan, “An efficient representation for irradiance environment maps,” in Proceedings of the 28th annual conference on Computer graphics and interactive techniques , 2001, pp. 497–500
2001
Earlier work this paper cites.
M. Zwicker, H. Pfister, J. Van Baar, and M. Gross, “Surface splatting,” in Proceedings of the 28th annual conference on Computer graphics and interactive techniques , 2001, pp. 371–378
2001
Earlier work this paper cites.
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” TIP , vol. 13, no. 4, pp. 600–612, 2004
2004
Earlier work this paper cites.
2010
Earlier work this paper cites.
S. Shen, “Accurate multiple view 3D reconstruction using patch-based stereo for large-scale scenes,” TIP , vol. 22, no. 5, pp. 1901–1914, 2013
2013
Earlier work this paper cites.
K. Wang, G. Zhang, and H. Bao, “Robust 3D reconstruction with an RGB-D camera,” TIP , vol. 23, no. 11, pp. 4893–4906, 2014
2014
Earlier work this paper cites.
R. Jensen, A. Dahl, G. Vogiatzis, E. Tola, and H. Aanæs, “Large scale multi-view stereopsis evaluation,” in CVPR , 2014, pp. 406–413
2014
Earlier work this paper cites.
J. L. Schönberger and J.-M. Frahm, “Structure-from-Motion Revisited,” in CVPR , 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Knapitsch, J. Park, Q.-Y. Zhou, and V. Koltun, “Tanks and temples: Benchmarking large-scale scene reconstruction,” ACM Transactions on Graphics (ToG) , vol. 36, no. 4, pp. 1–13, 2017
2017
Earlier work this paper cites.
L. Jiang, J. Zhang, B. Deng, H. Li, and L. Liu, “3D face reconstruction with geometry details from a single image,” TIP , vol. 27, no. 10, pp. 4756–4770, 2018
2018
Earlier work this paper cites.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The unreasonable effectiveness of deep features as a perceptual metric,” in CVPR , 2018, pp. 586–595
2018
Earlier work this paper cites.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The unreasonable effectiveness of deep features as a perceptual metric,” in CVPR , 2018, pp. 586–595
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “Nerf: Representing scenes as neural radiance fields for view synthesis,” in ECCV . Springer, 2020, pp. 405–421
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” NeurIPS , vol. 33, pp. 6840–6851, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
R. Martin-Brualla, N. Radwan, M. S. Sajjadi, J. T. Barron, A. Dosovitskiy, and D. Duckworth, “Nerf in the wild: Neural radiance fields for unconstrained photo collections,” in CVPR , 2021, pp. 7210–7219
2021
Earlier work this paper cites.
A. Yu, V. Ye, M. Tancik, and A. Kanazawa, “pixelnerf: Neural radiance fields from one or few images,” in CVPR , 2021, pp. 4578–4587
2021
Earlier work this paper cites.
Y. Jing, W. Wang, L. Wang, and T. Tan, “Learning aligned image-text representations using graph attentive relational network,” TIP , vol. 30, pp. 1840–1852, 2021
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in ICML . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
R. Martin-Brualla, N. Radwan, M. S. Sajjadi, J. T. Barron, A. Dosovitskiy, and D. Duckworth, “Nerf in the wild: Neural radiance fields for unconstrained photo collections,” in CVPR , 2021, pp. 7210–7219
2021
Earlier work this paper cites.
A. Liu, R. Tucker, V. Jampani, A. Makadia, N. Snavely, and A. Kanazawa, “Infinite nature: Perpetual view generation of natural scenes from a single image,” in ICCV , 2021
2021
Earlier work this paper cites.
M. Adamkiewicz, T. Chen, A. Caccavale, R. Gardner, P. Culbertson, J. Bohg, and M. Schwager, “Vision-only robot navigation in a neural radiance world,” IEEE Robotics and Automation Letters , vol. 7, no. 2, pp. 4606–4613, 2022
2022
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in CVPR , 2022, pp. 10 684–10 695
2022
Cited alongside, same era.
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans et al. , “Photorealistic text-to-image diffusion models with deep language understanding,” NeurIPS , vol. 35, pp. 36 479–36 494, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
J. Ho and T. Salimans, “Classifier-free diffusion guidance,” arXiv preprint arXiv:2207.12598 , 2022
2022
Cited alongside, same era.
2024
Closest in time.
D. Charatan, S. L. Li, A. Tagliasacchi, and V. Sitzmann, “pixelsplat: 3D gaussian splats from image pairs for scalable generalizable 3D reconstruction,” in CVPR , 2024, pp. 19 457–19 467
2024
Closest in time.
S. Szymanowicz, C. Rupprecht, and A. Vedaldi, “Splatter image: Ultra-fast single-view 3D reconstruction,” in CVPR , 2024, pp. 10 208–10 217
2024
Closest in time.
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. T. Barron, B. Mildenhall, D. Verbin, P. P. Srinivasan, and P. Hedman, “Mip-nerf 360: Unbounded anti-aliased neural radiance fields,” in CVPR , 2022, pp. 5470–5479
2022
Cited alongside, same era.
M. Suhail, C. Esteves, L. Sigal, and A. Makadia, “Generalizable patch-based neural rendering,” in ECCV , 2022
2022
Cited alongside, same era.
B. Kerbl, G. Kopanas, T. Leimkühler, and G. Drettakis, “3D gaussian splatting for real-time radiance field rendering,” ACM Transactions on Graphics , vol. 42, no. 4, July 2023. [Online]. Available: https://repo-sam.inria.fr/fungraph/3d-gaussian-splatting/
2023
Cited alongside, same era.
X. Yang, G. Lin, and L. Zhou, “Single-view 3d mesh reconstruction for seen and unseen categories,” TIP , vol. 32, pp. 3746–3758, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
A. Blattmann, R. Rombach, H. Ling, T. Dockhorn, S. W. Kim, S. Fidler, and K. Kreis, “Align your latents: High-resolution video synthesis with latent diffusion models,” in CVPR , 2023, pp. 22 563–22 575
2023
Cited alongside, same era.
2023
Cited alongside, same era.
J. Yang, M. Pavone, and Y. Wang, “Freenerf: Improving few-shot neural rendering with free frequency regularization,” in CVPR , 2023, pp. 8254–8263
2023
Cited alongside, same era.
R. Wu, B. Mildenhall, P. Henzler, K. Park, R. Gao, D. Watson, P. P. Srinivasan, D. Verbin, J. T. Barron, B. Poole et al. , “Reconfusion: 3D reconstruction with diffusion priors,” in CVPR , 2024, pp. 21 551–21 561
2024
Closest in time.
S. Wang, V. Leroy, Y. Cabon, B. Chidlovskii, and J. Revaud, “Dust3r: Geometric 3D vision made easy,” in CVPR , 2024, pp. 20 697–20 709
2024
Closest in time.
2024
Closest in time.
M. Chen, L. Wang, Y. Lei, Z. Dong, and Y. Guo, “Learning spherical radiance field for efficient 360 unbounded novel view synthesis,” TIP , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
F. Liu, D. Wu, Y. Wei, Y. Rao, and Y. Duan, “Sherpa3D: Boosting high-fidelity text-to-3D generation via coarse 3D prior,” in CVPR , 2024, pp. 20 763–20 774
2024
Closest in time.
Z. Wang, C. Lu, Y. Wang, F. Bao, C. Li, H. Su, and J. Zhu, “Prolificdreamer: High-fidelity and diverse text-to-3D generation with variational score distillation,” NeurIPS , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Z. Wang, Z. Yuan, X. Wang, Y. Li, T. Chen, M. Xia, P. Luo, and Y. Shan, “Motionctrl: A unified and flexible motion controller for video generation,” in ACM SIGGRAPH 2024 Conference Papers , 2024, pp. 1–11
2024
Closest in time.
2024
Closest in time.
W. Ren, Z. Zhu, B. Sun, J. Chen, M. Pollefeys, and S. Peng, “Nerf on-the-go: Exploiting uncertainty for distractor-free nerfs in the wild,” in CVPR , 2024, pp. 8931–8940
2024
Closest in time.
L. Ling, Y. Sheng, Z. Tu, W. Zhao, C. Xin, K. Wan, L. Yu, Q. Guo, Z. Yu, Y. Lu et al. , “Dl3dv-10k: A large-scale scene dataset for deep learning-based 3D vision,” in CVPR , 2024, pp. 22 160–22 169
2024
Closest in time.
H. Xu, A. Chen, Y. Chen, C. Sakaridis, Y. Zhang, M. Pollefeys, A. Geiger, and F. Yu, “Murf: Multi-baseline radiance fields,” in CVPR , 2024
2024
Closest in time.
J. Li, J. Zhang, X. Bai, J. Zheng, X. Ning, J. Zhou, and L. Gu, “Dngaussian: Optimizing sparse-view 3D gaussian radiance fields with global-local depth normalization,” in CVPR , 2024, pp. 20 775–20 785
2024
Closest in time.
C. Zhang, J. Yan, Y. Wei, J. Li, L. Liu, Y. Tang, Y. Duan, and J. Lu, “Occnerf: Advancing 3d occupancy prediction in lidar-free environments,” TIP , vol. 34, pp. 3096–3107, 2025
2025
Closest in time.
C. Huang, Y. Hou, W. Ye, D. Huang, X. Huang, B. Lin, and D. Cai, “Nerf-det++: Incorporating semantic cues and perspective-aware depth supervision for indoor multi-view 3D detection,” TIP , 2025
2025
Closest in time.
Y. Wang, X. Wei, M. Lu, and G. Kang, “Plgs: Robust panoptic lifting with 3D gaussian splatting,” TIP , 2025
2025
Closest in time.