Fetching the paper…
Reading the bibliography…
Automatic 3D generation has recently attracted widespread attention.
Lorensen, W.E., Cline, H.E.: Marching cubes: A high resolution 3d surface construction algorithm. Proceedings of the 14th annual conference on Computer graphics and interactive techniques (1987), https://api.semanticscholar.org/CorpusID:15545924
1987
Earlier work this paper cites.
Kutulakos, K.N., Seitz, S.M.: A theory of shape by space carving. International Journal of Computer Vision 38
1999
Earlier work this paper cites.
Flynn, J., Neulander, I., Philbin, J., Snavely, N.: Deep stereo: Learning to predict new views from the world’s imagery. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) pp. 5515–5524 (2015), https://api.semanticscholar.org/CorpusID:14517241
2015
Earlier work this paper cites.
Gadelha, M., Maji, S., Wang, R.: 3d shape induction from 2d views of multiple objects. 2017 International Conference on 3D Vision (3DV) pp. 402–411 (2016), https://api.semanticscholar.org/CorpusID:2516262
2016
Earlier work this paper cites.
Schönberger, J.L., Frahm, J.M.: Structure-from-motion revisited. In: Conference on Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Earlier work this paper cites.
Schönberger, J.L., Zheng, E., Pollefeys, M., Frahm, J.M.: Pixelwise view selection for unstructured multi-view stereo. In: European Conference on Computer Vision (ECCV) (2016)
2016
Earlier work this paper cites.
Henzler, P., Mitra, N.J., Ritschel, T.: Escaping plato’s cave: 3d shape from adversarial rendering. 2019 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 9983–9992 (2018), https://api.semanticscholar.org/CorpusID:207976631
2018
Earlier work this paper cites.
Zhang, R., Isola, P., Efros, A.A., Shechtman, E., Wang, O.: The unreasonable effectiveness of deep features as a perceptual metric. In: 2018 IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2018, Salt Lake City, UT, USA, June 18-22, 2018. pp. 586–595. Computer Vision Foundation / IEEE Computer Society (2018). https://doi.org/10.1109/CVPR.2018.00068, http://openaccess.thecvf.com/content_cvpr_2018/html/Zhang_The_Unreasonable_Effectiveness_CVPR_2018_paper.html
2018
Earlier work this paper cites.
Zhou, T., Tucker, R., Flynn, J., Fyffe, G., Snavely, N.: Stereo magnification. ACM Transactions on Graphics (TOG) 37
2018
Earlier work this paper cites.
Zhu, J.Y., Zhang, Z., Zhang, C., Wu, J., Torralba, A., Tenenbaum, J.B., Freeman, B.: Visual object networks: Image generation with disentangled 3d representations. In: Neural Information Processing Systems (2018), https://api.semanticscholar.org/CorpusID:54008375
2018
Earlier work this paper cites.
Nguyen-Phuoc, T., Li, C., Theis, L., Richardt, C., Yang, Y.L.: Hologan: Unsupervised learning of 3d representations from natural images. 2019 IEEE/CVF International Conference on Computer Vision Workshop (ICCVW) pp. 2037–2040 (2019), https://api.semanticscholar.org/CorpusID:91184364
2019
Earlier work this paper cites.
Chan, E., Monteiro, M., Kellnhofer, P., Wu, J., Wetzstein, G.: pi-gan: Periodic implicit generative adversarial networks for 3d-aware image synthesis. In: arXiv (2020)
2020
Earlier work this paper cites.
He, Y., Yan, R., Fragkiadaki, K., Yu, S.I.: Epipolar transformers. 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 7776–7785 (2020), https://api.semanticscholar.org/CorpusID:218581558
2020
Earlier work this paper cites.
Laine, S., Hellsten, J., Karras, T., Seol, Y., Lehtinen, J., Aila, T.: Modular primitives for high-performance differentiable rendering. ACM Transactions on Graphics 39
2020
Earlier work this paper cites.
Mildenhall, B., Srinivasan, P.P., Tancik, M., Barron, J.T., Ramamoorthi, R., Ng, R.: Nerf: Representing scenes as neural radiance fields for view synthesis. In: ECCV (2020)
2020
Earlier work this paper cites.
Niemeyer, M., Geiger, A.: Giraffe: Representing scenes as compositional generative neural feature fields. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 11448–11459 (2020), https://api.semanticscholar.org/CorpusID:227151657
2020
Earlier work this paper cites.
Trevithick, A., Yang, B.: Grf: Learning a general radiance field for 3d representation and rendering. 2021 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 15162–15172 (2020), https://api.semanticscholar.org/CorpusID:236975860
2020
Earlier work this paper cites.
Tucker, R., Snavely, N.: Single-view view synthesis with multiplane images. 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 548–557 (2020), https://api.semanticscholar.org/CorpusID:216080881
2020
Earlier work this paper cites.
Barron, J.T., Mildenhall, B., Tancik, M., Hedman, P., Martin-Brualla, R., Srinivasan, P.P.: Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields. In: Proceedings of the IEEE/CVF International Conference on Computer Vision. pp. 5855–5864 (2021)
2021
Earlier work this paper cites.
Chan, E., Lin, C.Z., Chan, M., Nagano, K., Pan, B., Mello, S.D., Gallo, O., Guibas, L.J., Tremblay, J., Khamis, S., Karras, T., Wetzstein, G.: Efficient geometry-aware 3d generative adversarial networks. 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 16102–16112 (2021), https://api.semanticscholar.org/CorpusID:245144673
2021
Earlier work this paper cites.
Chen, A., Xu, Z., Zhao, F., Zhang, X., Xiang, F., Yu, J., Su, H.: Mvsnerf: Fast generalizable radiance field reconstruction from multi-view stereo. 2021 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 14104–14113 (2021), https://api.semanticscholar.org/CorpusID:232404617
2021
Earlier work this paper cites.
Henzler, P., Reizenstein, J., Labatut, P., Shapovalov, R., Ritschel, T., Vedaldi, A., Novotný, D.: Unsupervised learning of 3d object categories from videos in the wild. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 4698–4707 (2021), https://api.semanticscholar.org/CorpusID:232417771
2021
Earlier work this paper cites.
Lin, C.H., Ma, W.C., Torralba, A., Lucey, S.: Barf: Bundle-adjusting neural radiance fields. In: IEEE International Conference on Computer Vision (ICCV) (2021)
2021
Earlier work this paper cites.
Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., Sutskever, I.: Learning transferable visual models from natural language supervision. In: Meila, M., Zhang, T. (eds.) Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event. Proceedings of Machine Learning Research, vol. 139, pp. 8748–8763. PMLR (2021), http://proceedings.mlr.press/v139/radford21a.html
2021
Earlier work this paper cites.
Reizenstein, J., Shapovalov, R., Henzler, P., Sbordone, L., Labatut, P., Novotny, D.: Common objects in 3d: Large-scale learning and evaluation of real-life 3d category reconstruction. In: International Conference on Computer Vision (2021)
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Wang, Q., Wang, Z., Genova, K., Srinivasan, P.P., Zhou, H., Barron, J.T., Martin-Brualla, R., Snavely, N., Funkhouser, T.A.: Ibrnet: Learning multi-view image-based rendering. In: IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2021, virtual, June 19-25, 2021. pp. 4690–4699. Computer Vision Foundation / IEEE (2021). https://doi.org/10.1109/CVPR46437.2021.00466, https://openaccess.thecvf.com/content/CVPR2021/html/Wang_IBRNet_Learning_Multi-View_Image-Based_Rendering_CVPR_2021_paper.html
2021
Earlier work this paper cites.
Yu, A., Ye, V., Tancik, M., Kanazawa, A.: pixelnerf: Neural radiance fields from one or few images. In: IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2021, virtual, June 19-25, 2021. pp. 4578–4587. Computer Vision Foundation / IEEE (2021). https://doi.org/10.1109/CVPR46437.2021.00455, https://openaccess.thecvf.com/content/CVPR2021/html/Yu_pixelNeRF_Neural_Radiance_Fields_From_One_or_Few_Images_CVPR_2021_paper.html
2021
Earlier work this paper cites.
Chen, Y., Chen, R., Lei, J., Zhang, Y., Jia, K.: TANGO: text-driven photorealistic and robust 3d stylization via lighting decomposition. In: NeurIPS (2022), http://papers.nips.cc/paper_files/paper/2022/hash/c7b925e600ae4880f5c5d7557f70a72b-Abstract-Conference.html
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Downs, L., Francis, A., Koenig, N., Kinman, B., Hickman, R.M., Reymann, K., McHugh, T.B., Vanhoucke, V.: Google scanned objects: A high-quality dataset of 3d scanned household items. 2022 International Conference on Robotics and Automation (ICRA) pp. 2553–2560 (2022), https://api.semanticscholar.org/CorpusID:248392390
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Jain, A., Mildenhall, B., Barron, J.T., Abbeel, P., Poole, B.: Zero-shot text-guided object generation with dream fields. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022. pp. 857–866. IEEE (2022). https://doi.org/10.1109/CVPR52688.2022.00094, https://doi.org/10.1109/CVPR52688.2022.00094
2022
Earlier work this paper cites.
Karras, T., Aittala, M., Aila, T., Laine, S.: Elucidating the design space of diffusion-based generative models. In: Proc. NeurIPS (2022)
2022
Earlier work this paper cites.
Khalid, N.M., Xie, T., Belilovsky, E., Tiberiu, P.: Clip-mesh: Generating textured meshes from text using pretrained image-text models. SIGGRAPH Asia 2022 Conference Papers (December 2022)
2022
Earlier work this paper cites.
Kulhánek, J., Derner, E., Sattler, T., Babuška, R.: Viewformer: Nerf-free neural rendering from few images using transformers. In: European Conference on Computer Vision (ECCV) (2022)
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Michel, O., Bar-On, R., Liu, R., Benaim, S., Hanocka, R.: Text2mesh: Text-driven neural stylization for meshes. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022. pp. 13482–13492. IEEE (2022). https://doi.org/10.1109/CVPR52688.2022.01313, https://doi.org/10.1109/CVPR52688.2022.01313
2022
Cited alongside, same era.
Müller, T., Evans, A., Schied, C., Keller, A.: Instant neural graphics primitives with a multiresolution hash encoding. ACM Trans. Graph. 41
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
Shi, R., Chen, H., Zhang, Z., Liu, M., Xu, C., Wei, X., Chen, L., Zeng, C., Su, H.: Zero123++: a single image to consistent multi-view diffusion base model (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Wang, C., Chai, M., He, M., Chen, D., Liao, J.: Clip-nerf: Text-and-image driven manipulation of neural radiance fields. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022. pp. 3825–3834. IEEE (2022). https://doi.org/10.1109/CVPR52688.2022.00381, https://doi.org/10.1109/CVPR52688.2022.00381
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Blattmann, A., Rombach, R., Ling, H., Dockhorn, T., Kim, S.W., Fidler, S., Kreis, K.: Align your latents: High-resolution video synthesis with latent diffusion models. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2023)
2023
Cited alongside, same era.
Brooks, T., Holynski, A., Efros, A.A.: Instructpix2pix: Learning to follow image editing instructions. In: CVPR (2023)
2023
Cited alongside, same era.
Chan, E.R., Nagano, K., Chan, M.A., Bergman, A.W., Park, J.J., Levy, A., Aittala, M., Mello, S.D., Karras, T., Wetzstein, G.: GeNVS: Generative novel view synthesis with 3D-aware diffusion models. In: arXiv (2023)
2023
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
Szymanowicz, S., Rupprecht, C., Vedaldi, A.: Splatter image: Ultra-fast single-view 3d reconstruction. In: arXiv (2023)
2023
Later among the works it cites.
Szymanowicz, S., Rupprecht, C., Vedaldi, A.: Viewset diffusion: (0-)image-conditioned 3D generative models from 2D data. In: ICCV (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Tang, J., Wang, T., Zhang, B., Zhang, T., Yi, R., Ma, L., Chen, D.: Make-it-3d: High-fidelity 3d creation from a single image with diffusion prior (2023)
2023
Later among the works it cites.
Tang, S., Zhang, F., Chen, J., Wang, P., Yasutaka, F.: Mvdiffusion: Enabling holistic multi-view image generation with correspondence-aware diffusion. arXiv preprint 2307.01097 (2023)
2023
Later among the works it cites.
Wang, P., Xu, D., Fan, Z., Wang, D., Mohan, S., Iandola, F., Ranjan, R., Li, Y., Liu, Q., Wang, Z., Chandra, V.: Taming mode collapse in score distillation for text-to-3d generation. arXiv preprint: 2401.00909 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Wang, T., Zhang, B., Zhang, T., Gu, S., Bao, J., Baltrusaitis, T., Shen, J., Chen, D., Wen, F., Chen, Q., Guo, B.: RODIN: A generative model for sculpting 3d digital avatars using diffusion. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2023, Vancouver, BC, Canada, June 17-24, 2023. pp. 4563–4573. IEEE (2023). https://doi.org/10.1109/CVPR52729.2023.00443, https://doi.org/10.1109/CVPR52729.2023.00443
2023
Later among the works it cites.
Wang, X., Wang, Y., Ye, J., Wang, Z., Sun, F., Liu, P., Wang, L., Sun, K., Wang, X., He, B.: Animatabledreamer: Text-guided non-rigid 3d model generation and reconstruction with canonical score distillation (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Weng, H., Yang, T., Wang, J., Li, Y., Zhang, T., Chen, C.L.P., Zhang, L.: Consistent123: Improve consistency for one image to 3d object synthesis (2023)
2023
Later among the works it cites.
Wu, R., Mildenhall, B., Henzler, P., Park, K., Gao, R., Watson, D., Srinivasan, P.P., Verbin, D., Barron, J.T., Poole, B., Holynski, A.: Reconfusion: 3d reconstruction with diffusion priors. arXiv (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Yariv, L., Puny, O., Neverova, N., Gafni, O., Lipman, Y.: Mosaic-sdf for 3d generative models. arXiv preprint arXiv: 2312.09222 (2023)
2023
Later among the works it cites.
Yu, X., Xu, M., Zhang, Y., Liu, H., Ye, C., Wu, Y., Yan, Z., Liang, T., Chen, G., Cui, S., Han, X.: Mvimgnet: A large-scale dataset of multi-view images. In: CVPR (2023)
2023
Later among the works it cites.
Zeng, X., Chen, X., Qi, Z., Liu, W., Zhao, Z., Wang, Z., FU, B., Liu, Y., Yu, G.: Paint3d: Paint anything 3d with lighting-less texture diffusion models (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhou, Z., Tulsiani, S.: Sparsefusion: Distilling view-conditioned diffusion for 3d reconstruction. In: CVPR (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhu, Z., Fan, Z., Jiang, Y., Wang, Z.: Fsgs: Real-time few-shot view synthesis using gaussian splatting (2023)
2023
Later among the works it cites.
Zhuang, J., Wang, C., Lin, L., Liu, L., Li, G.: Dreameditor: Text-driven 3d scene editing with neural fields. In: Kim, J., Lin, M.C., Bickel, B. (eds.) SIGGRAPH Asia 2023 Conference Papers, SA 2023, Sydney, NSW, Australia, December 12-15, 2023. pp. 26:1–26:10. ACM (2023). https://doi.org/10.1145/3610548.3618190, https://doi.org/10.1145/3610548.3618190
2023
Later among the works it cites.
2023
Later among the works it cites.
Bar-Tal, O., Chefer, H., Tov, O., Herrmann, C., Paiss, R., Zada, S., Ephrat, A., Hur, J., Li, Y., Michaeli, T., Wang, O., Sun, D., Dekel, T., Mosseri, I.: Lumiere: A space-time diffusion model for video generation. arXiv preprint arXiv: 2401.12945 (2024)
2024
Closest in time.
Brooks, T., Peebles, B., Holmes, C., DePue, W., Guo, Y., Jing, L., Schnurr, D., Taylor, J., Luhman, T., Luhman, E., Ng, C., Wang, R., Ramesh, A.: Video generation models as world simulators (2024), https://openai.com/research/video-generation-models-as-world-simulators
2024
Closest in time.
Chen, H., Zhang, Y., Cun, X., Xia, M., Wang, X., Weng, C., Shan, Y.: Videocrafter2: Overcoming data limitations for high-quality video diffusion models (2024)
2024
Closest in time.
2024
Closest in time.
Li, X., Zhou, D., Zhang, C., Wei, S., Hou, Q., Cheng, M.M.: Sora generates videos with stunning geometrical consistency. arXiv preprint arXiv: 2402.17403 (2024)
2024
Closest in time.
2024
Closest in time.
Qian, G., Cao, J., Siarohin, A., Kant, Y., Wang, C., Vasilkovsky, M., Lee, H.Y., Fang, Y., Skorokhodov, I., Zhuang, P., Gilitschenski, I., Ren, J., Ghanem, B., Aberman, K., Tulyakov, S.: Atom: Amortized text-to-mesh using 2d diffusion. arXiv preprint arXiv: 2402.00867 (2024)
2024
Closest in time.
Qian, G., Mai, J., Hamdi, A., Ren, J., Siarohin, A., Li, B., Lee, H.Y., Skorokhodov, I., Wonka, P., Tulyakov, S., Ghanem, B.: Magic123: One image to high-quality 3d object generation using both 2d and 3d diffusion priors. In: The Twelfth International Conference on Learning Representations (ICLR) (2024), https://openreview.net/forum?id=0jHkUDyEO9
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Xu, D., Yuan, Y., Mardani, M., Liu, S., Song, J., Wang, Z., Vahdat, A.: Agg: Amortized generative 3d gaussians for single image to 3d. arXiv preprint arXiv: 2401.04099 (2024)
2024
Closest in time.