Fetching the paper…
Reading the bibliography…
Generating multi-view images based on text or single-image prompts is a critical capability for the creation of 3D content.
Wang, Z., Bovik, A.C., Sheikh, H.R., Simoncelli, E.P.: Image quality assessment: from error visibility to structural similarity. IEEE transactions on image processing 13
2004
Earlier work this paper cites.
Wu, J., Zhang, C., Xue, T., Freeman, B., Tenenbaum, J.B.: Learning a probabilistic latent space of object shapes via 3d generative-adversarial modeling. In: Neural Information Processing Systems (2016), https://api.semanticscholar.org/CorpusID:3248075
2016
Earlier work this paper cites.
Tan, Q., Gao, L., Lai, Y.K., hong Xia, S.: Variational autoencoders for deforming 3d mesh models. 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition pp. 5841–5850 (2017), https://api.semanticscholar.org/CorpusID:4379989
2017
Earlier work this paper cites.
Mescheder, L.M., Oechsle, M., Niemeyer, M., Nowozin, S., Geiger, A.: Occupancy networks: Learning 3d reconstruction in function space. 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 4455–4465 (2018), https://api.semanticscholar.org/CorpusID:54465161
2018
Earlier work this paper cites.
Zhang, R., Isola, P., Efros, A.A., Shechtman, E., Wang, O.: The unreasonable effectiveness of deep features as a perceptual metric. In: Proceedings of the IEEE conference on computer vision and pattern recognition. pp. 586–595 (2018)
2018
Earlier work this paper cites.
Nguyen-Phuoc, T., Li, C., Theis, L., Richardt, C., Yang, Y.L.: Hologan: Unsupervised learning of 3d representations from natural images. 2019 IEEE/CVF International Conference on Computer Vision Workshop (ICCVW) pp. 2037–2040 (2019), https://api.semanticscholar.org/CorpusID:91184364
2019
Earlier work this paper cites.
Wiles, O., Gkioxari, G., Szeliski, R., Johnson, J.: Synsin: End-to-end view synthesis from a single image. 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 7465–7475 (2019), https://api.semanticscholar.org/CorpusID:209405397
2019
Earlier work this paper cites.
Chan, E., Monteiro, M., Kellnhofer, P., Wu, J., Wetzstein, G.: pi-gan: Periodic implicit generative adversarial networks for 3d-aware image synthesis. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 5795–5805 (2020), https://api.semanticscholar.org/CorpusID:227247980
2020
Earlier work this paper cites.
Ho, J., Jain, A., Abbeel, P.: Denoising diffusion probabilistic models (2020)
2020
Earlier work this paper cites.
Mildenhall, B., Srinivasan, P.P., Tancik, M., Barron, J.T., Ramamoorthi, R., Ng, R.: Nerf: Representing scenes as neural radiance fields for view synthesis. In: Vedaldi, A., Bischof, H., Brox, T., Frahm, J. (eds.) Computer Vision - ECCV 2020 - 16th European Conference, Glasgow, UK, August 23-28, 2020, Proceedings, Part I. Lecture Notes in Computer Science, vol. 12346, pp. 405–421. Springer, New York, NY, USA (2020). https://doi.org/10.1007/978-3-030-58452-8_24, https://doi.org/10.1007/978-3-030-58452-8_24
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Niemeyer, M., Geiger, A.: Giraffe: Representing scenes as compositional generative neural feature fields. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 11448–11459 (2020), https://api.semanticscholar.org/CorpusID:227151657
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Teed, Z., Deng, J.: Raft: Recurrent all-pairs field transforms for optical flow. In: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part II 16. pp. 402–419. Springer (2020)
2020
Earlier work this paper cites.
Chan, E., Lin, C.Z., Chan, M., Nagano, K., Pan, B., Mello, S.D., Gallo, O., Guibas, L.J., Tremblay, J., Khamis, S., Karras, T., Wetzstein, G.: Efficient geometry-aware 3d generative adversarial networks. 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 16102–16112 (2021), https://api.semanticscholar.org/CorpusID:245144673
2021
Earlier work this paper cites.
Deng, Y., Yang, J., Xiang, J., Tong, X.: Gram: Generative radiance manifolds for 3d-aware image generation. 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 10663–10673 (2021), https://api.semanticscholar.org/CorpusID:245218753
2021
Earlier work this paper cites.
Esser, P., Rombach, R., Ommer, B.: Taming transformers for high-resolution image synthesis. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). pp. 12873–12883 (June 2021)
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Luo, S., Hu, W.: Diffusion probabilistic models for 3d point cloud generation. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 2836–2844 (2021), https://api.semanticscholar.org/CorpusID:232092778
2021
Earlier work this paper cites.
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., Ommer, B.: High-resolution image synthesis with latent diffusion models. 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 10674–10685 (2021), https://api.semanticscholar.org/CorpusID:245335280
2021
Earlier work this paper cites.
Shen, T., Gao, J., Yin, K., Liu, M.Y., Fidler, S.: Deep marching tetrahedra: a hybrid representation for high-resolution 3d shape synthesis. In: Neural Information Processing Systems (2021), https://api.semanticscholar.org/CorpusID:243848115
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Xu, Y., Peng, S., Yang, C., Shen, Y., Zhou, B.: 3d-aware image synthesis via learning structural and textural representations. 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 18409–18418 (2021), https://api.semanticscholar.org/CorpusID:245334914
2021
Earlier work this paper cites.
Zhou, L., Du, Y., Wu, J.: 3d shape generation and completion through point-voxel diffusion. 2021 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 5806–5815 (2021), https://api.semanticscholar.org/CorpusID:233182041
2021
Earlier work this paper cites.
Cheng, Y.C., Lee, H.Y., Tulyakov, S., Schwing, A.G., Gui, L.: Sdfusion: Multimodal 3d shape completion, reconstruction, and generation. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 4456–4465 (2022), https://api.semanticscholar.org/CorpusID:254408516
2022
Earlier work this paper cites.
Chou, G., Bahat, Y., Heide, F.: Diffusion-sdf: Conditional generative modeling of signed distance functions. 2023 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 2262–2272 (2022), https://api.semanticscholar.org/CorpusID:254017862
2022
Earlier work this paper cites.
Deitke, M., Schwenk, D., Salvador, J., Weihs, L., Michel, O., VanderBilt, E., Schmidt, L., Ehsani, K., Kembhavi, A., Farhadi, A.: Objaverse: A universe of annotated 3d objects. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 13142–13153 (2022), https://api.semanticscholar.org/CorpusID:254685588
2022
Earlier work this paper cites.
Downs, L., Francis, A., Koenig, N., Kinman, B., Hickman, R., Reymann, K., McHugh, T.B., Vanhoucke, V.: Google scanned objects: A high-quality dataset of 3d scanned household items. In: 2022 International Conference on Robotics and Automation (ICRA). pp. 2553–2560. IEEE (2022)
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Kalischek, N., Peters, T., Wegner, J.D., Schindler, K.: Tetradiffusion: Tetrahedral diffusion models for 3d shape generation (2022), https://api.semanticscholar.org/CorpusID:253802117
2022
Earlier work this paper cites.
Kulhánek, J., Derner, E., Sattler, T., Babuška, R.: Viewformer: Nerf-free neural rendering from few images using transformers (2022)
2022
Earlier work this paper cites.
Li, M., Duan, Y., Zhou, J., Lu, J.: Diffusion-sdf: Text-to-shape via voxelized diffusion. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 12642–12651 (2022), https://api.semanticscholar.org/CorpusID:254366593
2022
Earlier work this paper cites.
Muller, N., Siddiqui, Y., Porzi, L., Bulò, S.R., Kontschieder, P., Nießner, M.: Diffrf: Rendering-guided 3d radiance field diffusion. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 4328–4338 (2022), https://api.semanticscholar.org/CorpusID:254221225
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Poole, B., Jain, A., Barron, J.T., Mildenhall, B.: Dreamfusion: Text-to-3d using 2d diffusion (2022)
2022
Cited alongside, same era.
Liu, K., Zhan, F., Chen, Y., Zhang, J., Yu, Y., El Saddik, A., Lu, S., Xing, E.P.: Stylerf: Zero-shot 3d style transfer of neural radiance fields. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 8338–8348 (2023)
2023
Later among the works it cites.
Liu, R., Wu, R., Hoorick, B.V., Tokmakov, P., Zakharov, S., Vondrick, C.: Zero-1-to-3: Zero-shot one image to 3d object (2023)
2023
Later among the works it cites.
Liu, Y., Lin, C., Zeng, Z., Long, X., Liu, L., Komura, T., Wang, W.: Syncdreamer: Generating multiview-consistent images from a single-view image (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sajjadi, M.S., Meyer, H., Pot, E., Bergmann, U., Greff, K., Radwan, N., Vora, S., Lučić, M., Duckworth, D., Dosovitskiy, A., et al.: Scene representation transformer: Geometry-free novel view synthesis through set-latent scene representations. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 6229–6238 (2022)
2022
Cited alongside, same era.
Schuhmann, C., Beaumont, R., Vencu, R., Gordon, C., Wightman, R., Cherti, M., Coombes, T., Katta, A., Mullis, C., Wortsman, M., et al.: Laion-5b: An open large-scale dataset for training next generation image-text models. Advances in Neural Information Processing Systems 35
2022
Cited alongside, same era.
Shue, J., Chan, E., Po, R., Ankner, Z., Wu, J., Wetzstein, G.: 3d neural field generation using triplane diffusion. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 20875–20886 (2022), https://api.semanticscholar.org/CorpusID:254095843
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Watson, D., Chan, W., Martin-Brualla, R., Ho, J., Tagliasacchi, A., Norouzi, M.: Novel view synthesis with diffusion models (2022)
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Zhang, X., Zheng, Z., Gao, D., Zhang, B., Pan, P., Yang, Y.: Multi-view consistent generative adversarial networks for 3d-aware image synthesis. 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 18429–18438 (2022), https://api.semanticscholar.org/CorpusID:248157233
2022
Cited alongside, same era.
Zhou, Z., Tulsiani, S.: Sparsefusion: Distilling view-conditioned diffusion for 3d reconstruction. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 12588–12597 (2022), https://api.semanticscholar.org/CorpusID:254125457
2022
Cited alongside, same era.
2023
Later among the works it cites.
Melas-Kyriazi, L., Rupprecht, C., Laina, I., Vedaldi, A.: Realfusion: 360 reconstruction of any object from a single image (2023)
2023
Later among the works it cites.
Qian, G., Mai, J., Hamdi, A., Ren, J., Siarohin, A., Li, B., Lee, H.Y., Skorokhodov, I., Wonka, P., Tulyakov, S., Ghanem, B.: Magic123: One image to high-quality 3d object generation using both 2d and 3d diffusion priors (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Szymanowicz, S., Rupprecht, C., Vedaldi, A.: Viewset diffusion: (0-)image-conditioned 3d generative models from 2d data. 2023 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 8829–8839 (2023), https://api.semanticscholar.org/CorpusID:259144886
2023
Later among the works it cites.
2023
Later among the works it cites.
Tang, J., Ren, J., Zhou, H., Liu, Z., Zeng, G.: Dreamgaussian: Generative gaussian splatting for efficient 3d content creation (2023)
2023
Later among the works it cites.
Tang, J., Wang, T., Zhang, B., Zhang, T., Yi, R., Ma, L., Chen, D.: Make-it-3d: High-fidelity 3d creation from a single image with diffusion prior (2023)
2023
Later among the works it cites.
Tseng, H.Y., Li, Q., Kim, C., Alsisan, S., Huang, J.B., Kopf, J.: Consistent view synthesis with pose-guided diffusion models. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 16773–16783 (2023), https://api.semanticscholar.org/CorpusID:257833813
2023
Later among the works it cites.
2023
Later among the works it cites.
Wang, P., Fan, Z., Xu, D., Wang, D., Mohan, S., Iandola, F., Ranjan, R., Li, Y., Liu, Q., Wang, Z., Chandra, V.: Steindreamer: Variance reduction for text-to-3d score distillation via stein identity (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Wang, Y., Han, Q., Habermann, M., Daniilidis, K., Theobalt, C., Liu, L.: Neus2: Fast learning of neural implicit surfaces for multi-view reconstruction. In: Proceedings of the IEEE/CVF International Conference on Computer Vision. pp. 3295–3306 (2023)
2023
Later among the works it cites.
Wang, Z., Lu, C., Wang, Y., Bao, F., Li, C., Su, H., Zhu, J.: Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation (2023)
2023
Later among the works it cites.
Xu, Y., Tan, H., Luan, F., Bi, S., Wang, P., Li, J., Shi, Z., Sunkavalli, K., Wetzstein, G., Xu, Z., Zhang, K.: Dmv3d: Denoising multi-view diffusion using 3d large reconstruction model (2023)
2023
Later among the works it cites.
Yi, T., Fang, J., Wang, J., Wu, G., Xie, L., Zhang, X., Liu, W., Tian, Q., Wang, X.: Gaussiandreamer: Fast generation from text to 3d gaussians by bridging 2d and 3d diffusion models (2023)
2023
Later among the works it cites.
Yu, J.J., Forghani, F., Derpanis, K.G., Brubaker, M.A.: Long-term photometric consistent novel view synthesis with diffusion models. 2023 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 7071–7081 (2023), https://api.semanticscholar.org/CorpusID:258291651
2023
Later among the works it cites.
Yu, W., Yuan, L., Cao, Y.P., Gao, X., Li, X., Quan, L., Shan, Y., Tian, Y.: Hifi-123: Towards high-fidelity one image to 3d content generation (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Zheng, X., Pan, H., Wang, P.S., Tong, X., Liu, Y., yeung Shum, H.: Locally attentional sdf diffusion for controllable 3d shape generation. ACM Transactions on Graphics (TOG) 42
2023
Later among the works it cites.
Zhu, J., Zhuang, P.: Hifa: High-fidelity text-to-3d generation with advanced diffusion guidance (2023)
2023
Later among the works it cites.
Zuo, Q., Song, Y., Li, J., Liu, L., Bo, L.: Dg3d: Generating high quality 3d textured shapes by learning to discriminate multi-modal diffusion-renderings. 2023 IEEE/CVF International Conference on Computer Vision (ICCV) pp. 14529–14538 (2023), https://api.semanticscholar.org/CorpusID:266438917
2023
Later among the works it cites.
Deitke, M., Liu, R., Wallingford, M., Ngo, H., Michel, O., Kusupati, A., Fan, A., Laforte, C., Voleti, V., Gadre, S.Y., et al.: Objaverse-xl: A universe of 10m+ 3d objects. Advances in Neural Information Processing Systems 36
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.