Fetching the paper…
Reading the bibliography…
Generative AI has made significant progress in recent years, with text-guided content generation being the most practical as it facilitates interaction between human instructions and AI-generated content (AIGC).
L. W. Peng and S. M. Shamsuddin, “3d object reconstruction and representation using neural networks,” in Proceedings of the 2nd international conference on Computer graphics and interactive techniques in Australasia and South East Asia , 2004, pp. 139–147
2004
Earlier work this paper cites.
J. F. Blinn, “What is a pixel?” IEEE computer graphics and applications , vol. 25, no. 5, pp. 82–87, 2005
2005
Earlier work this paper cites.
B. Cyganek and J. P. Siebert, An introduction to 3D computer vision techniques and algorithms . John Wiley & Sons, 2011
2011
Earlier work this paper cites.
Y. Furukawa, C. Hernández et al. , “Multi-view stereo: A tutorial,” Foundations and Trends® in Computer Graphics and Vision , vol. 9, no. 1-2, pp. 1–148, 2015
2015
Earlier work this paper cites.
F. Bogo, A. Kanazawa, C. Lassner, P. Gehler, J. Romero, and M. J. Black, “Keep it smpl: Automatic estimation of 3d human pose and shape from a single image,” in Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11-14, 2016, Proceedings, Part V 14 . Springer, 2016, pp. 561–578
2016
Earlier work this paper cites.
C. R. Qi, H. Su, K. Mo, and L. J. Guibas, “Pointnet: Deep learning on point sets for 3d classification and segmentation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 652–660
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
L. Minto, P. Zanuttigh, and G. Pagnutti, “Deep learning for 3d shape classification based on volumetric density and surface approximation clues.” in VISIGRAPP (5: VISAPP) , 2018, pp. 317–324
2018
Earlier work this paper cites.
B. Zhao, X. Wu, Z.-Q. Cheng, H. Liu, Z. Jie, and J. Feng, “Multi-view image generation from a single-view,” in Proceedings of the 26th ACM international conference on Multimedia , 2018, pp. 383–391
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
K. Chen, C. B. Choy, M. Savva, A. X. Chang, T. Funkhouser, and S. Savarese, “Text2shape: Generating shapes from natural language by learning joint embeddings,” in Computer Vision–ACCV 2018: 14th Asian Conference on Computer Vision, Perth, Australia, December 2–6, 2018, Revised Selected Papers, Part III 14 . Springer, 2019, pp. 100–116
2019
Earlier work this paper cites.
P. Achlioptas, J. Fan, R. Hawkins, N. Goodman, and L. J. Guibas, “Shapeglot: Learning language for shape differentiation,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 8938–8947
2019
Earlier work this paper cites.
J. J. Park, P. Florence, J. Straub, R. Newcombe, and S. Lovegrove, “Deepsdf: Learning continuous signed distance functions for shape representation,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 165–174
2019
Earlier work this paper cites.
R. Hanocka, A. Hertz, N. Fish, R. Giryes, S. Fleishman, and D. Cohen-Or, “Meshcnn: a network with an edge,” ACM Transactions on Graphics (TOG) , vol. 38, no. 4, pp. 1–12, 2019
2019
Earlier work this paper cites.
L. Mescheder, M. Oechsle, M. Niemeyer, S. Nowozin, and A. Geiger, “Occupancy networks: Learning 3d reconstruction in function space,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 4460–4470
2019
Earlier work this paper cites.
C. Wang, M. Cheng, F. Sohel, M. Bennamoun, and J. Li, “Normalnet: A voxel-based cnn for 3d object classification and retrieval,” Neurocomputing , vol. 323, pp. 139–147, 2019
2019
Earlier work this paper cites.
W. Liu, J. Sun, W. Li, T. Hu, and P. Wang, “Deep learning on point clouds and its application: A survey,” Sensors , vol. 19, no. 19, p. 4188, 2019
2019
Earlier work this paper cites.
G. Pavlakos, V. Choutas, N. Ghorbani, T. Bolkart, A. A. Osman, D. Tzionas, and M. J. Black, “Expressive body capture: 3d hands, face, and body from a single image,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 10 975–10 985
2019
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” Advances in Neural Information Processing Systems , vol. 33, pp. 6840–6851, 2020
2020
Earlier work this paper cites.
Y. Guo, H. Wang, Q. Hu, H. Liu, L. Liu, and M. Bennamoun, “Deep learning for 3d point clouds: A survey,” IEEE transactions on pattern analysis and machine intelligence , vol. 43, no. 12, pp. 4338–4364, 2020
2020
Earlier work this paper cites.
Y. Jin, D. Jiang, and M. Cai, “3d reconstruction using deep learning: a survey,” Communications in Information and Systems , vol. 20, no. 4, pp. 389–413, 2020
2020
Earlier work this paper cites.
W. Cao, Z. Yan, Z. He, and Z. He, “A comprehensive survey on geometric deep learning,” IEEE Access , vol. 8, pp. 35 929–35 949, 2020
2020
Earlier work this paper cites.
K. Rematas and V. Ferrari, “Neural voxel renderer: Learning an accurate and controllable rendering tool,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 5417–5427
2020
Earlier work this paper cites.
Z. Wu, S. Pan, F. Chen, G. Long, C. Zhang, and S. Y. Philip, “A comprehensive survey on graph neural networks,” IEEE transactions on neural networks and learning systems , vol. 32, no. 1, pp. 4–24, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick, “Momentum contrast for unsupervised visual representation learning,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 9729–9738
2020
Earlier work this paper cites.
R. Ranftl, K. Lasinger, D. Hafner, K. Schindler, and V. Koltun, “Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer,” IEEE transactions on pattern analysis and machine intelligence , vol. 44, no. 3, pp. 1623–1637, 2020
2020
Earlier work this paper cites.
B. Kye, N. Han, E. Kim, Y. Park, and S. Jo, “Educational applications of metaverse: possibilities and limitations,” Journal of educational evaluation for health professions , vol. 18, 2021
2021
Earlier work this paper cites.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “Nerf: Representing scenes as neural radiance fields for view synthesis,” Communications of the ACM , vol. 65, no. 1, pp. 99–106, 2021
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
H. Zhou, W. Zhang, K. Chen, W. Li, and N. Yu, “Three-dimensional mesh steganography and steganalysis: a review,” IEEE Transactions on Visualization and Computer Graphics , 2021
2021
Earlier work this paper cites.
T. Shen, J. Gao, K. Yin, M.-Y. Liu, and S. Fidler, “Deep marching tetrahedra: a hybrid representation for high-resolution 3d shape synthesis,” Advances in Neural Information Processing Systems , vol. 34, pp. 6087–6101, 2021
2021
Earlier work this paper cites.
P. Dhariwal and A. Nichol, “Diffusion models beat gans on image synthesis,” Advances in neural information processing systems , vol. 34, pp. 8780–8794, 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
R. Ranftl, A. Bochkovskiy, and V. Koltun, “Vision transformers for dense prediction,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 12 179–12 188
2021
Earlier work this paper cites.
M. Caron, H. Touvron, I. Misra, H. Jégou, J. Mairal, P. Bojanowski, and A. Joulin, “Emerging properties in self-supervised vision transformers,” in Proceedings of the IEEE/CVF international conference on computer vision , 2021, pp. 9650–9660
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
Q. Yang, Y. Zhao, H. Huang, Z. Xiong, J. Kang, and Z. Zheng, “Fusing blockchain and ai with metaverse: A survey,” IEEE Open Journal of the Computer Society , vol. 3, pp. 122–136, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
H. Jeon, H. Youn, S. Ko, and T. Kim, “Blockchain and ai meet in the metaverse,” Advances in the Convergence of Blockchain and Artificial Intelligence , vol. 73, no. 10.5772, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
A. Jain, B. Mildenhall, J. T. Barron, P. Abbeel, and B. Poole, “Zero-shot text-guided object generation with dream fields,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 867–876
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
C. Wang, M. Chai, M. He, D. Chen, and J. Liao, “Clip-nerf: Text-and-image driven manipulation of neural radiance fields,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 3835–3844
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
H. Wang and J. Zhang, “A survey of deep learning-based mesh processing,” Communications in Mathematics and Statistics , vol. 10, no. 1, pp. 163–194, 2022
2022
Earlier work this paper cites.
T. Müller, A. Evans, C. Schied, and A. Keller, “Instant neural graphics primitives with a multiresolution hash encoding,” ACM transactions on graphics (TOG) , vol. 41, no. 4, pp. 1–15, 2022
2022
Earlier work this paper cites.
E. R. Chan, C. Z. Lin, M. A. Chan, K. Nagano, B. Pan, S. De Mello, O. Gallo, L. J. Guibas, J. Tremblay, S. Khamis et al. , “Efficient geometry-aware 3d generative adversarial networks,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, pp. 16 123–16 133
2022
Earlier work this paper cites.
J. Gao, T. Shen, Z. Wang, W. Chen, K. Yin, D. Li, O. Litany, Z. Gojcic, and S. Fidler, “Get3d: A generative model of high quality 3d textured shapes learned from images,” Advances In Neural Information Processing Systems , vol. 35, pp. 31 841–31 854, 2022
2022
Earlier work this paper cites.
B. Zhang, M. Nießner, and P. Wonka, “3dilg: Irregular latent grids for 3d generative modeling,” Advances in Neural Information Processing Systems , vol. 35, pp. 21 871–21 885, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
M. A. Bautista, P. Guo, S. Abnar, W. Talbott, A. Toshev, Z. Chen, L. Dinh, S. Zhai, H. Goh, D. Ulbricht et al. , “Gaudi: A neural architect for immersive 3d scene generation,” Advances in Neural Information Processing Systems , vol. 35, pp. 25 102–25 116, 2022
2022
Earlier work this paper cites.
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans et al. , “Photorealistic text-to-image diffusion models with deep language understanding,” Advances in Neural Information Processing Systems , vol. 35, pp. 36 479–36 494, 2022
2022
Earlier work this paper cites.
N. Mohammad Khalid, T. Xie, E. Belilovsky, and T. Popa, “Clip-mesh: Generating textured meshes from text using pretrained image-text models,” in SIGGRAPH Asia 2022 Conference Papers , 2022, pp. 1–8
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
O. Michel, R. Bar-On, R. Liu, S. Benaim, and R. Hanocka, “Text2mesh: Text-driven neural stylization for meshes,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 13 492–13 502
2022
Earlier work this paper cites.
G. Tevet, B. Gordon, A. Hertz, A. H. Bermano, and D. Cohen-Or, “Motionclip: Exposing human motion generation to clip space,” in Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XXII . Springer, 2022, pp. 358–374
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 10 684–10 695
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
H. Wang, H. Ning, Y. Lin, W. Wang, S. Dhelim, F. Farha, J. Ding, and M. Daneshmand, “A survey on the metaverse: The state-of-the-art, technologies, applications, and challenges,” IEEE Internet of Things Journal , vol. 10, no. 16, pp. 14 671–14 688, 2023
2023
Cited alongside, same era.
Z. Lv, “Generative artificial intelligence in the metaverse era,” Cognitive Robotics , vol. 3, pp. 208–217, 2023
2023
Cited alongside, same era.
W. Xia and J.-H. Xue, “A survey on deep generative 3d-aware image synthesis,” ACM Computing Surveys , vol. 56, no. 4, pp. 1–34, 2023
2023
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Wu, C. Wen, S. Shi, X. Li, and C. Wang, “Virtual sparse convolution for multimodal 3d object detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 21 653–21 662
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
B. Kerbl, G. Kopanas, T. Leimkühler, and G. Drettakis, “3d gaussian splatting for real-time radiance field rendering,” 2023
2023
Cited alongside, same era.
J. R. Shue, E. R. Chan, R. Po, Z. Ankner, J. Wu, and G. Wetzstein, “3d neural field generation using triplane diffusion,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 20 875–20 886
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
T. Cao, K. Kreis, S. Fidler, N. Sharp, and K. Yin, “Texfusion: Synthesizing 3d textures with text-guided image diffusion models,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 4169–4181
2023
Closest in time.
2023
Closest in time.
S. Aneja, J. Thies, A. Dai, and M. Nießner, “Clipface: Text-guided editing of textured 3d morphable models,” in ACM SIGGRAPH 2023 Conference Proceedings , 2023, pp. 1–11
2023
Closest in time.
2023
Closest in time.
J. Zhuang, C. Wang, L. Lin, L. Liu, and G. Li, “Dreameditor: Text-driven 3d scene editing with neural fields,” in SIGGRAPH Asia 2023 Conference Papers , 2023, pp. 1–10
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
E. Sella, G. Fiebelman, P. Hedman, and H. Averbuch-Elor, “Vox-e: Text-guided voxel editing of 3d objects,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 430–440
2023
Closest in time.
2023
Closest in time.
H. Zohny, J. McMillan, and M. King, “Ethics of generative ai,” pp. 79–80, 2023
2023
Closest in time.
K. Wach, C. D. Duong, J. Ejdys, R. Kazlauskaitė, P. Korzynski, G. Mazurek, J. Paliszkiewicz, and E. Ziemba, “The dark side of generative artificial intelligence: A critical analysis of controversies and risks of chatgpt,” Entrepreneurial Business and Economics Review , vol. 11, no. 2, pp. 7–30, 2023
2023
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
H. Lee, M. Savva, and A. X. Chang, “Text-to-3d shape generation,” in Computer Graphics Forum . Wiley Online Library, 2024, p. e15061
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
J. Wu, X. Gao, X. Liu, Z. Shen, C. Zhao, H. Feng, J. Liu, and E. Ding, “Hd-fusion: Detailed text-to-3d generation leveraging multiple noise estimation,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2024, pp. 3202–3211
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Y. Chen, C. Zhang, X. Yang, Z. Cai, G. Yu, L. Yang, and G. Lin, “It3d: Improved text-to-3d generation with explicit view synthesis,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 2, 2024, pp. 1237–1244
2024
Closest in time.
Z. Wang, C. Lu, Y. Wang, F. Bao, C. Li, H. Su, and J. Zhu, “Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
H. Zhang, B. Chen, H. Yang, L. Qu, X. Wang, L. Chen, C. Long, F. Zhu, D. Du, and M. Zheng, “Avatarverse: High-quality & stable 3d avatar creation from text and pose,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 7, 2024, pp. 7124–7132
2024
Closest in time.
N. Kolotouros, T. Alldieck, A. Zanfir, E. Bazavan, M. Fieraru, and C. Sminchisescu, “Dreamhuman: Animatable 3d avatars from text,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Y. Huang, J. Wang, A. Zeng, H. Cao, X. Qi, Y. Shi, Z.-J. Zha, and L. Zhang, “Dreamwaltz: Make a scene with complex 3d animatable avatars,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
X. Han, Y. Cao, K. Han, X. Zhu, J. Deng, Y.-Z. Song, T. Xiang, and K.-Y. K. Wong, “Headsculpt: Crafting 3d head avatars with text,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
S. Bahmani, I. Skorokhodov, V. Rong, G. Wetzstein, L. Guibas, P. Wonka, S. Tulyakov, J. J. Park, A. Tagliasacchi, and D. B. Lindell, “4d-fy: Text-to-4d generation using hybrid score distillation sampling,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 7996–8006
2024
Closest in time.
H. Xu, Y. Wu, X. Tang, J. Zhang, Y. Zhang, Z. Zhang, C. Li, and X. Jin, “Fusiondeformer: text-guided mesh deformation using diffusion models,” The Visual Computer , pp. 1–12, 2024
2024
Closest in time.
Y. Li, Y. Dou, Y. Shi, Y. Lei, X. Chen, Y. Zhang, P. Zhou, and B. Ni, “Focaldreamer: Text-driven 3d editing via focal-fusion assembly,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 4, 2024, pp. 3279–3287
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
T. Wu, G. Yang, Z. Li, K. Zhang, Z. Liu, L. Guibas, D. Lin, and G. Wetzstein, “Gpt-4v (ision) is a human-aligned evaluator for text-to-3d generation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 22 227–22 238
2024
Closest in time.
2024
Closest in time.
Z. Qi, M. Yu, R. Dong, and K. Ma, “Vpp: Efficient conditional 3d generation via voxel-point progressive representation,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.