Fetching the paper…
Reading the bibliography…
In the realm of digital creativity, our potential to craft intricate 3D worlds from imagination is often hampered by the limitations of existing digital tools, which demand extensive expertise and efforts.
ManifoldPlus: A Robust and Scalable Watertight Manifold Surface Generation Method for Triangle Soups
Jingwei Huang, Yichao Zhou, and Leonidas Guibas. 2020 · 2005
Earlier work this paper cites.
ShapeNet: An Information-Rich 3D Model Repository
Angel X. Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, Jianxiong Xiao, Li Yi, and Fisher Yu. 2015 · 2015
Earlier work this paper cites.
3d-r2n2: A unified approach for single and multi-view 3d object reconstruction. In Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11-14, 2016, Proceedings, Part VIII 14 . Springer, 628–644
Christopher B Choy, Danfei Xu, JunYoung Gwak, Kevin Chen, and Silvio Savarese. 2016 · 2016
Earlier work this paper cites.
A point set generation network for 3d object reconstruction from a single image. In Proceedings of the IEEE conference on computer vision and pattern recognition . 605–613
Haoqiang Fan, Hao Su, and Leonidas J Guibas. 2017 · 2017
Earlier work this paper cites.
PointNet++: deep hierarchical feature learning on point sets in a metric space
Charles R. Qi, Li Yi, Hao Su, and Leonidas J. Guibas. 2017 · 2017
Earlier work this paper cites.
Attention is All you Need. In Advances in Neural Information Processing Systems , I. Guyon, U. Von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett (Eds.), Vol. 30. Curran Associates, Inc
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
A papier-mâché approach to learning 3d surface generation. In Proceedings of the IEEE conference on computer vision and pattern recognition . 216–224
Thibault Groueix, Matthew Fisher, Vladimir G Kim, Bryan C Russell, and Mathieu Aubry. 2018 · 2018
Earlier work this paper cites.
QuadriFlow: A Scalable and Robust Method for Quadrangulation
Jingwei Huang, Yichao Zhou, Matthias Niessner, Jonathan Richard Shewchuk, and Leonidas J. Guibas. 2018b · 2018
Earlier work this paper cites.
Learning implicit fields for generative shape modeling. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 5939–5948
Zhiqin Chen and Hao Zhang. 2019 · 2019
Earlier work this paper cites.
Occupancy networks: Learning 3d reconstruction in function space. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 4460–4470
Lars Mescheder, Michael Oechsle, Michael Niemeyer, Sebastian Nowozin, and Andreas Geiger. 2019 · 2019
Earlier work this paper cites.
DeepSDF: Learning Continuous Signed Distance Functions for Shape Representation. In 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE Computer Society, Los Alamitos, CA, USA, 165–174
J. Park, P. Florence, J. Straub, R. Newcombe, and S. Lovegrove. 2019 · 2019
Earlier work this paper cites.
A skeleton-bridged deep learning approach for generating meshes of complex topologies from single rgb images. In Proceedings of the ieee/cvf conference on computer vision and pattern recognition . 4541–4550
Jiapeng Tang, Xiaoguang Han, Junyi Pan, Kui Jia, and Xin Tong. 2019 · 2019
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models. In Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M.F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 6840–6851
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
Polygen: An autoregressive generative model of 3d meshes. In International conference on machine learning . PMLR, 7220–7229
Charlie Nash, Yaroslav Ganin, SM Ali Eslami, and Peter Battaglia. 2020 · 2020
Earlier work this paper cites.
Convolutional occupancy networks. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part III 16 . Springer, 523–540
Songyou Peng, Michael Niemeyer, Lars Mescheder, Marc Pollefeys, and Andreas Geiger. 2020 · 2020
Earlier work this paper cites.
On layer normalization in the transformer architecture. In International Conference on Machine Learning . PMLR, 10524–10533
Ruibin Xiong, Yunchang Yang, Di He, Kai Zheng, Shuxin Zheng, Chen Xing, Huishuai Zhang, Yanyan Lan, Liwei Wang, and Tieyan Liu. 2020 · 2020
Earlier work this paper cites.
mesh_to_sdf: Calculate signed distance fields for arbitrary meshes
Kleineberg Marian. 2021 · 2021
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T Barron, Ravi Ramamoorthi, and Ren Ng. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision. In International conference on machine learning . PMLR, 8748–8763
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Earlier work this paper cites.
Zero-Shot Text-to-Image Generation. In Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event (Proceedings of Machine Learning Research, Vol. 139) , Marina Meila and Tong Zhang (Eds.). PMLR, 8821–8831
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021 · 2021
Earlier work this paper cites.
Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category Reconstruction. In 2021 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE Computer Society, Los Alamitos, CA, USA, 10881–10891
J. Reizenstein, R. Shapovalov, P. Henzler, L. Sbordone, P. Labatut, and D. Novotny. 2021 · 2021
Earlier work this paper cites.
Deep Marching Tetrahedra: a Hybrid Representation for High-Resolution 3D Shape Synthesis. In Advances in Neural Information Processing Systems , A. Beygelzimer, Y. Dauphin, P. Liang, and J. Wortman Vaughan (Eds.)
Tianchang Shen, Jun Gao, Kangxue Yin, Ming-Yu Liu, and Sanja Fidler. 2021 · 2021
Earlier work this paper cites.
Skeletonnet: A topology-preserving solution for learning mesh reconstruction of object surfaces from rgb images
Jiapeng Tang, Xiaoguang Han, Mingkui Tan, Xin Tong, and Kui Jia. 2021a · 2021
Earlier work this paper cites.
NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view Reconstruction
Peng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt, Taku Komura, and Wenping Wang. 2021a · 2021
Earlier work this paper cites.
Real-ESRGAN: Training Real-World Blind Super-Resolution with Pure Synthetic Data. In 2021 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW) . IEEE Computer Society, Los Alamitos, CA, USA, 1905–1914
X. Wang, L. Xie, C. Dong, and Y. Shan. 2021b · 2021
Earlier work this paper cites.
LoRA: Low-Rank Adaptation of Large Language Models. In International Conference on Learning Representations
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Earlier work this paper cites.
SparseNeuS: Fast Generalizable Neural Surface Reconstruction from Sparse Views. In Computer Vision – ECCV 2022 , Shai Avidan, Gabriel Brostow, Moustapha Cissé, Giovanni Maria Farinella, and Tal Hassner (Eds.). Springer Nature Switzerland, Cham, 210–227
Xiaoxiao Long, Cheng Lin, Peng Wang, Taku Komura, and Wenping Wang. 2022 · 2022
Earlier work this paper cites.
Point-E: A System for Generating 3D Point Clouds from Complex Prompts
Alex Nichol, Heewoo Jun, Prafulla Dhariwal, Pamela Mishkin, and Mark Chen. 2022 · 2022
Earlier work this paper cites.
High-Resolution Image Synthesis with Latent Diffusion Models. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE Computer Society, Los Alamitos, CA, USA, 10674–10685
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer. 2022 · 2022
Earlier work this paper cites.
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding. In Advances in Neural Information Processing Systems , S. Koyejo, S. Mohamed, A. Agarwal, D. Belgrave, K. Cho, and A. Oh (Eds.), Vol. 35. Curran Associates, Inc., 36479–36494
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily L Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, Jonathan Ho, David J Fleet, and Mohammad Norouzi. 2022 · 2022
Earlier work this paper cites.
mesh2sdf
Peng-Shuai Wang. 2022 · 2022
Cited alongside, same era.
Dual octree graph networks for learning adaptive volumetric shape representations
Peng-Shuai Wang, Yang Liu, and Xin Tong. 2022 · 2022
Cited alongside, same era.
MultiDiffusion: fusing diffusion paths for controlled image generation
Omer Bar-Tal, Lior Yariv, Yaron Lipman, and Tali Dekel. 2023 · 2023
Cited alongside, same era.
Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Andreas Blattmann, Tim Dockhorn, Sumith Kulal, Daniel Mendelevitch, Maciej Kilian, Dominik Lorenz, Yam Levi, Zion English, Vikram Voleti, Adam Letts, Varun Jampani, and Robin Rombach. 2023 · 2023
Cited alongside, same era.
Text2Tex: Text-driven Texture Synthesis via Diffusion Models. In 2023 IEEE/CVF International Conference on Computer Vision (ICCV) . 18512–18522
Dave Zhenyu Chen, Yawar Siddiqui, Hsin-Ying Lee, Sergey Tulyakov, and Matthias Nießner. 2023b · 2023
Cited alongside, same era.
ShapeGPT: 3D Shape Generation with A Unified Multi-modal Language Model
Fukun Yin, Xin Chen, Chi Zhang, Biao Jiang, Zibo Zhao, Jiayuan Fan, Gang Yu, Taihao Li, and Tao Chen. 2023 · 2023
Later among the works it cites.
MVImgNet: A Large-scale Dataset of Multi-view Images. In 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE Computer Society, Los Alamitos, CA, USA, 9150–9161
X. Yu, M. Xu, Y. Zhang, H. Liu, C. Ye, Y. Wu, Z. Yan, C. Zhu, Z. Xiong, T. Liang, G. Chen, S. Cui, and X. Han. 2023a · 2023
Later among the works it cites.
3DShape2VecSet: A 3D Shape Representation for Neural Fields and Generative Diffusion Models
Biao Zhang, Jiapeng Tang, Matthias Nießner, and Peter Wonka. 2023c · 2023
Later among the works it cites.
DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance
Longwen Zhang, Qiwei Qiu, Hongyang Lin, Qixuan Zhang, Cheng Shi, Wei Yang, Ye Shi, Sibei Yang, Lan Xu, and Jingyi Yu. 2023a · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fantasia3D: Disentangling Geometry and Appearance for High-quality Text-to-3D Content Creation. In 2023 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE Computer Society, Los Alamitos, CA, USA, 22189–22199
R. Chen, Y. Chen, N. Jiao, and K. Jia. 2023a · 2023
Cited alongside, same era.
SDFusion: Multimodal 3D Shape Completion, Reconstruction, and Generation. In 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE Computer Society, Los Alamitos, CA, USA, 4456–4465
Y. Cheng, H. Lee, S. Tulyakov, A. Schwing, and L. Gui. 2023 · 2023
Cited alongside, same era.
Objaverse: A Universe of Annotated 3D Objects. In 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE Computer Society, Los Alamitos, CA, USA, 13142–13153
M. Deitke, D. Schwenk, J. Salvador, L. Weihs, O. Michel, E. VanderBilt, L. Schmidt, K. Ehsanit, A. Kembhavi, and A. Farhadi. 2023 · 2023
Cited alongside, same era.
Composable Function-preserving Expansions for Transformer Architectures
Andrea Gesmundo and Kaitlin Maile. 2023 · 2023
Cited alongside, same era.
threestudio: A unified framework for 3D content generation
Yuan-Chen Guo, Ying-Tian Liu, Ruizhi Shao, Christian Laforte, Vikram Voleti, Guan Luo, Chia-Hao Chen, Zi-Xin Zou, Chen Wang, Yan-Pei Cao, and Song-Hai Zhang. 2023 · 2023
Cited alongside, same era.
3DGen: Triplane Latent Diffusion for Textured Mesh Generation
Anchit Gupta, Wenhan Xiong, Yixin Nie, Ian Jones, and Barlas Oğuz. 2023 · 2023
Cited alongside, same era.
Shap-E: Generating Conditional 3D Implicit Functions
Heewoo Jun and Alex Nichol. 2023 · 2023
Cited alongside, same era.
Adding Conditional Control to Text-to-Image Diffusion Models. In 2023 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE Computer Society, Los Alamitos, CA, USA, 3813–3824
L. Zhang, A. Rao, and M. Agrawala. 2023b · 2023
Later among the works it cites.
Michelangelo: Conditional 3d shape generation based on shape-image-text aligned latent representation
Zibo Zhao, Wen Liu, Xin Chen, Xianfang Zeng, Rui Wang, Pei Cheng, Bin Fu, Tao Chen, Gang Yu, and Shenghua Gao. 2023 · 2023
Later among the works it cites.
Locally Attentional SDF Diffusion for Controllable 3D Shape Generation
Xin-Yang Zheng, Hao Pan, Peng-Shuai Wang, Xin Tong, Yang Liu, and Heung-Yeung Shum. 2023 · 2023
Later among the works it cites.
Blender - a 3D modelling and rendering package
Blender Online Community. 2024 · 2024
Closest in time.
Text-to-3D using Gaussian Splatting. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Zilong Chen, Feng Wang, and Huaping Liu. 2024 · 2024
Closest in time.
LRM: Large Reconstruction Model for Single Image to 3D. In The Twelfth International Conference on Learning Representations
Yicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi, Yang Zhou, Difan Liu, Feng Liu, Kalyan Sunkavalli, Trung Bui, and Hao Tan. 2024 · 2024
Closest in time.
DreamTime: An Improved Optimization Strategy for Diffusion-Guided 3D Generation. In The Twelfth International Conference on Learning Representations
Yukun Huang, Jianan Wang, Yukai Shi, Boshi Tang, Xianbiao Qi, and Lei Zhang. 2024 · 2024
Closest in time.
SweetDreamer: Aligning Geometric Priors in 2D diffusion for Consistent Text-to-3D. In The Twelfth International Conference on Learning Representations
Weiyu Li, Rui Chen, Xuelin Chen, and Ping Tan. 2024 · 2024
Closest in time.
Common diffusion noise schedules and sample steps are flawed. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision . 5404–5411
Shanchuan Lin, Bingchen Liu, Jiashi Li, and Xiao Yang. 2024 · 2024
Closest in time.
Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Huan Ling, Seung Wook Kim, Antonio Torralba, Sanja Fidler, and Karsten Kreis. 2024 · 2024
Closest in time.
Wonder3D: Single Image to 3D using Cross-Domain Diffusion. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin, Yuan Liu, Zhiyang Dou, Lingjie Liu, Yuexin Ma, Song-Hai Zhang, Marc Habermann, Christian Theobalt, and Wenping Wang. 2024 · 2024
Closest in time.
DINOv2: Learning Robust Visual Features without Supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy V. Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel HAZIZA, Francisco Massa, Alaaeldin El-Nouby, Mido Assran, Nicolas Ballas, Wojciech Galuba, Russell Howes, Po-Yao Huang, Shang-Wen Li, Ishan Misra, Michael Rabbat, Vasu Sharma, Gabriel Synnaeve, Hu Xu, Herve Jegou, Julien Mairal, Patrick Labatut, Armand Joulin, and Piotr Bojanowski. 2024 · 2024
Closest in time.
Magic123: One Image to High-Quality 3D Object Generation Using Both 2D and 3D Diffusion Priors. In The Twelfth International Conference on Learning Representations
Guocheng Qian, Jinjie Mai, Abdullah Hamdi, Jian Ren, Aliaksandr Siarohin, Bing Li, Hsin-Ying Lee, Ivan Skorokhodov, Peter Wonka, Sergey Tulyakov, and Bernard Ghanem. 2024 · 2024
Closest in time.
RichDreamer: A Generalizable Normal-Depth Diffusion Model for Detail Richness in Text-to-3D. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Lingteng Qiu, Guanying Chen, Xiaodong Gu, Qi Zuo, Mutian Xu, Yushuang Wu, Weihao Yuan, Zilong Dong, Liefeng Bo, and Xiaoguang Han. 2024 · 2024
Closest in time.
XCube ( 𝒳 3 \mathcal{X}^{3} ): Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Xuanchi Ren, Jiahui Huang, Xiaohui Zeng, Ken Museth, Sanja Fidler, and Francis Williams. 2024 · 2024
Closest in time.
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation. In The Twelfth International Conference on Learning Representations
Junyoung Seo, Wooseok Jang, Min-Seop Kwak, Hyeonsu Kim, Jaehoon Ko, Junho Kim, Jin-Hwa Kim, Jiyoung Lee, and Seungryong Kim. 2024 · 2024
Closest in time.
MVDream: Multi-view Diffusion for 3D Generation. In The Twelfth International Conference on Learning Representations
Yichun Shi, Peng Wang, Jianglong Ye, Long Mai, Kejie Li, and Xiao Yang. 2024 · 2024
Closest in time.
MeshGPT: Generating Triangle Meshes with Decoder-Only Transformers. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Yawar Siddiqui, Antonio Alliegro, Alexey Artemov, Tatiana Tommasi, Daniele Sirigatti, Vladislav Rosov, Angela Dai, and Matthias Nießner. 2024 · 2024
Closest in time.
DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior. In The Twelfth International Conference on Learning Representations
Jingxiang Sun, Bo Zhang, Ruizhi Shao, Lizhen Wang, Wen Liu, Zhenda Xie, and Yebin Liu. 2024 · 2024
Closest in time.
DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation. In The Twelfth International Conference on Learning Representations
Jiaxiang Tang, Jiawei Ren, Hang Zhou, Ziwei Liu, and Gang Zeng. 2024 · 2024
Closest in time.
PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction. In The Twelfth International Conference on Learning Representations
Peng Wang, Hao Tan, Sai Bi, Yinghao Xu, Fujun Luan, Kalyan Sunkavalli, Wenping Wang, Zexiang Xu, and Kai Zhang. 2024 · 2024
Closest in time.
Hd-fusion: Detailed text-to-3d generation leveraging multiple noise estimation. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision . 3202–3211
Jinbo Wu, Xiaobo Gao, Xing Liu, Zhengyang Shen, Chen Zhao, Haocheng Feng, Jingtuo Liu, and Errui Ding. 2024 · 2024
Closest in time.
DMV3D: Denoising Multi-view Diffusion Using 3D Large Reconstruction Model. In The Twelfth International Conference on Learning Representations
Yinghao Xu, Hao Tan, Fujun Luan, Sai Bi, Peng Wang, Jiahao Li, Zifan Shi, Kalyan Sunkavalli, Gordon Wetzstein, Zexiang Xu, and Kai Zhang. 2024 · 2024
Closest in time.
Mosaic-SDF for 3D Generative Models. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Lior Yariv, Omri Puny, Natalia Neverova, Oran Gafni, and Yaron Lipman. 2024 · 2024
Closest in time.
HIFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance. In The Twelfth International Conference on Learning Representations
Junzhe Zhu, Peiye Zhuang, and Sanmi Koyejo. 2024 · 2024
Closest in time.
Triplane Meets Gaussian Splatting: Fast and Generalizable Single-View 3D Reconstruction with Transformers. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Zi-Xin Zou, Zhipeng Yu, Yuan-Chen Guo, Yangguang Li, Ding Liang, Yan-Pei Cao, and Song-Hai Zhang. 2024 · 2024
Closest in time.