Fetching the paper…
Reading the bibliography…
Generating high-quality 3D objects from textual descriptions remains a challenging problem due to computational cost, the scarcity of 3D data, and complex 3D representations.
Marching cubes: A high resolution 3d surface construction algorithm
William E. Lorensen and Harvey E. Cline · 1987
Earlier work this paper cites.
An efficient method of triangulating equi-valued surfaces by using tetrahedral cells
Akio Doi and Akio Koide · 1991
Earlier work this paper cites.
Shape modeling with front propagation: A level set approach
Ravi Malladi, James A. Sethian, and Baba C. Vemuri · 1995
Earlier work this paper cites.
Geometry images
Xianfeng Gu, Steven J. Gortler, and Hugues Hoppe · 2002
Earlier work this paper cites.
Multi-chart geometry images
Pedro V. Sander, Zoë J. Wood, Steven J. Gortler, John M. Snyder, and Hugues Hoppe · 2003
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric A Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
Deep learning 3d shape surfaces using geometry images
Ayan Sinha, Jing Bai, and Karthik Ramani · 2016
Earlier work this paper cites.
Boundary first flattening
Rohan Sawhney and Keenan Crane · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam M. Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Learning implicit fields for generative shape modeling
Zhiqin Chen and Hao Zhang · 2018
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Tero Karras, Samuli Laine, and Timo Aila · 2018
Earlier work this paper cites.
Occupancy networks: Learning 3d reconstruction in function space
Lars M. Mescheder, Michael Oechsle, Michael Niemeyer, Sebastian Nowozin, and Andreas Geiger · 2018
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Narain Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2020
Earlier work this paper cites.
Efficient geometry-aware 3d generative adversarial networks
Eric Chan, Connor Z. Lin, Matthew Chan, Koki Nagano, Boxiao Pan, Shalini De Mello, Orazio Gallo, Leonidas J. Guibas, Jonathan Tremblay, S. Khamis, Tero Karras, and Gordon Wetzstein · 2021
Earlier work this paper cites.
Abo: Dataset and benchmarks for real-world 3d object understanding
Jasmine Collins, Shubham Goel, Achleshwar Luthra, Leon L. Xu, Kenan Deng, Xi Zhang, T. F. Y. Vicente, Himanshu Arora, T. L. Dideriksen, Matthieu Guillaumin, and Jitendra Malik · 2021
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, A. Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2021
Earlier work this paper cites.
Maurice Weiler, Patrick Forr’e, Erik P. Verlinde, and Max Welling · 2021
Earlier work this paper cites.
Neural fields in visual computing and beyond
Yiheng Xie, Towaki Takikawa, Shunsuke Saito, Or Litany, Shiqin Yan, Numair Khan, Federico Tombari, James Tompkin, Vincent Sitzmann, and Srinath Sridhar · 2021
Earlier work this paper cites.
Xdgan: Multi-modal 3d shape generation in 2d space
Hassan Abu Alhaija, Alara Dirik, Andr’e Knorig, Sanja Fidler, and Maria Shugrina · 2022
Earlier work this paper cites.
Objaverse: A universe of annotated 3d objects
Matt Deitke, Dustin Schwenk, Jordi Salvador, Luca Weihs, Oscar Michel, Eli VanderBilt, Ludwig Schmidt, Kiana Ehsani, Aniruddha Kembhavi, and Ali Farhadi · 2022
Earlier work this paper cites.
Flow matching for generative modeling
Yaron Lipman, Ricky T Q Chen, Heli Ben-Hamu, Maximilian Nickel, and Matt Le · 2022
Earlier work this paper cites.
Point-e: A system for generating 3d point clouds from complex prompts
Alex Nichol, Heewoo Jun, Prafulla Dhariwal, Pamela Mishkin, and Mark Chen · 2022
Earlier work this paper cites.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T. Barron, and Ben Mildenhall · 2022
Cited alongside, same era.
Laion-5b: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, Patrick Schramowski, Srivatsa Kundurthy, Katherine Crowson, Ludwig Schmidt, Robert Kaczmarczyk, and Jenia Jitsev · 2022
Cited alongside, same era.
Score jacobian chaining: Lifting pretrained 2d diffusion models for 3d generation
Haochen Wang, Xiaodan Du, Jiahao Li, Raymond A. Yeh, and Gregory Shakhnarovich · 2022
Cited alongside, same era.
Xatlas: Mesh parameterization / uv unwrapping library
Jonathan Young · 2022
Cited alongside, same era.
Denis Zavadski, Johann-Friedrich Feiden, and Carsten Rother · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala · 2023
Later among the works it cites.
Free3d: Consistent novel view synthesis without 3d representation
Chuanxia Zheng and Andrea Vedaldi · 2023
Later among the works it cites.
Hifa: High-fidelity text-to-3d generation with advanced diffusion guidance
Junzhe Zhu, Peiye Zhuang, and Oluwasanmi Koyejo · 2023
Later among the works it cites.
Score distillation sampling with learned manifold corrective
Thiemo Alldieck, Nikos Kolotouros, and Cristian Sminchisescu · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xiaohui Zeng, Arash Vahdat, Francis Williams, Zan Gojcic, Or Litany, Sanja Fidler, and Karsten Kreis · 2022
Cited alongside, same era.
Sdf-stylegan: Implicit sdf-based stylegan for 3d shape generation
Xin Zheng, Yang Liu, Peng-Shuai Wang, and Xin Tong · 2022
Cited alongside, same era.
Diffusiondepth: Diffusion denoising approach for monocular depth estimation
Yiqun Duan, Xianda Guo, and Zhengbiao Zhu · 2023
Cited alongside, same era.
Lrm: Large reconstruction model for single image to 3d
Yicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi, Yang Zhou, Difan Liu, Feng Liu, Kalyan Sunkavalli, Trung Bui, and Hao Tan · 2023
Cited alongside, same era.
Animate anyone: Consistent and controllable image-to-video synthesis for character animation
Liucheng Hu, Xin Gao, Peng Zhang, Ke Sun, Bang Zhang, and Liefeng Bo · 2023
Cited alongside, same era.
Shap-e: Generating conditional 3d implicit functions
Heewoo Jun and Alex Nichol · 2023
Cited alongside, same era.
Oren Katzir, Or Patashnik, Daniel Cohen-Or, and Dani Lischinski · 2023
Cited alongside, same era.
Repurposing diffusion-based image generators for monocular depth estimation
Bing Wen Ke, Anton Obukhov, Shengyu Huang, Nando Metzger, Rodrigo Caye Daudt, and Konrad Schindler · 2023
Cited alongside, same era.
Closest in time.
Meta 3d gen, 2024
Raphael Bensadoun, Tom Monnier, Yanir Kleiman, Filippos Kokkinos, Yawar Siddiqui, Mahendra Kariya, Omri Harosh, Roman Shapovalov, Benjamin Graham, Emilien Garreau, Animesh Karnewar, Ang Cao, Idan Azuri, Iurii Makarov, Eric-Tuan Le, Antoine Toisoul, David Novotny, Oran Gafni, Natalia Neverova, and Andrea Vedaldi · 2024
Closest in time.
Sf3d: Stable fast 3d mesh reconstruction with uv-unwrapping and illumination disentanglement
Mark Boss, Zixuan Huang, Aaryaman Vasishta, and Varun Jampani · 2024
Closest in time.
Scaling rectified flow transformers for high-resolution image synthesis
Patrick Esser, Sumith Kulal, A. Blattmann, Rahim Entezari, Jonas Muller, Harry Saini, Yam Levi, Dominik Lorenz, Axel Sauer, Frederic Boesel, Dustin Podell, Tim Dockhorn, Zion English, Kyle Lacey, Alex Goodwin, Yannik Marek, and Robin Rombach · 2024
Closest in time.
Viewdiff: 3d-consistent image generation with text-to-image models
Lukas Höllein, Aljavz Bovzivc, Norman Muller, David Novotny, Hung-Yu Tseng, Christian Richardt, Michael Zollhofer, and Matthias Nießner · 2024
Closest in time.
3dtopia: Large text-to-3d generation model with hybrid diffusion priors
Fangzhou Hong, Jiaxiang Tang, Ziang Cao, Min Shi, Tong Wu, Zhaoxi Chen, Tengfei Wang, Liang Pan, Dahua Lin, and Ziwei Liu · 2024
Closest in time.
Pointinfinity: Resolution-invariant point diffusion models
Zixuan Huang, Justin Johnson, Shoubhik Debnath, James M. Rehg, and Chao-Yuan Wu · 2024
Closest in time.
Spad : Spatially aware multiview diffusers
Yash Kant, Ziyi Wu, Michael Vasilkovsky, Guocheng Qian, Jian Ren, Riza Alp Guler, Bernard Ghanem, S. Tulyakov, Igor Gilitschenski, and Aliaksandr Siarohin · 2024
Closest in time.
Common diffusion noise schedules and sample steps are flawed
Shanchuan Lin, Bingchen Liu, Jiashi Li, and Xiao Yang · 2024
Closest in time.
Würstchen: An efficient architecture for large-scale text-to-image diffusion models
Pablo Pernias, Dominic Rampas, Mats L. Richter, Christopher Pal, and Marc Aubreville · 2024
Closest in time.
Ipadapter-instruct: Resolving ambiguity in image-based conditioning using instruct prompts, 2024
Ciara Rowles, Shimon Vainer, Dante De Nigris, Slava Elizarov, Konstantin Kutsy, and Simon Donné · 2024
Closest in time.
Meta 3d assetgen: Text-to-mesh generation with high-quality geometry, texture, and pbr materials
Yawar Siddiqui, Tom Monnier, Filippos Kokkinos, Mahendra Kariya, Yanir Kleiman, Emilien Garreau, Oran Gafni, Natalia V. Neverova, Andrea Vedaldi, Roman Shapovalov, and David Novotny · 2024
Closest in time.
Triposr: Fast 3d object reconstruction from a single image
Dmitry Tochilkin, David Pankratz, Zexiang Liu, Zixuan Huang, Adam Letts, Yangguang Li, Ding Liang, Christian Laforte, Varun Jampani, and Yan-Pei Cao · 2024
Closest in time.
Huggingface zerodiffusion model weights v0.9
2024
Closest in time.
Collaborative control for geometry-conditioned pbr image generation
Shimon Vainer, Mark Boss, Mathias Parger, Konstantin Kutsy, Dante De Nigris, Ciara Rowles, Nicolas Perony, and Simon Donn’e · 2024
Closest in time.
Zike Wu, Pan Zhou, Xuanyu Yi, Xiaoding Yuan, and Hanwang Zhang · 2024
Closest in time.
Latte3d: Large-scale amortized text-to-enhanced3d synthesis
Kevin Xie, Jonathan Lorraine, Tianshi Cao, Jun Gao, James Lucas, Antonio Torralba, Sanja Fidler, and Xiaohui Zeng · 2024
Closest in time.
An object is worth 64x64 pixels: Generating 3d object via image diffusion
Xingguang Yan, Han-Hung Lee, Ziyu Wan, and Angel X. Chang · 2024
Closest in time.