Fetching the paper…
Reading the bibliography…
While 3D generative models have greatly improved artists' workflows, the existing diffusion models for 3D generation suffer from slow generation and poor generalization.
Marching cubes: A high resolution 3d surface construction algorithm
William E. Lorensen and Harvey E. Cline · 1987
Earlier work this paper cites.
Photorealistic scene reconstruction by voxel coloring
S.M. Seitz and C.R. Dyer · 1997
Earlier work this paper cites.
Poxels: Probabilistic voxelized volume reconstruction
Jeremy S. De Bonet · 1999
Earlier work this paper cites.
A theory of shape by space carving
K.N. Kutulakos and S.M. Seitz · 1999
Earlier work this paper cites.
A probabilistic framework for surface reconstruction from multiple images
M. Agrawal and L.S. Davis · 2001
Earlier work this paper cites.
A probabilistic framework for space carving
A. Broadhurst, T.W. Drummond, and R. Cipolla · 2001
Earlier work this paper cites.
A global linear method for camera pose registration
Nianjuan Jiang, Zhaopeng Cui, and Ping Tan · 2013
Earlier work this paper cites.
Learning to compare image patches via convolutional neural networks
Sergey Zagoruyko and Nikos Komodakis · 2015
Earlier work this paper cites.
Efficient deep learning for stereo matching
Wenjie Luo, Alexander G. Schwing, and Raquel Urtasun · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L. Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Learned multi-patch similarity
Wilfried Hartmann, Silvano Galliani, Michal Havlena, Luc Van Gool, and Konrad Schindler · 2017
Earlier work this paper cites.
Octnetfusion: Learning depth fusion from data
Gernot Riegler, Ali Osman Ulusoy, Horst Bischof, and Andreas Geiger · 2017
Earlier work this paper cites.
Multi-view supervision for single-view reconstruction via differentiable ray consistency
Shubham Tulsiani, Tinghui Zhou, Alexei A Efros, and Jitendra Malik · 2017
Earlier work this paper cites.
Demon: Depth and motion network for learning monocular stereo
Benjamin Ummenhofer, Huizhong Zhou, Jonas Uhrig, Nikolaus Mayer, Eddy Ilg, Alexey Dosovitskiy, and Thomas Brox · 2017
Earlier work this paper cites.
Deepmvs: Learning multi-view stereopsis
Po-Han Huang, Kevin Matzen, Johannes Kopf, Narendra Ahuja, and Jia-Bin Huang · 2018
Earlier work this paper cites.
Shape reconstruction using volume sweeping and learned photoconsistency
Vincent Leroy, Jean-Sébastien Franco, and Edmond Boyer · 2018
Earlier work this paper cites.
Raynet: Learning volumetric 3d reconstruction with ray potentials
Despoina Paschalidou, Osman Ulusoy, Carolin Schmitt, Luc Van Gool, and Andreas Geiger · 2018
Earlier work this paper cites.
Defusr: Learning non-volumetric depth fusion using successive reprojections
Simon Donne and Andreas Geiger · 2019
Earlier work this paper cites.
Recurrent mvsnet for high-resolution multi-view stereo depth inference
Yao Yao, Zixin Luo, Shiwei Li, Tianwei Shen, Tian Fang, and Long Quan · 2019
Earlier work this paper cites.
DIST: Rendering Deep Implicit Signed Distance Function With Differentiable Sphere Tracing
Shaohui Liu, Yinda Zhang, Songyou Peng, Boxin Shi, Marc Pollefeys, and Zhaopeng Cui · 2020
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Cited alongside, same era.
Differentiable volumetric rendering: Learning implicit 3d representations without 3d supervision
Michael Niemeyer, Lars Mescheder, Michael Oechsle, and Andreas Geiger · 2020
Cited alongside, same era.
Multiview neural surface reconstruction by disentangling geometry and appearance
Lior Yariv, Yoni Kasten, Dror Moran, Meirav Galun, Matan Atzmon, Basri Ronen, and Yaron Lipman · 2020
Cited alongside, same era.
Fast-mvsnet: Sparse-to-dense multi-view stereo with learned propagation and gauss-newton refinement
Zehao Yu and Shenghua Gao · 2020
Cited alongside, same era.
Classifier-free diffusion guidance
Jonathan Ho and Tim Salimans · 2021
Cited alongside, same era.
Mvdiffusion: Enabling holistic multi-view image generation with correspondence-aware diffusion
Shitao Tang, Fuayng Zhang, Jiacheng Chen, Peng Wang, and Furukawa Yasutaka · 2023
Later among the works it cites.
Consistent view synthesis with pose-guided diffusion models
Hung-Yu Tseng, Qinbo Li, Changil Kim, Suhib Alsisan, Jia-Bin Huang, and Johannes Kopf · 2023
Later among the works it cites.
Imagedream: Image-prompt multi-view diffusion for 3d generation
Peng Wang and Yichun Shi · 2023
Later among the works it cites.
Consistent123: Improve consistency for one image to 3d object synthesis
Haohan Weng, Tianyu Yang, Jianan Wang, Yu Li, Tong Zhang, CL Chen, and Lei Zhang · 2023
Later among the works it cites.
Omniobject3d: Large-vocabulary 3d object dataset for realistic perception, reconstruction and generation
Tong Wu, Jiarui Zhang, Xiao Fu, Yuxin Wang, Jiawei Ren, Liang Pan, Wayne Wu, Lei Yang, Jiaqi Wang, Chen Qian, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Michael Oechsle, Songyou Peng, and Andreas Geiger · 2021
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models, 2021
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2021
Cited alongside, same era.
Neus: Learning neural implicit surfaces by volume rendering for multi-view reconstruction
Peng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt, Taku Komura, and Wenping Wang · 2021
Cited alongside, same era.
Volume rendering of neural implicit surfaces
Lior Yariv, Jiatao Gu, Yoni Kasten, and Yaron Lipman · 2021
Cited alongside, same era.
Google scanned objects: A high-quality dataset of 3d scanned household items
Laura Downs, Anthony Francis, Nate Koenig, Brandon Kinman, Ryan Hickman, Krista Reymann, Thomas B McHugh, and Vincent Vanhoucke · 2022
Cited alongside, same era.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T Barron, and Ben Mildenhall · 2022
Cited alongside, same era.
Novel view synthesis with diffusion models, 2022
Daniel Watson, William Chan, Ricardo Martin-Brualla, Jonathan Ho, Andrea Tagliasacchi, and Mohammad Norouzi · 2022
Cited alongside, same era.
Later among the works it cites.
Dmv3d: Denoising multi-view diffusion using 3d large reconstruction model
Yinghao Xu, Hao Tan, Fujun Luan, Sai Bi, Peng Wang, Jiahao Li, Zifan Shi, Kalyan Sunkavalli, Gordon Wetzstein, Zexiang Xu, et al · 2023
Later among the works it cites.
https://huggingface.co/stabilityai/stable-zero123
Stablezero123 · 2024
Closest in time.
Sf3d: Stable fast 3d mesh reconstruction with uv-unwrapping and illumination disentanglement
Mark Boss, Zixuan Huang, Aaryaman Vasishta, and Varun Jampani · 2024
Closest in time.
Objaverse-xl: A universe of 10m+ 3d objects
Matt Deitke, Ruoshi Liu, Matthew Wallingford, Huong Ngo, Oscar Michel, Aditya Kusupati, Alan Fan, Christian Laforte, Vikram Voleti, Samir Yitzhak Gadre, et al · 2024
Closest in time.
Spad : Spatially aware multiview diffusers, 2024
Yash Kant, Ziyi Wu, Michael Vasilkovsky, Guocheng Qian, Jian Ren, Riza Alp Guler, Bernard Ghanem, Sergey Tulyakov, Igor Gilitschenski, and Aliaksandr Siarohin · 2024
Closest in time.
Grounding image matching in 3d with mast3r
Vincent Leroy, Yohann Cabon, and Jérôme Revaud · 2024
Closest in time.
One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization
Minghua Liu, Chao Xu, Haian Jin, Linghao Chen, Mukund Varma T, Zexiang Xu, and Hao Su · 2024
Closest in time.
Lgm: Large multi-view gaussian model for high-resolution 3d content creation
Jiaxiang Tang, Zhaoxi Chen, Xiaokang Chen, Tengfei Wang, Gang Zeng, and Ziwei Liu · 2024
Closest in time.
Triposr: Fast 3d object reconstruction from a single image
Dmitry Tochilkin, David Pankratz, Zexiang Liu, Zixuan Huang, , Adam Letts, Yangguang Li, Ding Liang, Christian Laforte, Varun Jampani, and Yan-Pei Cao · 2024
Closest in time.
SV3D: Novel multi-view synthesis and 3D generation from a single image using latent video diffusion
Vikram Voleti, Chun-Han Yao, Mark Boss, Adam Letts, David Pankratz, Dmitrii Tochilkin, Christian Laforte, Robin Rombach, and Varun Jampani · 2024
Closest in time.
Meshlrm: Large reconstruction model for high-quality mesh
Xinyue Wei, Kai Zhang, Sai Bi, Hao Tan, Fujun Luan, Valentin Deschaintre, Kalyan Sunkavalli, Hao Su, and Zexiang Xu · 2024
Closest in time.
Instantmesh: Efficient 3d mesh generation from a single image with sparse-view large reconstruction models, 2024
Jiale Xu, Weihao Cheng, Yiming Gao, Xintao Wang, Shenghua Gao, and Ying Shan · 2024
Closest in time.
Viewfusion: Towards multi-view consistency via interpolated denoising, 2024
Xianghui Yang, Yan Zuo, Sameera Ramasinghe, Loris Bazzani, Gil Avraham, and Anton van den Hengel · 2024
Closest in time.
Gs-lrm: Large reconstruction model for 3d gaussian splatting
Kai Zhang, Sai Bi, Hao Tan, Yuanbo Xiangli, Nanxuan Zhao, Kalyan Sunkavalli, and Zexiang Xu · 2025
Closest in time.