Fetching the paper…
Reading the bibliography…
We propose Flash3D, a method for scene reconstruction and novel view synthesis from a single image which is both very generalisable and efficient.
Monocular depth perception from optical flow by space time signal processing
B. F. Buxton and H. Buxton · 1983
Earlier work this paper cites.
Layered representations for vision and video
E. Adelson · 1995
Earlier work this paper cites.
Optical models for direct volume rendering
Nelson L. Max · 1995
Earlier work this paper cites.
A layered approach to stereo reconstruction
Simon Baker, R. Szeliski, and P. Anandan · 1998
Earlier work this paper cites.
Depth discontinuities by pixel-to-pixel stereo
Stan Birchfield and Carlo Tomasi · 1998
Earlier work this paper cites.
Layered depth images
J. Shade, S. Gortler, Li wei He, and R. Szeliski · 1998
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, Alan C. Bovik, Hamid R. Sheikh, and Eero P. Simoncelli · 2004
Earlier work this paper cites.
Are we ready for autonomous driving? the KITTI vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from RGBD images
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus · 2012
Earlier work this paper cites.
Vision meets robotics: The KITTI dataset
Andreas Geiger, Philip Lenz, Christoph Stiller, and Raquel Urtasun · 2013
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
David Eigen, Christian Puhrsch, and Rob Fergus · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Unsupervised monocular depth estimation with left-right consistency
Clément Godard, Oisin Mac Aodha, and Gabriel J. Brostow · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Deeper depth prediction with fully convolutional residual networks
Iro Laina, Christian Rupprecht, Vasileios Belagiannis, Federico Tombari, and Nassir Navab · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Learning a multi-view stereo machine
Abhishek Kar, Christian Häne, and Jitendra Malik · 2017
Earlier work this paper cites.
Multi-view supervision for single-view reconstruction via differentiable ray consistency
Shubham Tulsiani, Tinghui Zhou, Alexei A. Efros, and Jitendra Malik · 2017
Earlier work this paper cites.
Unsupervised learning of depth and ego-motion from video
Tinghui Zhou, Matthew Brown, Noah Snavely, and David G. Lowe · 2017
Earlier work this paper cites.
Estimating depth from monocular images as classification using deep fully convolutional residual networks
Yuanzhouhan Cao, Zifeng Wu, and Chunhua Shen · 2018
Earlier work this paper cites.
Stereo magnification: Learning view synthesis using multiplane images
Zhou Tinghui, Tucker Richard, Flynn John, Fyffe Graham, and Snavely Noah · 2018
Earlier work this paper cites.
Layer-structured 3d scene inference via view synthesis
Shubham Tulsiani, Richard Tucker, and Noah Snavely · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Richard Zhang, Phillip Isola, Alexei A. Efros, Eli Shechtman, and Oliver Wang · 2018
Earlier work this paper cites.
Learning the depths of moving people by watching frozen people
Bill Freeman, Ce Liu, Forrester Cole, Noah Snavely, Richard Tucker, Tali Dekel, and Zhengqi Li · 2019
Earlier work this paper cites.
Digging into self-supervised monocular depth prediction
Clément Godard, Oisin Mac Aodha, Michael Firman, and Gabriel J. Brostow · 2019
Earlier work this paper cites.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
Katrin Lasinger, René Ranftl, Konrad Schindler, and Vladlen Koltun · 2019
Earlier work this paper cites.
Local light field fusion: practical view synthesis with prescriptive sampling guidelines
Ben Mildenhall, Pratul P. Srinivasan, Rodrigo Ortiz Cayon, Nima Khademi Kalantari, Ravi Ramamoorthi, Ren Ng, and Abhishek Kar · 2019
Earlier work this paper cites.
Scene representation networks: Continuous 3D-structure-aware neural scene representations
Vincent Sitzmann, Michael Zollhöfer, and Gordon Wetzstein · 2019
Earlier work this paper cites.
Pushing the boundaries of view extrapolation with multiplane images
Pratul P. Srinivasan, Richard Tucker, Jonathan T. Barron, Ravi Ramamoorthi, Ren Ng, and Noah Snavely · 2019
Earlier work this paper cites.
NeRF: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Earlier work this paper cites.
RealMonoDepth: Self-supervised monocular depth estimation for general scenes
Mertalp Ocal and Armin Mustafa · 2020
Cited alongside, same era.
Single-view view synthesis with multiplane images
Richard Tucker and Noah Snavely · 2020
Cited alongside, same era.
Synsin: End-to-end view synthesis from a single image
Olivia Wiles, Georgia Gkioxari, Richard Szeliski, and Justin Johnson · 2020
Cited alongside, same era.
MVSNeRF: Fast generalizable radiance field reconstruction from multi-view stereo
Anpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang, Fanbo Xiang, Jingyi Yu, and Hao Su · 2021
Cited alongside, same era.
Stereo radiance fields (SRF): learning view synthesis for sparse views of novel scenes
Julian Chibane, Aayush Bansal, Verica Lazova, and Gerard Pons-Moll · 2021
Cited alongside, same era.
3D gaussian splatting for real-time radiance field rendering
Bernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, and George Drettakis · 2023
Later among the works it cites.
Zero-1-to-3: Zero-shot one image to 3D object
Ruoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov, Sergey Zakharov, and Carl Vondrick · 2023
Later among the works it cites.
RealFusion: 360 reconstruction of any object from a single image
Luke Melas-Kyriazi, Christian Rupprecht, Iro Laina, and Andrea Vedaldi · 2023
Later among the works it cites.
GTA: A geometry-aware attention mechanism for multi-view transformers
Takeru Miyato, Bernhard Jaeger, Max Welling, and Andreas Geiger · 2023
Later among the works it cites.
SDXL: improving latent diffusion models for high-resolution image synthesis
Dustin Podell, Zion English, Kyle Lacey, Andreas Blattmann, Tim Dockhorn, Jonas Müller, Joe Penna, and Robin Rombach · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Philipp Henzler, Jeremy Reizenstein, Patrick Labatut, Roman Shapovalov, Tobias Ritschel, Andrea Vedaldi, and David Novotny · 2021
Cited alongside, same era.
Putting nerf on a diet: Semantically consistent few-shot view synthesis
Ajay Jain, Matthew Tancik, and Pieter Abbeel · 2021
Cited alongside, same era.
MINE: towards continuous depth MPI with nerf for novel view synthesis
Jiaxin Li, Zijian Feng, Qi She, Henghui Ding, Changhu Wang, and Gim Hee Lee · 2021
Cited alongside, same era.
Vision transformers for dense prediction
René Ranftl, Alexey Bochkovskiy, and Vladlen Koltun · 2021
Cited alongside, same era.
Common Objects in 3D: Large-scale learning and evaluation of real-life 3D category reconstruction
Jeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone, Patrick Labatut, and David Novotny · 2021
Cited alongside, same era.
PixelSynth: Generating a 3D-consistent experience from a single image
Chris Rockwell, David F. Fouhey, and Justin Johnson · 2021
Cited alongside, same era.
Geometry-free view synthesis: Transformers and no 3d priors
Robin Rombach, Patrick Esser, and Björn Ommer · 2021
Cited alongside, same era.
DreamFusion: Text-to-3D using 2D diffusion
Ben Poole, Ajay Jain, Jonathan T. Barron, and Ben Mildenhall · 2023
Later among the works it cites.
ZeroNVS: Zero-shot 360-degree view synthesis from a single real image
Kyle Sargent, Zizhang Li, Tanmay Shah, Charles Herrmann, Hong-Xing Yu, Yunzhi Zhang, Eric Ryan Chan, Dmitry Lagun, Li Fei-Fei, Deqing Sun, and Jiajun Wu · 2023
Later among the works it cites.
Monocular depth estimation using diffusion models
Saurabh Saxena, Abhishek Kar, Mohammad Norouzi, and David J. Fleet · 2023
Later among the works it cites.
Nddepth: Normal-distance assisted monocular depth estimation
Shuwei Shao, Zhongcai Pei, Weihai Chen, Xingming Wu, and Zhengguo Li · 2023
Later among the works it cites.
Viewset diffusion: (0-)image-conditioned 3D generative models from 2D data
Stanislaw Szymanowicz, Christian Rupprecht, and Andrea Vedaldi · 2023
Later among the works it cites.
Make-It-3D: High-fidelity 3d creation from A single image with diffusion prior
Junshu Tang, Tengfei Wang, Bo Zhang, Ting Zhang, Ran Yi, Lizhuang Ma, and Dong Chen · 2023
Later among the works it cites.
Diffusion with forward models: Solving stochastic inverse problems without direct supervision
Ayush Tewari, Tianwei Yin, George Cazenavette, Semon Rezchikov, Joshua B. Tenenbaum, Frédo Durand, William T. Freeman, and Vincent Sitzmann · 2023
Later among the works it cites.
Consistent view synthesis with pose-guided diffusion models
Hung-Yu Tseng, Qinbo Li, Changil Kim, Suhib Alsisan, Jia-Bin Huang, and Johannes Kopf · 2023
Later among the works it cites.
Novel view synthesis with diffusion models
Daniel Watson, William Chan, Ricardo Martin-Brualla, Jonathan Ho, Andrea Tagliasacchi, and Mohammad Norouzi · 2023
Later among the works it cites.
Behind the scenes: Density fields for single view reconstruction
Felix Wimbauer, Nan Yang, Christian Rupprecht, and Daniel Cremers · 2023
Later among the works it cites.
DiffusioNeRF: regularizing neural radiance fields with denoising diffusion models
Jamie Wynn and Daniyar Turmukhambetov · 2023
Later among the works it cites.
MuRF: Multi-baseline radiance fields
Haofei Xu, Anpei Chen, Yuedong Chen, Christos Sakaridis, Yulun Zhang, Marc Pollefeys, Andreas Geiger, and Fisher Yu · 2023
Later among the works it cites.
Metric3d: Towards zero-shot metric 3d prediction from a single image
Wei Yin, Chi Zhang, Hao Chen, Zhipeng Cai, Gang Yu, Kaixuan Wang, Xiaozhi Chen, and Chunhua Shen · 2023
Later among the works it cites.
Explicit correspondence matching for generalizable neural radiance fields
Chen Yuedong, Xu Haofei, Wu Qianyi, Zheng Chuanxia, Cham Tat-Jen, and Cai Jianfei · 2023
Later among the works it cites.
MVSplat: efficient 3d gaussian splatting from sparse multi-view images
Yuedong Chen, Haofei Xu, Chuanxia Zheng, Bohan Zhuang, Marc Pollefeys, Andreas Geiger, Tat-Jen Cham, and Jianfei Cai · 2024
Closest in time.
Cat3d: Create anything in 3d with multi-view diffusion models
Ruiqi Gao, Aleksander Holynski, Philipp Henzler, Arthur Brussee, Ricardo Martin-Brualla, Pratul Srinivasan, Jonathan T Barron, and Ben Poole · 2024
Closest in time.
LRM: Large reconstruction model for single image to 3D
Yicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi, Yang Zhou, Difan Liu, Feng Liu, Kalyan Sunkavalli, Trung Bui, and Hao Tan · 2024
Closest in time.
Mu Hu, Wei Yin, Chi Zhang, Zhipeng Cai, Xiaoxiao Long, Hao Chen, Kaixuan Wang, Gang Yu, Chunhua Shen, and Shaojie Shen · 2024
Closest in time.
Instant3D: Fast text-to-3D with sparse-view generation and large reconstruction model
Jiahao Li, Hao Tan, Kai Zhang, Zexiang Xu, Fujun Luan, Yinghao Xu, Yicong Hong, Kalyan Sunkavalli, Greg Shakhnarovich, and Sai Bi · 2024
Closest in time.
IM-3D: Iterative multiview diffusion and reconstruction for high-quality 3D generation
Luke Melas-Kyriazi, Iro Laina, Christian Rupprecht, Natalia Neverova, Andrea Vedaldi, Oran Gafni, and Filippos Kokkinos · 2024
Closest in time.
MVDream: Multi-view diffusion for 3D generation
Yichun Shi, Peng Wang, Jianglong Ye, Mai Long, Kejie Li, and Xiao Yang · 2024
Closest in time.
Splatter Image: Ultra-fast single-view 3D reconstruction
Stanislaw Szymanowicz, Christian Rupprecht, and Andrea Vedaldi · 2024
Closest in time.
LGM: Large multi-view Gaussian model for high-resolution 3D content creation
Jiaxiang Tang, Zhaoxi Chen, Xiaokang Chen, Tengfei Wang, Gang Zeng, and Ziwei Liu · 2024
Closest in time.
latentsplat: Autoencoding variational gaussians for fast generalizable 3d reconstruction
Christopher Wewer, Kevin Raj, Eddy Ilg, Bernt Schiele, and Jan Eric Lenssen · 2024
Closest in time.
Reconfusion: 3d reconstruction with diffusion priors
Rundi Wu, Ben Mildenhall, Philipp Henzler, Keunhong Park, Ruiqi Gao, Daniel Watson, Pratul P Srinivasan, Dor Verbin, Jonathan T Barron, Ben Poole, et al · 2024
Closest in time.
GRM: Large gaussian reconstruction model for efficient 3D reconstruction and generation
Yinghao Xu, Zifan Shi, Wang Yifan, Hansheng Chen, Ceyuan Yang, Sida Peng, Yujun Shen, and Gordon Wetzstein · 2024
Closest in time.
Free3D: Consistent novel view synthesis without 3D representation
Chuanxia Zheng and Andrea Vedaldi · 2024
Closest in time.