Fetching the paper…
Reading the bibliography…
We present SCube, a novel method for reconstructing large-scale 3D scenes (geometry, appearance, and semantics) from a sparse set of posed images.
Image quality assessment: from error visibility to structural similarity
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli · 2004
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Earlier work this paper cites.
Structure-from-motion revisited
J. L. Schonberger and J.-M. Frahm · 2016
Earlier work this paper cites.
Submanifold sparse convolutional networks
B. Graham and L. Van der Maaten · 2017
Earlier work this paper cites.
Focal loss for dense object detection
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár · 2017
Earlier work this paper cites.
Semantic scene completion from a single depth image
S. Song, F. Yu, A. Zeng, A. X. Chang, M. Savva, and T. Funkhouser · 2017
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang · 2018
Earlier work this paper cites.
PyTorch Lightning, Mar. 2019
W. Falcon and The PyTorch Lightning team · 2019
Earlier work this paper cites.
Singan: Learning a generative model from a single natural image
T. R. Shaham, T. Dekel, and T. Michaeli · 2019
Earlier work this paper cites.
Deepvoxels: Learning persistent 3d feature embeddings
V. Sitzmann, J. Thies, F. Heide, M. Nießner, G. Wetzstein, and M. Zollhofer · 2019
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
J. Ho, A. Jain, and P. Abbeel · 2020
Earlier work this paper cites.
Convolutional occupancy networks
S. Peng, M. Niemeyer, L. Mescheder, M. Pollefeys, and A. Geiger · 2020
Earlier work this paper cites.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
J. Philion and S. Fidler · 2020
Earlier work this paper cites.
Video deblurring by fitting to test data
X. Ren, Z. Qian, and Q. Chen · 2020
Earlier work this paper cites.
Metasdf: Meta-learning signed distance functions
V. Sitzmann, E. Chan, R. Tucker, N. Snavely, and G. Wetzstein · 2020
Earlier work this paper cites.
Scalability in perception for autonomous driving: Waymo open dataset
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine, et al · 2020
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng · 2021
Earlier work this paper cites.
Neuralrecon: Real-time coherent 3d reconstruction from monocular video
J. Sun, Y. Xie, L. Chen, X. Zhou, and H. Bao · 2021
Earlier work this paper cites.
Segformer: Simple and efficient design for semantic segmentation with transformers
E. Xie, W. Wang, Z. Yu, A. Anandkumar, J. M. Alvarez, and P. Luo · 2021
Earlier work this paper cites.
pixelnerf: Neural radiance fields from one or few images
A. Yu, V. Ye, M. Tancik, and A. Kanazawa · 2021
Earlier work this paper cites.
Gensdf: Two-stage learning of generalizable signed distance functions
G. Chou, I. Chugunov, and F. Heide · 2022
Earlier work this paper cites.
Depth-supervised nerf: Fewer views and faster training for free
K. Deng, A. Liu, J.-Y. Zhu, and D. Ramanan · 2022
Earlier work this paper cites.
Get3d: A generative model of high quality 3d textured shapes learned from images
J. Gao, T. Shen, Z. Wang, W. Chen, K. Yin, D. Li, O. Litany, Z. Gojcic, and S. Fidler · 2022
Earlier work this paper cites.
Shape, light, and material decomposition from images using monte carlo rendering and denoising
J. Hasselgren, N. Hofmann, and J. Munkberg · 2022
Earlier work this paper cites.
Subdivision-based mesh convolution networks
S.-M. Hu, Z.-N. Liu, M.-H. Guo, J.-X. Cai, J. Huang, T.-J. Mu, and R. R. Martin · 2022
Cited alongside, same era.
Dynamic 3d scene analysis by point cloud accumulation
S. Huang, Z. Gojcic, J. Huang, A. Wieser, and K. Schindler · 2022
Cited alongside, same era.
Instant neural graphics primitives with a multiresolution hash encoding
T. Müller, A. Evans, C. Schied, and A. Keller · 2022
Cited alongside, same era.
Dreamfusion: Text-to-3d using 2d diffusion
B. Poole, A. Jain, J. T. Barron, and B. Mildenhall · 2022
Cited alongside, same era.
Look outside the room: Synthesizing A consistent long-term 3d scene video from A single image
X. Ren and X. Wang · 2022
Cited alongside, same era.
Torchsparse: Efficient point cloud inference engine
Cogvlm: Visual expert for pretrained language models
W. Wang, Q. Lv, W. Yu, W. Hong, J. Qi, Y. Wang, J. Ji, Z. Yang, L. Zhao, X. Song, et al · 2023
Later among the works it cites.
Surroundocc: Multi-camera 3d occupancy prediction for autonomous driving
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, J. Zhou, and J. Lu · 2023
Later among the works it cites.
Reconfusion: 3d reconstruction with diffusion priors
R. Wu, B. Mildenhall, P. Henzler, K. Park, R. Gao, D. Watson, P. P. Srinivasan, D. Verbin, J. T. Barron, B. Poole, et al · 2023
Later among the works it cites.
Dmv3d: Denoising multi-view diffusion using 3d large reconstruction model
Y. Xu, H. Tan, F. Luan, S. Bi, P. Wang, J. Li, Z. Shi, K. Sunkavalli, G. Wetzstein, Z. Xu, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Tang, Z. Liu, X. Li, Y. Lin, and S. Han · 2022
Cited alongside, same era.
Monosdf: Exploring monocular geometric cues for neural implicit surface reconstruction
Z. Yu, S. Peng, M. Niemeyer, T. Sattler, and A. Geiger · 2022
Cited alongside, same era.
Zip-nerf: Anti-aliased grid-based neural radiance fields
J. T. Barron, B. Mildenhall, D. Verbin, P. P. Srinivasan, and P. Hedman · 2023
Cited alongside, same era.
Align your latents: High-resolution video synthesis with latent diffusion models
A. Blattmann, R. Rombach, H. Ling, T. Dockhorn, S. W. Kim, S. Fidler, and K. Kreis · 2023
Cited alongside, same era.
pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction
D. Charatan, S. Li, A. Tagliasacchi, and V. Sitzmann · 2023
Cited alongside, same era.
Flashattention-2: Faster attention with better parallelism and work partitioning
T. Dao · 2023
Cited alongside, same era.
Relightable 3d gaussian: Real-time point cloud relighting with brdf decomposition and ray tracing
J. Gao, C. Gu, Y. Lin, H. Zhu, X. Cao, L. Zhang, and Y. Yao · 2023
Cited alongside, same era.
X. Yu, Y.-C. Guo, Y. Li, D. Liang, S.-H. Zhang, and X. Qi · 2023
Later among the works it cites.
Locally attentional sdf diffusion for controllable 3d shape generation
X.-Y. Zheng, H. Pan, P.-S. Wang, X. Tong, Y. Liu, and H.-Y. Shum · 2023
Later among the works it cites.
Drivinggaussian: Composite gaussian splatting for surrounding dynamic autonomous driving scenes
X. Zhou, Z. Lin, X. Shan, Y. Wang, D. Sun, and M.-H. Yang · 2023
Later among the works it cites.
G3r: Gradient guided generalizable reconstruction
Y. Chen, J. Wang, Z. Yang, S. Manivasagam, and R. Urtasun · 2024
Closest in time.
Mvsplat: Efficient 3d gaussian splatting from sparse multi-view images
Y. Chen, H. Xu, C. Zheng, B. Zhuang, M. Pollefeys, A. Geiger, T.-J. Cham, and J. Cai · 2024
Closest in time.
6img-to-3d: Few-image large-scale outdoor driving scene reconstruction
T. Gieruc, M. Kästingschäfer, S. Bernhard, and M. Salzmann · 2024
Closest in time.
M. Hu, W. Yin, C. Zhang, Z. Cai, X. Long, H. Chen, K. Wang, G. Yu, C. Shen, and S. Shen · 2024
Closest in time.
Mvsgaussian: Fast generalizable gaussian splatting reconstruction from multi-view stereo
T. Liu, G. Wang, S. Hu, L. Shen, X. Ye, Y. Zang, Z. Cao, W. Li, and Z. Liu · 2024
Closest in time.
One-step image translation with text-to-image models
G. Parmar, T. Park, S. Narasimhan, and J.-Y. Zhu · 2024
Closest in time.
Octree-gs: Towards consistent real-time rendering with lod-structured 3d gaussians
K. Ren, L. Jiang, T. Lu, M. Yu, L. Xu, Z. Ni, and B. Dai · 2024
Closest in time.
Xcube: Large-scale 3d generative modeling using sparse voxel hierarchies
X. Ren, J. Huang, X. Zeng, K. Museth, S. Fidler, and F. Williams · 2024
Closest in time.
Realmdreamer: Text-driven 3d scene generation with inpainting and depth diffusion
J. Shriram, A. Trevithick, L. Liu, and R. Ramamoorthi · 2024
Closest in time.
Triposr: Fast 3d object reconstruction from a single image
D. Tochilkin, D. Pankratz, Z. Liu, Z. Huang, A. Letts, Y. Li, D. Liang, C. Laforte, V. Jampani, and Y.-P. Cao · 2024
Closest in time.
Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion
V. Voleti, C.-H. Yao, M. Boss, A. Letts, D. Pankratz, D. Tochilkin, C. Laforte, R. Rombach, and V. Jampani · 2024
Closest in time.
Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation
Z. Wang, C. Lu, Y. Wang, F. Bao, C. Li, H. Su, and J. Zhu · 2024
Closest in time.
Street gaussians for modeling dynamic urban scenes
Y. Yan, H. Lin, C. Zhou, W. Wang, H. Sun, K. Zhan, X. Lang, X. Zhou, and S. Peng · 2024
Closest in time.
Absgs: Recovering fine details for 3d gaussian splatting
Z. Ye, W. Li, S. Liu, P. Qiao, and Y. Dou · 2024
Closest in time.
Gaussian opacity fields: Efficient and compact surface reconstruction in unbounded scenes
Z. Yu, T. Sattler, and A. Geiger · 2024
Closest in time.
3d-scenedreamer: Text-driven 3d-consistent scene generation
F. Zhang, Y. Zhang, Q. Zheng, R. Ma, W. Hua, H. Bao, W. Xu, and C. Zou · 2024
Closest in time.
Gs-lrm: Large reconstruction model for 3d gaussian splatting
K. Zhang, S. Bi, H. Tan, Y. Xiangli, N. Zhao, K. Sunkavalli, and Z. Xu · 2024
Closest in time.
Lidardm: Generative lidar simulation in a generated world
V. Zyrianov, H. Che, Z. Liu, and S. Wang · 2024
Closest in time.