Fetching the paper…
Reading the bibliography…
High-fidelity 3D scene reconstruction from monocular videos continues to be challenging, especially for complete and fine-grained geometry reconstruction.
Z. Yu and S. Gao, “Fast-mvsnet: Sparse-to-dense multi-view stereo with learned propagation and gauss-newton refinement,” in IEEE CVPR , 2020, pp. 1946–1955
1955
Earlier work this paper cites.
W. E. Lorensen and H. E. Cline, “Marching cubes: A high resolution 3d surface construction algorithm,” in ACM SIGGRAPH , 1987, pp. 163–169
1987
Earlier work this paper cites.
B. Curless and M. Levoy, “A volumetric method for building complex models from range images,” in ACM SIGGRAPH , 1996, pp. 303–312
1996
Earlier work this paper cites.
N. Snavely, S. M. Seitz, and R. Szeliski, “Photo tourism: exploring photo collections in 3d,” in ACM SIGGRAPH , 2006, pp. 835–846
2006
Earlier work this paper cites.
R. A. Newcombe, S. Izadi, O. Hilliges, D. Molyneaux, D. Kim, A. J. Davison, P. Kohli, J. Shotton, S. Hodges, and A. W. Fitzgibbon, “Kinectfusion: Real-time dense surface mapping and tracking,” in IEEE ISMAR , 2011, pp. 127–136
2011
Earlier work this paper cites.
S. Agarwal, Y. Furukawa, N. Snavely, I. Simon, B. Curless, S. M. Seitz, and R. Szeliski, “Building rome in a day,” Communications of the ACM , vol. 54, no. 10, pp. 105–112, 2011
2011
Earlier work this paper cites.
J. Sturm, N. Engelhard, F. Endres, W. Burgard, and D. Cremers, “A benchmark for the evaluation of RGB-D SLAM systems,” in IEEE/RSJ IROS , 2012, pp. 573–580
2012
Earlier work this paper cites.
D. Eigen, C. Puhrsch, and R. Fergus, “Depth map prediction from a single image using a multi-scale deep network,” in NeurIPS , 2014, pp. 2366–2374
2014
Earlier work this paper cites.
O. Kähler, V. A. Prisacariu, C. Y. Ren, X. Sun, P. H. S. Torr, and D. W. Murray, “Very high frame rate volumetric integration of depth images on mobile devices,” IEEE Transactions on Visualization and Computer Graphics , vol. 21, no. 11, pp. 1241–1250, 2015
2015
Earlier work this paper cites.
J. L. Schönberger, E. Zheng, J. Frahm, and M. Pollefeys, “Pixelwise view selection for unstructured multi-view stereo,” in ECCV , 2016, pp. 501–518
2016
Earlier work this paper cites.
J. L. Schönberger and J. Frahm, “Structure-from-motion revisited,” in IEEE CVPR , 2016, pp. 4104–4113
2016
Earlier work this paper cites.
A. Dai, M. Nießner, M. Zollhöfer, S. Izadi, and C. Theobalt, “Bundlefusion: Real-time globally consistent 3d reconstruction using on-the-fly surface reintegration,” ACM Transactions on Graphics , vol. 36, no. 3, pp. 24:1–24:18, 2017
2017
Earlier work this paper cites.
K. Tateno, F. Tombari, I. Laina, and N. Navab, “Cnn-slam: Real-time dense monocular slam with learned depth prediction,” in IEEE CVPR , 2017, pp. 6243–6252
2017
Earlier work this paper cites.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. A. Funkhouser, and M. Nießner, “Scannet: Richly-annotated 3d reconstructions of indoor scenes,” in IEEE CVPR , 2017, pp. 2432–2443
2017
Earlier work this paper cites.
M. Ji, J. Gall, H. Zheng, Y. Liu, and L. Fang, “Surfacenet: An end-to-end 3d neural network for multiview stereopsis,” in IEEE ICCV , 2017, pp. 2326–2334
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
T. Qin, P. Li, and S. Shen, “Vins-mono: A robust and versatile monocular visual-inertial state estimator,” IEEE Transactions on Robotics , vol. 34, no. 4, pp. 1004–1020, 2018
2018
Earlier work this paper cites.
Y.-P. Cao, L. Kobbelt, and S.-M. Hu, “Real-time high-accuracy three-dimensional reconstruction with consumer RGB-D cameras,” ACM Transactions on Graphics , vol. 37, no. 5, pp. 171:1–171:16, 2018
2018
Earlier work this paper cites.
J. P. C. Valentin, A. Kowdle, J. T. Barron, N. Wadhwa, M. Dzitsiuk, M. Schoenberg, V. Verma, A. Csaszar, E. Turner, I. Dryanovski, J. Afonso, J. Pascoal, K. Tsotsos, M. Leung, M. Schmidt, O. G. Guleryuz, S. Khamis, V. Tankovich, S. R. Fanello, S. Izadi, and C. Rhemann, “Depth from motion for smartphone AR,” ACM Transactions on Graphics , vol. 37, no. 6, pp. 193:1–193:19, 2018
2018
Earlier work this paper cites.
M. Bloesch, J. Czarnowski, R. Clark, S. Leutenegger, and A. J. Davison, “Codeslam - learning a compact, optimisable representation for dense visual SLAM,” in IEEE CVPR , 2018, pp. 2560–2568
2018
Earlier work this paper cites.
K. Wang and S. Shen, “Mvdepthnet: Real-time multiview depth estimation neural network,” in 3DV , 2018, pp. 248–257
2018
Earlier work this paper cites.
S. Zhi, M. Bloesch, S. Leutenegger, and A. J. Davison, “Scenecode: Monocular dense semantic reconstruction using learned encoded scene representations,” in IEEE CVPR , 2019, pp. 11 776–11 785
2019
Earlier work this paper cites.
A. Gordon, H. Li, R. Jonschkowski, and A. Angelova, “Depth from videos in the wild: Unsupervised monocular depth learning from unknown cameras,” in IEEE ICCV , 2019, pp. 8977–8986
2019
Earlier work this paper cites.
C. Godard, O. M. Aodha, M. Firman, and G. J. Brostow, “Digging into self-supervised monocular depth estimation,” in IEEE ICCV , 2019, pp. 3827–3837
2019
Cited alongside, same era.
2019
Cited alongside, same era.
C. Liu, J. Gu, K. Kim, S. G. Narasimhan, and J. Kautz, “Neural rgb(r)d sensing: Depth and uncertainty from a video camera,” in IEEE CVPR , 2019, pp. 10 986–10 995
2019
Cited alongside, same era.
Y. Hou, J. Kannala, and A. Solin, “Multi-view stereo by temporal nonparametric fusion,” in IEEE ICCV , 2019, pp. 2651–2660
2019
Cited alongside, same era.
P. Wang, L. Liu, Y. Liu, C. Theobalt, T. Komura, and W. Wang, “Neus: Learning neural implicit surfaces by volume rendering for multi-view reconstruction,” in NeurIPS , 2021, pp. 27 171–27 183
2021
Later among the works it cites.
L. Yariv, J. Gu, Y. Kasten, and Y. Lipman, “Volume rendering of neural implicit surfaces,” in NeurIPS , 2021, pp. 4805–4815
2021
Later among the works it cites.
H. Matsuki, R. Scona, J. Czarnowski, and A. J. Davison, “Codemapping: Real-time dense mapping for sparse slam using compact scene representations,” IEEE Robotics and Automation Letters , vol. 6, no. 4, pp. 7105–7112, 2021
2021
Later among the works it cites.
A. Düzçeker, S. Galliani, C. Vogel, P. Speciale, M. Dusmanu, and M. Pollefeys, “Deepvideomvs: Multi-view stereo on video with recurrent spatio-temporal fusion,” in IEEE CVPR , 2021, pp. 15 324–15 333
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. J. Park, P. Florence, J. Straub, R. A. Newcombe, and S. Lovegrove, “Deepsdf: Learning continuous signed distance functions for shape representation,” in IEEE CVPR , 2019, pp. 165–174
2019
Cited alongside, same era.
L. M. Mescheder, M. Oechsle, M. Niemeyer, S. Nowozin, and A. Geiger, “Occupancy networks: Learning 3d reconstruction in function space,” in IEEE CVPR , 2019, pp. 4460–4470
2019
Cited alongside, same era.
M. Tan, B. Chen, R. Pang, V. Vasudevan, M. Sandler, A. Howard, and Q. V. Le, “Mnasnet: Platform-aware neural architecture search for mobile,” in IEEE CVPR , 2019, pp. 2820–2828
2019
Cited alongside, same era.
R. Chen, S. Han, J. Xu, and H. Su, “Point-based multi-view stereo network,” in IEEE ICCV , 2019, pp. 1538–1547
2019
Cited alongside, same era.
X. Yang, L. Zhou, H. Jiang, Z. Tang, Y. Wang, H. Bao, and G. Zhang, “Mobile3drecon: real-time monocular 3d reconstruction on a mobile phone,” IEEE Transactions on Visualization and Computer Graphics , vol. 26, no. 12, pp. 3446–3456, 2020
2020
Cited alongside, same era.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “Nerf: Representing scenes as neural radiance fields for view synthesis,” in ECCV , 2020, pp. 405–421
2020
Cited alongside, same era.
L. Liu, J. Gu, K. Z. Lin, T. Chua, and C. Theobalt, “Neural sparse voxel fields,” in NeurIPS , 2020, pp. 15 651–15 663
2020
Cited alongside, same era.
Z. Murez, T. van As, J. Bartolozzi, A. Sinha, V. Badrinarayanan, and A. Rabinovich, “Atlas: End-to-end 3d scene reconstruction from posed images,” in ECCV , 2020, pp. 414–431
2020
Cited alongside, same era.
A. Rich, N. Stier, P. Sen, and T. Höllerer, “3dvnet: Multi-view depth prediction and volumetric refinement,” in 3DV , 2021, pp. 700–709
2021
Later among the works it cites.
J. Choe, S. Im, F. Rameau, M. Kang, and I. S. Kweon, “Volumefusion: Deep depth fusion for 3d scene reconstruction,” in IEEE CVPR , 2021, pp. 16 086–16 095
2021
Later among the works it cites.
N. Stier, A. Rich, P. Sen, and T. Höllerer, “Vortx: Volumetric 3d reconstruction with transformers for voxelwise view selection and fusion,” in 3DV , 2021, pp. 320–330
2021
Later among the works it cites.
S. J. Garbin, M. Kowalski, M. Johnson, J. Shotton, and J. P. C. Valentin, “Fastnerf: High-fidelity neural rendering at 200fps,” in IEEE ICCV , 2021, pp. 14 326–14 335
2021
Later among the works it cites.
J. T. Barron, B. Mildenhall, M. Tancik, P. Hedman, R. Martin-Brualla, and P. P. Srinivasan, “Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields,” in IEEE ICCV , 2021, pp. 5835–5844
2021
Later among the works it cites.
C. Lin, W. Ma, A. Torralba, and S. Lucey, “BARF: bundle-adjusting neural radiance fields,” in IEEE ICCV , 2021, pp. 5721–5731
2021
Later among the works it cites.
M. Oechsle, S. Peng, and A. Geiger, “UNISURF: unifying neural implicit surfaces and radiance fields for multi-view reconstruction,” in IEEE ICCV , 2021, pp. 5569–5579
2021
Later among the works it cites.
N. Stier, A. Rich, P. Sen, and T. Höllerer, “Vortx: Volumetric 3d reconstruction with transformers for voxelwise view selection and fusion,” in 3DV , 2021, pp. 320–330
2021
Later among the works it cites.
A. Eftekhar, A. Sax, J. Malik, and A. R. Zamir, “Omnidata: A scalable pipeline for making multi-task mid-level vision datasets from 3d scans,” in IEEE ICCV , 2021, pp. 10 766–10 776
2021
Later among the works it cites.
Z. Zhu, S. Peng, V. Larsson, W. Xu, H. Bao, Z. Cui, M. R. Oswald, and M. Pollefeys, “Nice-slam: Neural implicit scalable encoding for slam,” in IEEE CVPR , 2022, pp. 12 786–12 796
2022
Closest in time.
Y. Li, F. Luo, and C. Xiao, “Self-supervised coarse-to-fine monocular depth estimation using a lightweight attention module,” Computational Visual Media , vol. 8, no. 4, pp. 631–647, 2022
2022
Closest in time.
2022
Closest in time.
J. Wang, P. Wang, X. Long, C. Theobalt, T. Komura, L. Liu, and W. Wang, “Neuris: Neural reconstruction of indoor scenes using normal priors,” in ECCV , 2022
2022
Closest in time.
H. Guo, S. Peng, H. Lin, Q. Wang, G. Zhang, H. Bao, and X. Zhou, “Neural 3d scene reconstruction with the manhattan-world assumption,” in IEEE CVPR , 2022, pp. 5511–5520
2022
Closest in time.
M. Sayed, J. Gibson, J. Watson, V. Prisacariu, M. Firman, and C. Godard, “Simplerecon: 3d reconstruction without 3d convolutions,” in ECCV , 2022
2022
Closest in time.
H. Chen, J. Huang, T.-J. Mu, and S.-M. Hu, “Circle: Convolutional implicit reconstruction and completion for large-scale indoor scene,” in ECCV , 2022
2022
Closest in time.
T. Müller, A. Evans, C. Schied, and A. Keller, “Instant neural graphics primitives with a multiresolution hash encoding,” ACM Transactions on Graphics , vol. 41, no. 4, pp. 102:1–102:15, 2022
2022
Closest in time.
X. Zhang, S. Bi, K. Sunkavalli, H. Su, and Z. Xu, “Nerfusion: Fusing radiance fields for large-scale scene reconstruction,” in IEEE CVPR , 2022, pp. 5449–5458
2022
Closest in time.