Fetching the paper…
Reading the bibliography…
We introduce a novel monocular visual odometry (VO) system, NeRF-VO, that integrates learning-based sparse visual odometry for low-latency camera tracking and a neural radiance scene representation for fine-detailed dense reconstruction and novel view synthesis.
W. Kabsch, “A discussion of the solution for the best rotation to relate two sets of vectors,” Acta Crystallographica Section A , 1978
1978
Earlier work this paper cites.
S. Umeyama, “Least-squares estimation of transformation parameters between two point patterns,” TPAMI , 1991
1991
Earlier work this paper cites.
Z. Wang, A. Bovik, H. Sheikh, and E. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” IEEE Transactions on Image Processing , 2004
2004
Earlier work this paper cites.
A. J. Davison, I. D. Reid, N. D. Molton, and O. Stasse, “Monoslam: Real-time single camera slam,” TPAMI , 2007
2007
Earlier work this paper cites.
R. A. Newcombe, S. J. Lovegrove, and A. J. Davison, “DTAM: Dense tracking and mapping in real-time,” in ICCV , 2011
2011
Earlier work this paper cites.
R. A. Newcombe, S. Izadi, O. Hilliges, D. Molyneaux, D. Kim, A. J. Davison, P. Kohi, J. Shotton, S. Hodges, and A. Fitzgibbon, “KinectFusion: Real-time dense surface mapping and tracking,” in ISMAR , 2011
2011
Earlier work this paper cites.
J. Sturm, N. Engelhard, F. Endres, W. Burgard, and D. Cremers, “A benchmark for the evaluation of rgb-d slam systems,” in IROS , 2012
2012
Earlier work this paper cites.
J. Shotton, B. Glocker, C. Zach, S. Izadi, A. Criminisi, and A. Fitzgibbon, “Scene coordinate regression forests for camera relocalization in rgb-d images,” in CVPR , 2013
2013
Earlier work this paper cites.
J. Engel, T. Schöps, and D. Cremers, “Lsd-slam: Large-scale direct monocular slam,” in ECCV , 2014
2014
Earlier work this paper cites.
K. Tateno, F. Tombari, I. Laina, and N. Navab, “CNN-SLAM: Real-time dense monocular slam with learned depth prediction,” in CVPR , 2017
2017
Earlier work this paper cites.
B. Ummenhofer, H. Zhou, J. Uhrig, N. Mayer, E. Ilg, A. Dosovitskiy, and T. Brox, “DeMoN: Depth and motion network for learning monocular stereo,” in CVPR , 2017
2017
Earlier work this paper cites.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Niessner, “ScanNet: Richly-annotated 3d reconstructions of indoor scenes,” in CVPR , 2017
2017
Earlier work this paper cites.
A. Dai, M. Nießner, M. Zollhöfer, S. Izadi, and C. Theobalt, “BundleFusion: Real-time globally consistent 3d reconstruction using on-the-fly surface reintegration,” ACM Trans. Graph. , 2017
2017
Earlier work this paper cites.
J. Engel, V. Koltun, and D. Cremers, “Direct sparse odometry,” TPAMI , 2018
2018
Earlier work this paper cites.
H. Zhou, B. Ummenhofer, and T. Brox, “DeepTAM: Deep tracking and mapping,” in ECCV , 2018
2018
Earlier work this paper cites.
M. Bloesch, J. Czarnowski, R. Clark, S. Leutenegger, and A. J. Davison, “CodeSLAM — learning a compact, optimisable representation for dense visual slam,” in CVPR , 2018
2018
Earlier work this paper cites.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The unreasonable effectiveness of deep features as a perceptual metric,” in CVPR , 2018
2018
Earlier work this paper cites.
J. Straub, T. Whelan, L. Ma, Y. Chen, E. Wijmans, S. Green, J. J. Engel, R. Mur-Artal, C. Ren, S. Verma, A. Clarkson, M. Yan, B. Budge, Y. Yan, X. Pan, J. Yon, Y. Zou, K. Leon, N. Carter, J. Briales, T. Gillingham, E. Mueggler, L. Pesqueira, M. Savva, D. Batra, H. M. Strasdat, R. D. Nardi, M. Goesele, S. Lovegrove, and R. Newcombe, “The Replica dataset: A digital replica of indoor spaces,” 2019
2019
Earlier work this paper cites.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “Nerf: Representing scenes as neural radiance fields for view synthesis,” in ECCV , 2020
2020
Earlier work this paper cites.
J. Zubizarreta, I. Aguinaga, and J. M. M. Montiel, “Direct sparse mapping,” IEEE Transactions on Robotics , 2020
2020
Earlier work this paper cites.
N. Yang, L. v. Stumberg, R. Wang, and D. Cremers, “D3VO: Deep depth, deep pose and deep uncertainty for monocular visual odometry,” in CVPR , 2020
2020
Earlier work this paper cites.
J. Czarnowski, T. Laidlow, R. Clark, and A. J. Davison, “DeepFactors: Real-time probabilistic dense monocular slam,” IEEE Robotics and Automation Letters , 2020
2020
Cited alongside, same era.
E. Sucar, S. Liu, J. Ortiz, and A. J. Davison, “iMAP: Implicit mapping and positioning in real-time,” in ICCV , 2021
2021
Cited alongside, same era.
C. Campos, R. Elvira, J. J. G. Rodríguez, J. M. M. Montiel, and J. D. Tardós, “Orb-slam3: An accurate open-source library for visual, visual–inertial, and multimap slam,” IEEE Transactions on Robotics , 2021
2021
Cited alongside, same era.
W. Wang, Y. Hu, and S. Scherer, “TartanVO: A generalizable learning-based vo,” in CoRL , 2021
2021
Cited alongside, same era.
Z. Teed and J. Deng, “DROID-SLAM: Deep visual slam for monocular, stereo, and rgb-d cameras,” in Advances in Neural Information Processing Systems , 2021
2021
H. Wang, J. Wang, and L. Agapito, “Co-SLAM: Joint coordinate and sparse parametric encodings for neural real-time slam,” in CVPR , 2023
2023
Closest in time.
K. Mazur, E. Sucar, and A. J. Davison, “Feature-realistic neural fusion for real-time, open set scene understanding,” in ICRA , 2023
2023
Closest in time.
A. Rosinol, J. J. Leonard, and L. Carlone, “Nerf-slam: Real-time dense monocular slam with neural radiance fields,” in IROS , 2023
2023
Closest in time.
H. Li, X. Gu, W. Yuan, L. Yang, Z. Dong, and P. Tan, “Dense rgb slam with neural implicit maps,” in ICLR , 2023
2023
Closest in time.
C.-M. Chung, Y.-C. Tseng, Y.-C. Hsu, X.-Q. Shi, Y.-H. Hua, J.-F. Yeh, W.-C. Chen, Y.-T. Chen, and W. H. Hsu, “Orbeez-slam: A real-time monocular visual slam with orb features and nerf-realized mapping,” in ICRA , 2023
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
X. Zuo, N. Merrill, W. Li, Y. Liu, M. Pollefeys, and G. Huang, “CodeVIO: Visual-inertial odometry with learned optimizable dense depth,” in ICRA , 2021
2021
Cited alongside, same era.
L. Koestler, N. Yang, N. Zeller, and D. Cremers, “TANDEM: tracking and dense mapping in real-time using deep multi-view stereo,” in CoRL , 2021
2021
Cited alongside, same era.
A. Eftekhar, A. Sax, J. Malik, and A. Zamir, “Omnidata: A scalable pipeline for making multi-task mid-level vision datasets from 3d scans,” in ICCV , 2021
2021
Cited alongside, same era.
R. Ranftl, A. Bochkovskiy, and V. Koltun, “Vision transformers for dense prediction,” in ICCV , 2021
2021
Cited alongside, same era.
A. Chen, Z. Xu, A. Geiger, J. Yu, and H. Su, “TensoRF: Tensorial radiance fields,” in ECCV , 2022
2022
Cited alongside, same era.
S. Fridovich-Keil, A. Yu, M. Tancik, Q. Chen, B. Recht, and A. Kanazawa, “Plenoxels: Radiance fields without neural networks,” in CVPR , 2022
2022
Cited alongside, same era.
T. Müller, A. Evans, C. Schied, and A. Keller, “Instant neural graphics primitives with a multiresolution hash encoding,” ACM Trans. Graph. , 2022
2022
Cited alongside, same era.
Closest in time.
Y. Zhang, F. Tosi, S. Mattoccia, and M. Poggi, “GO-SLAM: Global optimization for consistent 3d instant reconstruction,” in ICCV , 2023
2023
Closest in time.
Z. Teed, L. Lipson, and J. Deng, “Deep patch visual odometry,” Advances in Neural Information Processing Systems , 2023
2023
Closest in time.
Y. Xin, X. Zuo, D. Lu, and S. Leutenegger, “SimpleMapping: Real-Time Visual-Inertial Dense Mapping with Deep Multi-View Stereo,” ISMAR , 2023
2023
Closest in time.
D. Lisus, C. Holmes, and S. Waslander, “Towards open world nerf-based slam,” in CRV , 2023
2023
Closest in time.
B. Kerbl, G. Kopanas, T. Leimkuehler, and G. Drettakis, “3d gaussian splatting for real-time radiance field rendering,” ACM Trans. Graph. , 2023
2023
Closest in time.
M. Tancik, E. Weber, E. Ng, R. Li, B. Yi, T. Wang, A. Kristoffersen, J. Austin, K. Salahi, A. Ahuja, D. Mcallister, J. Kerr, and A. Kanazawa, “Nerfstudio: A modular framework for neural radiance field development,” in SIGGRAPH , 2023
2023
Closest in time.
D. Wofk, R. Ranftl, M. Müller, and V. Koltun, “Monocular visual-inertial depth estimation,” in ICRA , 2023
2023
Closest in time.
Z. Zhu, S. Peng, V. Larsson, Z. Cui, M. R. Oswald, A. Geiger, and M. Pollefeys, “Nicer-slam: Neural implicit scene encoding for rgb slam,” in 3DV , 2024
2024
Closest in time.
W. Zhang, T. Sun, S. Wang, Q. Cheng, and N. Haala, “HI-SLAM: Monocular real-time dense mapping with hybrid implicit fields,” in IEEE Robotics and Automation Letters , 2024
2024
Closest in time.
F. Tosi, Y. Zhang, Z. Gong, E. Sandström, S. Mattoccia, M. R. Oswald, and M. Poggi, “How nerfs and 3d gaussian splatting are reshaping slam: a survey,” 2024
2024
Closest in time.
H. Matsuki, K. Tateno, M. Niemeyer, and F. Tombari, “NEWTON: Neural view-centric mapping for on-the-fly large-scale slam,” IEEE Robotics and Automation Letters , 2024
2024
Closest in time.
H. Matsuki, R. Murai, P. H. J. Kelly, and A. J. Davison, “Gaussian Splatting SLAM,” in CVPR , 2024
2024
Closest in time.
C. Yan, D. Qu, D. Wang, D. Xu, Z. Wang, B. Zhao, and X. Li, “Gs-slam: Dense visual slam with 3d gaussian splatting,” in CVPR , 2024
2024
Closest in time.
N. Keetha, J. Karhade, K. M. Jatavallabhula, G. Yang, S. Scherer, D. Ramanan, and J. Luiten, “Splatam: Splat, track & map 3d gaussians for dense rgb-d slam,” in CVPR , 2024
2024
Closest in time.
H. Huang, L. Li, C. Hui, and S.-K. Yeung, “Photo-slam: Real-time simultaneous localization and photorealistic mapping for monocular, stereo, and rgb-d cameras,” in CVPR , 2024
2024
Closest in time.