Fetching the paper…
Reading the bibliography…
Learning to understand dynamic 3D scenes from imagery is crucial for applications ranging from robotics to scene reconstruction.
Surface reconstruction from unorganized points
Hugues Hoppe, Tony DeRose, Tom Duchamp, John McDonald, and Werner Stuetzle · 1992
Earlier work this paper cites.
A volumetric method for building complex models from range images
Brian Curless and Marc Levoy · 1996
Earlier work this paper cites.
Depth discontinuities by pixel-to-pixel stereo
Stan Birchfield and Carlo Tomasi · 1999
Earlier work this paper cites.
Real-time correlation-based stereo vision with reduced border errors
Heiko Hirschmüller, Peter R Innocent, and Jon Garibaldi · 2002
Earlier work this paper cites.
A hierarchical symmetric stereo algorithm using dynamic programming
Geert Van Meerbergen, Maarten Vergauwen, Marc Pollefeys, and Luc Van Gool · 2002
Earlier work this paper cites.
Stereo matching using belief propagation
Jian Sun, Nan-Ning Zheng, and Heung-Yeung Shum · 2003
Earlier work this paper cites.
SIFT-the scale invariant feature transform
G Lowe · 2004
Earlier work this paper cites.
Visual modeling with a hand-held camera
Marc Pollefeys, Luc Van Gool, Maarten Vergauwen, Frank Verbiest, Kurt Cornelis, Jan Tops, and Reinhard Koch · 2004
Earlier work this paper cites.
Poisson surface reconstruction
Michael Kazhdan, Matthew Bolitho, and Hugues Hoppe · 2006
Earlier work this paper cites.
Segment-based stereo matching using belief propagation and a self-adapting dissimilarity measure
Andreas Klaus, Mario Sormann, and Konrad Karner · 2006
Earlier work this paper cites.
Photo tourism: exploring photo collections in 3D
Noah Snavely, Steven M Seitz, and Richard Szeliski · 2006
Earlier work this paper cites.
MonoSLAM: Real-time single camera SLAM
Andrew J Davison, Ian D Reid, Nicholas D Molton, and Olivier Stasse · 2007
Earlier work this paper cites.
Using multiple hypotheses to improve depth-maps for multi-view stereo
Neill DF Campbell, George Vogiatzis, Carlos Hernández, and Roberto Cipolla · 2008
Earlier work this paper cites.
Detailed real-time urban 3D reconstruction from video
Marc Pollefeys, David Nistér, J-M Frahm, Amir Akbarzadeh, Philippos Mordohai, Brian Clipp, Chris Engels, David Gallup, S-J Kim, Paul Merrell, et al · 2008
Earlier work this paper cites.
Accurate, dense, and robust multiview stereopsis
Yasutaka Furukawa and Jean Ponce · 2009
Earlier work this paper cites.
Towards internet-scale multi-view stereo
Yasutaka Furukawa, Brian Curless, Steven M Seitz, and Richard Szeliski · 2010
Earlier work this paper cites.
3D reconstruction of a moving point from a series of 2D projections
Hyun Soo Park, Takaaki Shiratori, Iain Matthews, and Yaser Sheikh · 2010
Earlier work this paper cites.
Building Rome in a day
Sameer Agarwal, Yasutaka Furukawa, Noah Snavely, Ian Simon, Brian Curless, Steven M Seitz, and Richard Szeliski · 2011
Earlier work this paper cites.
Multi-view reconstruction preserving weakly-supported surfaces
Michal Jancosek and Tomas Pajdla · 2011
Earlier work this paper cites.
A naturalistic open source movie for optical flow evaluation
Daniel J Butler, Jonas Wulff, Garrett B Stanley, and Michael J Black · 2012
Earlier work this paper cites.
Vision meets robotics: The KITTI dataset
Andreas Geiger, Philip Lenz, Christoph Stiller, and Raquel Urtasun · 2013
Earlier work this paper cites.
FlowNet: Learning optical flow with convolutional networks
Alexey Dosovitskiy, Philipp Fischer, Eddy Ilg, Philip Hausser, Caner Hazirbas, Vladimir Golkov, Patrick Van Der Smagt, Daniel Cremers, and Thomas Brox · 2015
Earlier work this paper cites.
Massively parallel multiview stereopsis by surface normal diffusion
Silvano Galliani, Katrin Lasinger, and Konrad Schindler · 2015
Earlier work this paper cites.
Panoptic studio: A massively multiview system for social motion capture
Hanbyul Joo, Hao Liu, Lei Tan, Lin Gui, Bart Nabbe, Iain Matthews, Takeo Kanade, Shohei Nobuhara, and Yaser Sheikh · 2015
Earlier work this paper cites.
ORB-SLAM: a versatile and accurate monocular SLAM system
Raúl Mur-Artal, J. M. M. Montiel, and Juan D. Tardós · 2015
Earlier work this paper cites.
DynamicFusion: Reconstruction and tracking of non-rigid scenes in real-time
Richard A Newcombe, Dieter Fox, and Steven M Seitz · 2015
Earlier work this paper cites.
A large dataset to train convolutional networks for disparity, optical flow, and scene flow estimation
Nikolaus Mayer, Eddy Ilg, Philip Hausser, Philipp Fischer, Daniel Cremers, Alexey Dosovitskiy, and Thomas Brox · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L Schonberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Pixelwise view selection for unstructured multi-view stereo
Johannes L Schönberger, Enliang Zheng, Jan-Michael Frahm, and Marc Pollefeys · 2016
Earlier work this paper cites.
Kronecker-Markov prior for dynamic 3D reconstruction
Tomas Simon, Jack Valmadre, Iain Matthews, and Yaser Sheikh · 2016
Earlier work this paper cites.
Spatiotemporal bundle adjustment for dynamic 3D reconstruction
Minh Vo, Srinivasa G Narasimhan, and Yaser Sheikh · 2016
Earlier work this paper cites.
Rethinking atrous convolution for semantic image segmentation
Liang-Chieh Chen · 2017
Earlier work this paper cites.
Direct sparse odometry
Jakob Engel, Vladlen Koltun, and Daniel Cremers · 2017
Earlier work this paper cites.
End-to-end learning of geometry and context for deep stereo regression
Alex Kendall, Hayk Martirosyan, Saumitro Dasgupta, Peter Henry, Ryan Kennedy, Abraham Bachrach, and Adam Bry · 2017
Earlier work this paper cites.
Cascade residual learning: A two-stage convolutional neural network for stereo matching
Jiahao Pang, Wenxiu Sun, Jimmy SJ Ren, Chengxi Yang, and Qiong Yan · 2017
Earlier work this paper cites.
Attention is all you need
A Vaswani · 2017
Cited alongside, same era.
Scene parsing through ADE20K dataset
Bolei Zhou, Hang Zhao, Xavier Puig, Sanja Fidler, Adela Barriuso, and Antonio Torralba · 2017
Cited alongside, same era.
CodeSLAM—learning a compact, optimisable representation for dense visual SLAM
Michael Bloesch, Jan Czarnowski, Ronald Clark, Stefan Leutenegger, and Andrew J Davison · 2018
Cited alongside, same era.
Pyramid stereo matching network
Jia-Ren Chang and Yong-Sheng Chen · 2018
Cited alongside, same era.
MegaDepth: Learning single-view depth prediction from internet photos
Zhengqi Li and Noah Snavely · 2018
Cited alongside, same era.
BA-Net: Dense bundle adjustment network
Chengzhou Tang and Ping Tan · 2018
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Later among the works it cites.
Accelerated coordinate encoding: Learning to relocalize in minutes using RGB and poses
Eric Brachmann, Tommaso Cavallari, and Victor Adrian Prisacariu · 2023
Later among the works it cites.
HumanRF: High-fidelity neural radiance fields for humans in motion
Mustafa Işık, Martin Rünz, Markos Georgopoulos, Taras Khakhulin, Jonathan Starck, Lourdes Agapito, and Matthias Nießner · 2023
Later among the works it cites.
DynamicStereo: Consistent dynamic depth from stereo videos
Nikita Karaev, Ignacio Rocco, Benjamin Graham, Natalia Neverova, Andrea Vedaldi, and Christian Rupprecht · 2023
Later among the works it cites.
NeRSemble: Multi-view radiance field reconstruction of human heads
Tobias Kirschstein, Shenhan Qian, Simon Giebenhain, Tim Walter, and Matthias Nießner · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
MVSNet: Depth inference for unstructured multi-view stereo
Yao Yao, Zixin Luo, Shiwei Li, Tian Fang, and Long Quan · 2018
Cited alongside, same era.
Stereo magnification: Learning view synthesis using multiplane images
Tinghui Zhou, Richard Tucker, John Flynn, Graham Fyffe, and Noah Snavely · 2018
Cited alongside, same era.
Learning the depths of moving people by watching frozen people
Zhengqi Li, Tali Dekel, Forrester Cole, Richard Tucker, Noah Snavely, Ce Liu, and William T Freeman · 2019
Cited alongside, same era.
ReFusion: 3D Reconstruction in Dynamic Environments for RGB-D Cameras Exploiting Residuals
E. Palazzolo, J. Behley, P. Lottes, P. Giguère, and C. Stachniss · 2019
Cited alongside, same era.
Structure from motion for panorama-style videos
Chris Sweeney, Aleksander Holynski, Brian Curless, and Steve M Seitz · 2019
Cited alongside, same era.
Recurrent MVSNet for high-resolution multi-view stereo depth inference
Yao Yao, Zixin Luo, Shiwei Li, Tianwei Shen, Tian Fang, and Long Quan · 2019
Cited alongside, same era.
Robust dynamic radiance fields
Yu-Lun Liu, Chen Gao, Andreas Meuleman, Hung-Yu Tseng, Ayush Saraf, Changil Kim, Yung-Yu Chuang, Johannes Kopf, and Jia-Bin Huang · 2023
Later among the works it cites.
Aria digital twin: A new benchmark dataset for egocentric 3d machine perception
Xiaqing Pan, Nicholas Charron, Yongqian Yang, Scott Peters, Thomas Whelan, Chen Kong, Omkar Parkhi, Richard Newcombe, and Yuheng Carl Ren · 2023
Later among the works it cites.
CamP: Camera preconditioning for neural radiance fields
Keunhong Park, Philipp Henzler, Ben Mildenhall, Jonathan T Barron, and Ricardo Martin-Brualla · 2023
Later among the works it cites.
DytanVO: Joint refinement of visual odometry and motion segmentation in dynamic environments
Shihao Shen, Yilin Cai, Wenshan Wang, and Sebastian Scherer · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, Katie Millican, et al · 2023
Later among the works it cites.
Metric3D: Towards zero-shot metric 3D prediction from a single image
Wei Yin, Chi Zhang, Hao Chen, Zhipeng Cai, Gang Yu, Kaixuan Wang, Xiaozhi Chen, and Chunhua Shen · 2023
Later among the works it cites.
TemporalStereo: Efficient spatial-temporal stereo matching network
Youmin Zhang, Matteo Poggi, and Stefano Mattoccia · 2023
Later among the works it cites.
PointOdyssey: A large-scale synthetic dataset for long-term point tracking
Yang Zheng, Adam W. Harley, Bokui Shen, Gordon Wetzstein, and Leonidas J. Guibas · 2023
Later among the works it cites.
Depth Pro: Sharp monocular metric depth in less than a second
Aleksei Bochkovskii, Amaël Delaunoy, Hugo Germain, Marcel Santos, Yichao Zhou, Stephan R Richter, and Vladlen Koltun · 2024
Closest in time.
Scene coordinate reconstruction: Posing of image collections via incremental learning of a relocalizer
Eric Brachmann, Jamie Wynn, Shuai Chen, Tommaso Cavallari, Áron Monszpart, Daniyar Turmukhambetov, and Victor Adrian Prisacariu · 2024
Closest in time.
BootsTAP: Bootstrapped training for tracking any point
Carl Doersch, Pauline Luc, Yi Yang, Dilara Gokay, Skanda Koppula, Ankush Gupta, Joseph Heyward, Ignacio Rocco, Ross Goroshin, João Carreira, and Andrew Zisserman · 2024
Closest in time.
COLMAP-Free 3D gaussian splatting
Yang Fu, Sifei Liu, Amey Kulkarni, Jan Kautz, Alexei A. Efros, and Xiaolong Wang · 2024
Closest in time.
CAT3D: Create anything in 3D with multi-view diffusion models
Ruiqi Gao, Aleksander Holynski, Philipp Henzler, Arthur Brussee, Ricardo Martin-Brualla, Pratul Srinivasan, Jonathan T Barron, and Ben Poole · 2024
Closest in time.
DepthCrafter: Generating consistent long depth sequences for open-world videos
Wenbo Hu, Xiangjun Gao, Xiaoyu Li, Sijie Zhao, Xiaodong Cun, Yong Zhang, Long Quan, and Ying Shan · 2024
Closest in time.
Repurposing diffusion-based image generators for monocular depth estimation
Bingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger, Rodrigo Caye Daudt, and Konrad Schindler · 2024
Closest in time.
TAPVid-3D: A benchmark for tracking any point in 3D
Skanda Koppula, Ignacio Rocco, Yi Yang, Joe Heyward, João Carreira, Andrew Zisserman, Gabriel Brostow, and Carl Doersch · 2024
Closest in time.
MoSca: Dynamic gaussian fusion from casual videos via 4D motion scaffolds
Jiahui Lei, Yijia Weng, Adam Harley, Leonidas Guibas, and Kostas Daniilidis · 2024
Closest in time.
Grounding image matching in 3D with MASt3R
Vincent Leroy, Yohann Cabon, and Jérôme Revaud · 2024
Closest in time.
MegaSaM: Accurate, fast, and robust structure and motion from casual dynamic videos
Zhengqi Li, Richard Tucker, Forrester Cole, Qianqian Wang, Linyi Jin, Vickie Ye, Angjoo Kanazawa, Aleksander Holynski, and Noah Snavely · 2024
Closest in time.
UniDepth: Universal monocular metric depth estimation
Luigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segu, Siyuan Li, Luc Van Gool, and Fisher Yu · 2024
Closest in time.
Movie Gen: A cast of media foundation models
Adam Polyak, Amit Zohar, Andrew Brown, Andros Tjandra, Animesh Sinha, Ann Lee, Apoorv Vyas, Bowen Shi, Chih-Yao Ma, Ching-Yao Chuang, et al · 2024
Closest in time.
Learning temporally consistent video depth from video diffusion priors
Jiahao Shao, Yuanbo Yang, Hongyu Zhou, Youmin Zhang, Yujun Shen, Matteo Poggi, and Yiyi Liao · 2024
Closest in time.
ExtraNeRF: Visibility-aware view extrapolation of neural radiance fields with diffusion models
Meng-Li Shih, Wei-Chiu Ma, Lorenzo Boyice, Aleksander Holynski, Forrester Cole, Brian Curless, and Janne Kontkanen · 2024
Closest in time.
Deep patch visual odometry
Zachary Teed, Lahav Lipson, and Jia Deng · 2024
Closest in time.
Nerfiller: Completing scenes via generative 3D inpainting
Ethan Weber, Aleksander Holynski, Varun Jampani, Saurabh Saxena, Noah Snavely, Abhishek Kar, and Angjoo Kanazawa · 2024
Closest in time.
CAT4D: Create anything in 4D with multi-view video diffusion models
Rundi Wu, Ruiqi Gao, Ben Poole, Alex Trevithick, Changxi Zheng, Jonathan T Barron, and Aleksander Holynski · 2024
Closest in time.
Match-stereo-videos: Bidirectional alignment for consistent dynamic stereo matching
Junpeng Jing, Ye Mao, and Krystian Mikolajczyk · 2025
Closest in time.
SEA-RAFT: Simple, efficient, accurate RAFT for optical flow
Yihan Wang, Lahav Lipson, and Jia Deng · 2025
Closest in time.
MonST3R: A simple approach for estimating geometry in the presence of motion
Junyi Zhang, Charles Herrmann, Junhwa Hur, Varun Jampani, Trevor Darrell, Forrester Cole, Deqing Sun, and Ming-Hsuan Yang · 2025
Closest in time.