Fetching the paper…
Reading the bibliography…
Recent advances in DUSt3R have enabled robust estimation of dense point clouds and camera parameters of static scenes, leveraging Transformer network architectures and direct supervision on large-scale 3D datasets.
A threshold selection method from gray-level histograms
Nobuyuki Otsu et al · 1975
Earlier work this paper cites.
Least-squares estimation of transformation parameters between two point patterns
Shinji Umeyama · 1991
Earlier work this paper cites.
The fundamental matrix: Theory, algorithms, and stability analysis
Quan-Tuan Luong and Olivier D Faugeras · 1996
Earlier work this paper cites.
Attentional modulation of visual motion processing in cortical areas mt and mst
Stefan Treue and John HR Maunsell · 1996
Earlier work this paper cites.
Self-calibration and metric reconstruction inspite of varying and unknown intrinsic camera parameters
Marc Pollefeys, Reinhard Koch, and Luc Van Gool · 1999
Earlier work this paper cites.
Bundle adjustment—a modern synthesis
Bill Triggs, Philip F McLauchlan, Richard I Hartley, and Andrew W Fitzgibbon · 2000
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
David G Lowe · 2004
Earlier work this paper cites.
Visual modeling with a hand-held camera
Marc Pollefeys, Luc Van Gool, Maarten Vergauwen, Frank Verbiest, Kurt Cornelis, Jan Tops, and Reinhard Koch · 2004
Earlier work this paper cites.
Photo tourism: exploring photo collections in 3d
Noah Snavely, Steven M Seitz, and Richard Szeliski · 2006
Earlier work this paper cites.
A general framework for motion segmentation: Independent, articulated, rigid, non-rigid, degenerate and non-degenerate
Jingyu Yan and Marc Pollefeys · 2006
Earlier work this paper cites.
Monoslam: Real-time single camera slam
Andrew J Davison, Ian D Reid, Nicholas D Molton, and Olivier Stasse · 2007
Earlier work this paper cites.
Speeded-up robust features (surf)
Herbert Bay, Andreas Ess, Tinne Tuytelaars, and Luc Van Gool · 2008
Earlier work this paper cites.
Modeling the world from internet photo collections
Noah Snavely, Steven M Seitz, and Richard Szeliski · 2008
Earlier work this paper cites.
Background subtraction for freely moving cameras
Yaser Sheikh, Omar Javed, and Takeo Kanade · 2009
Earlier work this paper cites.
Bundle adjustment in the large
Sameer Agarwal, Noah Snavely, Steven M Seitz, and Richard Szeliski · 2010
Earlier work this paper cites.
Object segmentation by long term analysis of point trajectories
Thomas Brox and Jitendra Malik · 2010
Earlier work this paper cites.
Building rome in a day
Sameer Agarwal, Yasutaka Furukawa, Noah Snavely, Ian Simon, Brian Curless, Steven M Seitz, and Richard Szeliski · 2011
Earlier work this paper cites.
Dtam: Dense tracking and mapping in real-time
Richard A Newcombe, Steven J Lovegrove, and Andrew J Davison · 2011
Earlier work this paper cites.
A benchmark for the evaluation of rgb-d slam systems
Jürgen Sturm, Nikolas Engelhard, Felix Endres, Wolfram Burgard, and Daniel Cremers · 2012
Earlier work this paper cites.
Segmentation of moving objects by long term video analysis
Peter Ochs, Jitendra Malik, and Thomas Brox · 2013
Earlier work this paper cites.
Lsd-slam: Large-scale direct monocular slam
Jakob Engel, Thomas Schöps, and Daniel Cremers · 2014
Earlier work this paper cites.
Orb-slam: a versatile and accurate monocular slam system
Raul Mur-Artal, Jose Maria Martinez Montiel, and Juan D Tardos · 2015
Earlier work this paper cites.
Large-scale data for multiple-view stereopsis
Henrik Aanæs, Rasmus Ramsbøl Jensen, George Vogiatzis, Engin Tola, and Anders Bjorholm Dahl · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Direct sparse odometry
Jakob Engel, Vladlen Koltun, and Daniel Cremers · 2017
Earlier work this paper cites.
Mask R-CNN
Kaiming He, Georgia Gkioxari, Piotr Dollar, and Ross Girshick · 2017
Cited alongside, same era.
The 2017 davis challenge on video object segmentation
Jordi Pont-Tuset, Federico Perazzi, Sergi Caelles, Pablo Arbeláez, Alex Sorkine-Hornung, and Luc Van Gool · 2017
Cited alongside, same era.
Superpoint: Self-supervised interest point detection and description
Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2018
Cited alongside, same era.
Ba-net: Dense bundle adjustment network
Chengzhou Tang and Ping Tan · 2018
Cited alongside, same era.
DS-SLAM: A semantic visual SLAM towards dynamic environments
Chao Yu, Zuxin Liu, Xin-Jun Liu, Fugui Xie, Yi Yang, Qi Wei, and Fei Qiao · 2018
Cited alongside, same era.
Cross-view completion models are zero-shot correspondence estimators
Honggyu An, Jinhyeon Kim, Seonghoon Park, Jaewoo Jung, Jisang Han, Sunghwan Hong, and Seungryong Kim · 2024
Later among the works it cites.
Scene coordinate reconstruction: Posing of image collections via incremental learning of a relocalizer
Eric Brachmann, Jamie Wynn, Shuai Chen, Tommaso Cavallari, Áron Monszpart, Daniyar Turmukhambetov, and Victor Adrian Prisacariu · 2024
Later among the works it cites.
Feat2gs: Probing visual foundation models with gaussian splatting
Yue Chen, Xingyu Chen, Anpei Chen, Gerard Pons-Moll, and Yuliang Xiu · 2024
Later among the works it cites.
Romo: Robust motion segmentation improves structure from motion
Lily Goli, Sara Sabour, Mark Matthews, Marcus Brubaker, Dmitry Lagun, Alec Jacobson, David J Fleet, Saurabh Saxena, and Andrea Tagliasacchi · 2024
Later among the works it cites.
Depthcrafter: Generating consistent long depth sequences for open-world videos
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
Consistent video depth estimation
Xuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen, and Johannes Kopf · 2020
Cited alongside, same era.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
René Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun · 2020
Cited alongside, same era.
Superglue: Learning feature matching with graph neural networks
Paul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2020
Cited alongside, same era.
Scalability in perception for autonomous driving: Waymo open dataset
Pei Sun, Henrik Kretzschmar, Xerxes Dotiwalla, Aurelien Chouard, Vijaysai Patnaik, Paul Tsui, James Guo, Yin Zhou, Yuning Chai, Benjamin Caine, et al · 2020
Cited alongside, same era.
Tartanair: A dataset to push the limits of visual slam
Wenshan Wang, Delong Zhu, Xiangwei Wang, Yaoyu Hu, Yuheng Qiu, Chen Wang, Yafei Hu, Ashish Kapoor, and Sebastian Scherer · 2020
Cited alongside, same era.
Robust consistent video depth estimation
Johannes Kopf, Xuejian Rong, and Jia-Bin Huang · 2021
Cited alongside, same era.
Wenbo Hu, Xiangjun Gao, Xiaoyu Li, Sijie Zhao, Xiaodong Cun, Yong Zhang, Long Quan, and Ying Shan · 2024
Later among the works it cites.
Stereo4d: Learning how things move in 3d from internet stereo videos
Linyi Jin, Richard Tucker, Zhengqi Li, David Fouhey, Noah Snavely, and Aleksander Holynski · 2024
Later among the works it cites.
Tapvid-3d: A benchmark for tracking any point in 3d, 2024
Skanda Koppula, Ignacio Rocco, Yi Yang, Joe Heyward, Joao Carreira, Andrew Zisserman, Gabriel Brostow, and Carl Doersch · 2024
Later among the works it cites.
Grounding image matching in 3d with MASt3R
Vincent Leroy, Yohann Cabon, and Jerome Revaud · 2024
Later among the works it cites.
MegaSaM: accurate, fast, and robust structure and motion from casual dynamic videos
Zhengqi Li, Richard Tucker, Forrester Cole, Qianqian Wang, Linyi Jin, Vickie Ye, Angjoo Kanazawa, Aleksander Holynski, and Noah Snavely · 2024
Later among the works it cites.
SLAM3R: real-time dense scene reconstruction from monocular RGB videos
Yuzheng Liu, Siyan Dong, Shuzhe Wang, Yanchao Yang, Qingnan Fan, and Baoquan Chen · 2024
Later among the works it cites.
MASt3R-SLAM: real-time dense SLAM with 3d reconstruction priors
Riku Murai, Eric Dexheimer, and Andrew J. Davison · 2024
Later among the works it cites.
Global structure-from-motion revisited
Linfei Pan, Dániel Baráth, Marc Pollefeys, and Johannes Lutz Schönberger · 2024
Later among the works it cites.
UniDepth: universal monocular metric depth estimation
Luigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segu, Siyuan Li, Luc Van Gool, and Fisher Yu · 2024
Later among the works it cites.
Sam 2: Segment anything in images and videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloe Rolland, Laura Gustafson, et al · 2024
Later among the works it cites.
Learning temporally consistent video depth from video diffusion priors
Jiahao Shao, Yuanbo Yang, Hongyu Zhou, Youmin Zhang, Yujun Shen, Vitor Guizilini, Yue Wang, Matteo Poggi, and Yiyi Liao · 2024
Later among the works it cites.
MV-DUSt3R+: single-stage scene reconstruction from sparse views in 2 seconds
Zhenggang Tang, Yuchen Fan, Dilin Wang, Hongyu Xu, Rakesh Ranjan, Alexander Schwing, and Zhicheng Yan · 2024
Later among the works it cites.
3d reconstruction with spatial memory
Hengyi Wang and Lourdes Agapito · 2024
Later among the works it cites.
Vggsfm: Visual geometry grounded deep structure from motion
Jianyuan Wang, Nikita Karaev, Christian Rupprecht, and David Novotny · 2024
Later among the works it cites.
DUSt3R: geometric 3d vision made easy
Shuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii, and Jerome Revaud · 2024
Later among the works it cites.
Sea-raft: Simple, efficient, accurate raft for optical flow
Yihan Wang, Lahav Lipson, and Jia Deng · 2024
Later among the works it cites.
DAS3R: dynamics-aware gaussian splatting for static scene reconstruction
Kai Xu, Tze Ho Elden Tse, Jizong Peng, and Angela Yao · 2024
Later among the works it cites.
Depth anything: Unleashing the power of large-scale unlabeled data
Lihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu, Jiashi Feng, and Hengshuang Zhao · 2024
Later among the works it cites.
MonST3R: a simple approach for estimating geometry in the presence of motion
Junyi Zhang, Charles Herrmann, Junhwa Hur, Varun Jampani, Trevor Darrell, Forrester Cole, Deqing Sun, and Ming-Hsuan Yang · 2024
Later among the works it cites.
Learning segmentation from point trajectories
Laurynas Karazija, Iro Laina, Christian Rupprecht, and Andrea Vedaldi · 2025
Closest in time.
Continuous 3d perception model with persistent state
Qianqian Wang, Yifei Zhang, Aleksander Holynski, Alexei A. Efros, and Angjoo Kanazawa · 2025
Closest in time.