Fetching the paper…
Reading the bibliography…
Structure-from-motion (SfM) is a long-standing problem in the computer vision community, which aims to reconstruct the camera poses and 3D structure of a scene from a set of unconstrained 2D images.
Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography
Martin A Fischler and Robert C Bolles · 1981
Earlier work this paper cites.
Deterministic edge-preserving regularization in computed imaging
Pierre Charbonnier, Laure Blanc-féraud, Gilles Aubert, and Michel Barlaud · 1997
Earlier work this paper cites.
Object Recognition from Local Scale-Invariant Features
David G. Lowe · 1999
Earlier work this paper cites.
Multiple View Geometry in Computer Vision
Richard Hartley and Andrew Zisserman · 2000
Earlier work this paper cites.
A critique of structure-from-motion algorithms
John Oliensis · 2000
Earlier work this paper cites.
Bundle Adjustment - A Modern Synthesis
Bill Triggs, Philip F. McLauchlan, Richard I. Hartley, and Andrew W. Fitzgibbon · 2000
Earlier work this paper cites.
Robust Wide Baseline Stereo from Maximally Stable Extremal Regions
Jiri Matas, Ondrej Chum, Martin Urban, and Tomás Pajdla · 2002
Earlier work this paper cites.
Multi-view Matching for Unordered Image Sets, or ”How Do I Organize My Holiday Snaps?”
Frederik Schaffalitzky and Andrew Zisserman · 2002
Earlier work this paper cites.
Linear multiview reconstruction of points, lines, planes and cameras using a reference plane
Rother · 2003
Earlier work this paper cites.
Distinctive Image Features from Scale-Invariant Keypoints
David G. Lowe · 2004
Earlier work this paper cites.
The levenberg-marquardt algorithm: implementation and theory
Jorge J Moré · 2006
Earlier work this paper cites.
Photo tourism: exploring photo collections in 3d
Noah Snavely, Steven M Seitz, and Richard Szeliski · 2006
Earlier work this paper cites.
Speeded-Up Robust Features (SURF)
Herbert Bay, Andreas Ess, Tinne Tuytelaars, and Luc Van Gool · 2008
Earlier work this paper cites.
Particle video: Long-range motion estimation using point trajectories
Peter Sand and Seth Teller · 2008
Earlier work this paper cites.
Bundle adjustment in the large
Sameer Agarwal, Noah Snavely, Steven M Seitz, and Richard Szeliski · 2010
Earlier work this paper cites.
Building rome on a cloudless day
Jan-Michael Frahm, Pierre Fite-Georgel, David Gallup, Tim Johnson, Rahul Raguram, Changchang Wu, Yi-Hung Jen, Enrique Dunn, Brian Clipp, Svetlana Lazebnik, et al · 2010
Earlier work this paper cites.
Towards Internet-scale multi-view stereo
Yasutaka Furukawa, Brian Curless, Steven M. Seitz, and Richard Szeliski · 2010
Earlier work this paper cites.
Building rome in a day
Sameer Agarwal, Yasutaka Furukawa, Noah Snavely, Ian Simon, Brian Curless, Steven M Seitz, and Richard Szeliski · 2011
Earlier work this paper cites.
Global motion estimation from point matches
Mica Arie-Nachimson, Shahar Z Kovalsky, Ira Kemelmacher-Shlizerman, Amit Singer, and Ronen Basri · 2012
Earlier work this paper cites.
Sfm with mrfs: Discrete-continuous optimization for large-scale structure from motion
David J Crandall, Andrew Owens, Noah Snavely, and Daniel P Huttenlocher · 2012
Earlier work this paper cites.
Matchminer: Efficient spanning structure mining in large image collections
Yin Lou, Noah Snavely, and Johannes Gehrke · 2012
Earlier work this paper cites.
‘structure-from-motion’photogrammetry: A low-cost, effective tool for geoscience applications
Matthew J Westoby, James Brasington, Niel F Glasser, Michael J Hambrey, and Jennifer M Reynolds · 2012
Earlier work this paper cites.
A global linear method for camera pose registration
Nianjuan Jiang, Zhaopeng Cui, and Ping Tan · 2013
Earlier work this paper cites.
Global fusion of relative motions for robust, accurate and scalable structure from motion
Pierre Moulon, Pascal Monasse, and Renaud Marlet · 2013
Earlier work this paper cites.
Towards linear-time incremental structure from motion
Changchang Wu · 2013
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context
Tsung-Yi Lin, Michael Maire, Serge J. Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick · 2014
Earlier work this paper cites.
Robust global translations with 1dsfm
Kyle Wilson and Noah Snavely · 2014
Earlier work this paper cites.
Global structure-from-motion by similarity averaging
Zhaopeng Cui and Ping Tan · 2015
Earlier work this paper cites.
Linear global translation estimation with feature tracks
Zhaopeng Cui, Nianjuan Jiang, Chengzhou Tang, and Ping Tan · 2015
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Reconstructing the world* in six days *(as captured by the yahoo 100 million image dataset)
Jared Heinly, Johannes L. Schonberger, Enrique Dunn, and Jan-Michael Frahm · 2015
Earlier work this paper cites.
Robust camera location estimation by convex programming
Onur Ozyesil and Amit Singer · 2015
Earlier work this paper cites.
Optimizing the viewing graph for structure-from-motion
Chris Sweeney, Torsten Sattler, Tobias Hollerer, Matthew Turk, and Marc Pollefeys · 2015
Cited alongside, same era.
Structure from Motion in the Geosciences
Jonathan L Carrivick, Mark W Smith, and Duncan J Quincey · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Cited alongside, same era.
LIFT: Learned Invariant Feature Transform
Kwang Moo Yi, Eduard Trulls, Vincent Lepetit, and Pascal Fua · 2016
Cited alongside, same era.
Dsac-differentiable ransac for camera localization
Eric Brachmann, Alexander Krull, Sebastian Nowozin, Jamie Shotton, Frank Michel, Stefan Gumhold, and Carsten Rother · 2017
Cited alongside, same era.
Emerging properties in self-supervised vision transformers
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin · 2021
Later among the works it cites.
Learning to match features with seeded graph matching network
Hongkai Chen, Zixin Luo, Jiahui Zhang, Lei Zhou, Xuyang Bai, Zeyu Hu, Chiew-Lan Tai, and Long Quan · 2021
Later among the works it cites.
Cotr: Correspondence transformer for matching across images
Wei Jiang, Eduard Trulls, Jan Hosang, Andrea Tagliasacchi, and Kwang Moo Yi · 2021
Later among the works it cites.
Image matching across wide baselines: From paper to practice
Yuhe Jin, Dmytro Mishkin, Anastasiia Mishchuk, Jiri Matas, Pascal Fua, Kwang Moo Yi, and Eduard Trulls · 2021
Later among the works it cites.
Barf: Bundle-adjusting neural radiance fields
Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, and Simon Lucey · 2021
Later among the works it cites.
Pixel-Perfect Structure-from-Motion with Featuremetric Refinement
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hsfm: Hybrid structure-from-motion
Hainan Cui, Xiang Gao, Shuhan Shen, and Zhanyi Hu · 2017
Cited alongside, same era.
What Uncertainties Do We Need in Bayesian Deep Learning for Computer Vision?
Alex Kendall and Yarin Gal · 2017
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Cited alongside, same era.
Learning 3d object categories by looking around them
David Novotny, Diane Larlus, and Andrea Vedaldi · 2017
Cited alongside, same era.
A survey of structure from motion*
Onur Özyeşil, Vladislav Voroninski, Ronen Basri, and Amit Singer · 2017
Cited alongside, same era.
A multi-view stereo benchmark with high-resolution images and multi-camera videos
Thomas Schops, Johannes L Schonberger, Silvano Galliani, Torsten Sattler, Konrad Schindler, Marc Pollefeys, and Andreas Geiger · 2017
Cited alongside, same era.
Philipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, and Marc Pollefeys · 2021
Later among the works it cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T Barron, Ravi Ramamoorthi, and Ren Ng · 2021
Later among the works it cites.
Common objects in 3d: Large-scale learning and evaluation of real-life 3d category reconstruction
Jeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone, Patrick Labatut, and David Novotny · 2021
Later among the works it cites.
Loftr: Detector-free local feature matching with transformers
Jiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao, and Xiaowei Zhou · 2021
Later among the works it cites.
Droid-slam: Deep visual slam for monocular, stereo, and rgb-d cameras
Zachary Teed and Jia Deng · 2021
Later among the works it cites.
Ceres Solver, 2022
Sameer Agarwal, Keir Mierle, and The Ceres Solver Team · 2022
Later among the works it cites.
Tap-vid: A benchmark for tracking any point in a video
Carl Doersch, Ankush Gupta, Larisa Markeeva, Adria Recasens Continente, Kucas Smaira, Yusuf Aytar, Joao Carreira, Andrew Zisserman, and Yi Yang · 2022
Later among the works it cites.
Kubric: a scalable dataset generator
Klaus Greff, Francois Belletti, Lucas Beyer, Carl Doersch, Yilun Du, Daniel Duckworth, David J Fleet, Dan Gnanapragasam, Florian Golemo, Charles Herrmann, Thomas Kipf, Abhijit Kundu, Dmitry Lagun, Issam Laradji, Hsueh-Ti (Derek) Liu, Henning Meyer, Yishu Miao, Derek Nowrouzezahrai, Cengiz Oztireli, Etienne Pot, Noha Radwan, Daniel Rebain, Sara Sabour, Mehdi S. M. Sajjadi, Matan Sela, Vincent Sitzmann, Austin Stone, Deqing Sun, Suhani Vora, Ziyu Wang, Tianhao Wu, Kwang Moo Yi, Fangcheng Zhong, and Andrea Tagliasacchi · 2022
Later among the works it cites.
Particle video revisited: Tracking through occlusions using point trajectories
Adam W Harley, Zhaoyuan Fang, and Katerina Fragkiadaki · 2022
Later among the works it cites.
Few-view object reconstruction with unknown categories and camera poses
Hanwen Jiang, Zhenyu Jiang, Kristen Grauman, and Yuke Zhu · 2022
Later among the works it cites.
Virtual correspondence: Humans as a cue for extreme-view geometry
Wei-Chiu Ma, Anqi Joyce Yang, Shenlong Wang, Raquel Urtasun, and Antonio Torralba · 2022
Later among the works it cites.
Theseus: A library for differentiable nonlinear optimization
Luis Pineda, Taosha Fan, Maurizio Monge, Shobha Venkataraman, Paloma Sodhi, Ricky TQ Chen, Joseph Ortiz, Daniel DeTone, Austin Wang, Stuart Anderson, et al · 2022
Later among the works it cites.
Clustergnn: Cluster-based coarse-to-fine graph neural network for efficient feature matching
Yan Shi, Jun-Xiong Cai, Yoli Shavit, Tai-Jiang Mu, Wensen Feng, and Kai Zhang · 2022
Later among the works it cites.
Matchformer: Interleaving attention in transformers for feature matching
Qing Wang, Jiaming Zhang, Kailun Yang, Kunyu Peng, and Rainer Stiefelhagen · 2022
Later among the works it cites.
Relpose: Predicting probabilistic relative rotation for single objects in the wild
Jason Y Zhang, Deva Ramanan, and Shubham Tulsiani · 2022
Later among the works it cites.
Tapir: Tracking any point with per-frame initialization and temporal refinement
Carl Doersch, Yi Yang, Mel Vecerik, Dilara Gokay, Ankush Gupta, Yusuf Aytar, Joao Carreira, and Andrew Zisserman · 2023
Closest in time.
Detector-free structure from motion
Xingyi He, Jiaming Sun, Yifan Wang, Sida Peng, Qixing Huang, Hujun Bao, and Xiaowei Zhou · 2023
Closest in time.
CoTracker: It is better to track together
Nikita Karaev, Ignacio Rocco, Benjamin Graham, Natalia Neverova, Andrea Vedaldi, and Christian Rupprecht · 2023
Closest in time.
3d gaussian splatting for real-time radiance field rendering
Bernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, and George Drettakis · 2023
Closest in time.
Relpose++: Recovering 6d poses from sparse-view observations
Amy Lin, Jason Y Zhang, Deva Ramanan, and Shubham Tulsiani · 2023
Closest in time.
Lightglue: Local feature matching at light speed
Philipp Lindenberger, Paul-Edouard Sarlin, and Marc Pollefeys · 2023
Closest in time.
Sparsepose: Sparse-view camera pose regression and refinement
Samarth Sinha, Jason Y Zhang, Andrea Tagliasacchi, Igor Gilitschenski, and David B Lindell · 2023
Closest in time.
Posediffusion: Solving pose estimation via diffusion-aided bundle adjustment
Jianyuan Wang, Christian Rupprecht, and David Novotny · 2023
Closest in time.
Generalized differentiable ransac
Tong Wei, Yash Patel, Alexander Shekhovtsov, Jiri Matas, and Daniel Barath · 2023
Closest in time.
MagicPony: Learning articulated 3d animals in the wild
Shangzhe Wu, Ruining Li, Tomas Jakab, Christian Rupprecht, and Andrea Vedaldi · 2023
Closest in time.
Pointodyssey: A large-scale synthetic dataset for long-term point tracking
Yang Zheng, Adam W Harley, Bokui Shen, Gordon Wetzstein, and Leonidas J Guibas · 2023
Closest in time.