Fetching the paper…
Reading the bibliography…
We present a simple baseline for directly estimating the relative pose (rotation and translation, including scale) between two images.
Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography
Martin A Fischler and Robert C Bolles · 1981
Earlier work this paper cites.
A computer algorithm for reconstructing a scene from two projections
H Christopher Longuet-Higgins · 1981
Earlier work this paper cites.
Approximation capabilities of multilayer feedforward networks
Kurt Hornik · 1991
Earlier work this paper cites.
What can be seen in three dimensions with an uncalibrated stereo rig?
Olivier D Faugeras · 1992
Earlier work this paper cites.
Estimation of relative camera positions for uncalibrated cameras
Richard I Hartley · 1992
Earlier work this paper cites.
In defense of the eight-point algorithm
Richard I Hartley · 1997
Earlier work this paper cites.
Wide baseline stereo matching
Phil Pritchett and Andrew Zisserman · 1998
Earlier work this paper cites.
Multiple view geometry in computer vision
Richard Hartley and Andrew Zisserman · 2003
Earlier work this paper cites.
Robust wide-baseline stereo from maximally stable extremal regions
Jiri Matas, Ondrej Chum, Martin Urban, and Tomás Pajdla · 2004
Earlier work this paper cites.
An efficient solution to the five-point relative pose problem
David Nistér · 2004
Earlier work this paper cites.
Surf: Speeded up robust features
Herbert Bay, Tinne Tuytelaars, and Luc Van Gool · 2006
Earlier work this paper cites.
Robust real-time visual odometry with a single camera and an imu
Laurent Kneip, Margarita Chli, and Roland Siegwart · 2011
Earlier work this paper cites.
Orb: An efficient alternative to sift or surf
Ethan Rublee, Vincent Rabaud, Kurt Konolige, and Gary Bradski · 2011
Earlier work this paper cites.
Hand waving away scale
Christopher Ham, Simon Lucey, and Surya Singh · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Wxbs: Wide baseline stereo generalizations
Dmytro Mishkin, Jiri Matas, Michal Perdoch, and Karel Lenc · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L Schonberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Dsac-differentiable ransac for camera localization
Eric Brachmann, Alexander Krull, Sebastian Nowozin, Jamie Shotton, Frank Michel, Stefan Gumhold, and Carsten Rother · 2017
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber, Matthias Niessner, Manolis Savva, Shuran Song, Andy Zeng, and Yinda Zhang · 2017
Earlier work this paper cites.
Camera relocalization by computing pairwise relative poses using convolutional neural network
Zakaria Laskar, Iaroslav Melekhov, Surya Kalia, and Juho Kannala · 2017
Earlier work this paper cites.
Relative camera pose estimation using convolutional neural networks
Iaroslav Melekhov, Juha Ylioinas, Juho Kannala, and Esa Rahtu · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Geometry-aware learning of maps for camera localization
Samarth Brahmbhatt, Jinwei Gu, Kihwan Kim, James Hays, and Jan Kautz · 2018
Earlier work this paper cites.
Superpoint: Self-supervised interest point detection and description
Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2018
Cited alongside, same era.
Rpnet: An end-to-end network for relative camera pose estimation
Sovann En, Alexis Lechervy, and Frédéric Jurie · 2018
Cited alongside, same era.
Minimal solutions for the rotational alignment of imu-camera systems using homography constraints
Banglei Guan, Qifeng Yu, and Friedrich Fraundorfer · 2018
Cited alongside, same era.
Bilinear attention networks
Jin-Hwa Kim, Jaehyun Jun, and Byoung-Tak Zhang · 2018
Cited alongside, same era.
Beyond grobner bases: Basis selection for minimal solvers
Viktor Larsson, Magnus Oskarsson, Kalle Astrom, Alge Wallis, Zuzana Kukelova, and Tomas Pajdla · 2018
Cited alongside, same era.
Interiornet: Mega-scale multi-sensor photo-realistic indoor scenes dataset
Wenbin Li, Sajad Saeedi, John McCormac, Ronald Clark, Dimos Tzoumanikas, Qing Ye, Yuzhong Huang, Rui Tang, and Stefan Leutenegger · 2018
Nerf: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Later among the works it cites.
Associative3d: Volumetric reconstruction from sparse views
Shengyi Qian, Linyi Jin, and David F Fouhey · 2020
Later among the works it cites.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
René Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun · 2020
Later among the works it cites.
Superglue: Learning feature matching with graph neural networks
Paul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2020
Later among the works it cites.
Disk: Learning local features with policy gradient
Michał Tyszkiewicz, Pascal Fua, and Eduard Trulls · 2020
Later among the works it cites.
Tartanair: A dataset to push the limits of visual slam
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep fundamental matrix estimation without correspondences
Omid Poursaeed, Guandao Yang, Aditya Prakash, Qiuren Fang, Hanqing Jiang, Bharath Hariharan, and Serge Belongie · 2018
Cited alongside, same era.
Vlocnet++: Deep multitask learning for semantic visual localization and odometry
Noha Radwan, Abhinav Valada, and Wolfram Burgard · 2018
Cited alongside, same era.
Deep fundamental matrix estimation
René Ranftl and Vladlen Koltun · 2018
Cited alongside, same era.
End-to-end weakly-supervised semantic alignment
I. Rocco, R. Arandjelović, and J. Sivic · 2018
Cited alongside, same era.
Pwc-net: Cnns for optical flow using pyramid, warping, and cost volume
Deqing Sun, Xiaodong Yang, Ming-Yu Liu, and Jan Kautz · 2018
Cited alongside, same era.
Stereo magnification: Learning view synthesis using multiplane images
Tinghui Zhou, Richard Tucker, John Flynn, Graham Fyffe, and Noah Snavely · 2018
Cited alongside, same era.
Wenshan Wang, Delong Zhu, Xiangwei Wang, Yaoyu Hu, Yuheng Qiu, Chen Wang, Yafei Hu, Ashish Kapoor, and Sebastian Scherer · 2020
Later among the works it cites.
Extreme relative pose network under hybrid representations
Zhenpei Yang, Siming Yan, and Qixing Huang · 2020
Later among the works it cites.
To learn or not to learn: Visual localization from essential matrices
Qunjie Zhou, Torsten Sattler, Marc Pollefeys, and Laura Leal-Taixe · 2020
Later among the works it cites.
Extreme rotation estimation using dense correlation volumes
Ruojin Cai, Bharath Hariharan, Noah Snavely, and Hadar Averbuch-Elor · 2021
Later among the works it cites.
Wide-baseline relative camera pose estimation with directional learning
Kefan Chen, Noah Snavely, and Ameesh Makadia · 2021
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2021
Later among the works it cites.
Unsupervisedr&r: Unsupervised point cloud registration via differentiable rendering
Mohamed El Banani, Luya Gao, and Justin Johnson · 2021
Later among the works it cites.
Planar surface reconstruction from sparse views
Linyi Jin, Shengyi Qian, Andrew Owens, and David F Fouhey · 2021
Later among the works it cites.
Transformers in vision: A survey
Salman Khan, Muzammal Naseer, Munawar Hayat, Syed Waqas Zamir, Fahad Shahbaz Khan, and Mubarak Shah · 2021
Later among the works it cites.
Barf: Bundle-adjusting neural radiance fields
Chen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, and Simon Lucey · 2021
Later among the works it cites.
Nerf in the wild: Neural radiance fields for unconstrained photo collections
Ricardo Martin-Brualla, Noha Radwan, Mehdi SM Sajjadi, Jonathan T Barron, Alexey Dosovitskiy, and Daniel Duckworth · 2021
Later among the works it cites.
How do vision transformers work?
Namuk Park and Songkuk Kim · 2021
Later among the works it cites.
Loftr: Detector-free local feature matching with transformers
Jiaming Sun, Zehong Shen, Yuang Wang, Hujun Bao, and Xiaowei Zhou · 2021
Later among the works it cites.
Droid-slam: Deep visual slam for monocular, stereo, and rgb-d cameras
Zachary Teed and Jia Deng · 2021
Later among the works it cites.
Tangent space backpropagation for 3d transformation groups
Zachary Teed and Jia Deng · 2021
Later among the works it cites.
Tartanvo: A generalizable learning-based vo
Wenshan Wang, Yaoyu Hu, and Sebastian Scherer · 2021
Later among the works it cites.
Input-level inductive biases for 3d reconstruction
Wang Yifan, Carl Doersch, Relja Arandjelović, João Carreira, and Andrew Zisserman · 2021
Later among the works it cites.
Planeformers: From sparse view planes to 3d reconstruction
Samir Agarwala, Linyi Jin, Chris Rockwell, and David F. Fouhey · 2022
Closest in time.