Fetching the paper…
Reading the bibliography…
Visual localization aims to determine the camera pose of a query image relative to a database of posed images.
Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography
Martin A Fischler and Robert C Bolles · 1981
Earlier work this paper cites.
Triangulation
Richard I Hartley and Peter Sturm · 1997
Earlier work this paper cites.
Object recognition from local scale-invariant features
David G Lowe · 1999
Earlier work this paper cites.
The opencv library
G Bradski · 2000
Earlier work this paper cites.
Locally optimized ransac
Ondřej Chum, Jiří Matas, and Josef Kittler · 2003
Earlier work this paper cites.
Multiple view geometry in computer vision
Richard Hartley and Andrew Zisserman · 2003
Earlier work this paper cites.
Location recognition using prioritized feature matching
Yunpeng Li, Noah Snavely, and Daniel P Huttenlocher · 2010
Earlier work this paper cites.
A novel parametrization of the perspective-three-point problem for a direct computation of absolute camera position and orientation
Laurent Kneip, Davide Scaramuzza, and Roland Siegwart · 2011
Earlier work this paper cites.
Fast image-based localization using direct 2d-to-3d matching
Torsten Sattler, Bastian Leibe, and Leif Kobbelt · 2011
Earlier work this paper cites.
Three things everyone should know to improve object retrieval
Relja Arandjelović and Andrew Zisserman · 2012
Earlier work this paper cites.
Fixing the locally optimized ransac–full experimental evaluation
Karel Lebeda, Jirı Matas, and Ondrej Chum · 2012
Earlier work this paper cites.
Improving image-based localization by active correspondence search
Torsten Sattler, Bastian Leibe, and Leif Kobbelt · 2012
Earlier work this paper cites.
Efficient and robust large-scale rotation averaging
Avishek Chatterjee and Venu Madhav Govindu · 2013
Earlier work this paper cites.
Scene coordinate regression forests for camera relocalization in rgb-d images
Jamie Shotton, Ben Glocker, Christopher Zach, Shahram Izadi, Antonio Criminisi, and Andrew Fitzgibbon · 2013
Earlier work this paper cites.
Robust global translations with 1dsfm
Kyle Wilson and Noah Snavely · 2014
Earlier work this paper cites.
Posenet: A convolutional network for real-time 6-dof camera relocalization
Alex Kendall, Matthew Grimes, and Roberto Cipolla · 2015
Earlier work this paper cites.
Camera pose voting for large-scale image-based localization
Bernhard Zeisl, Torsten Sattler, and Marc Pollefeys · 2015
Earlier work this paper cites.
Netvlad: Cnn architecture for weakly supervised place recognition
Relja Arandjelovic, Petr Gronat, Akihiko Torii, Tomas Pajdla, and Josef Sivic · 2016
Earlier work this paper cites.
Modelling uncertainty in deep learning for camera relocalization
Alex Kendall and Roberto Cipolla · 2016
Earlier work this paper cites.
Efficient & effective prioritized matching for large-scale image-based localization
Torsten Sattler, Bastian Leibe, and Leif Kobbelt · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes L Schonberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Dsac-differentiable ransac for camera localization
Eric Brachmann, Alexander Krull, Sebastian Nowozin, Jamie Shotton, Frank Michel, Stefan Gumhold, and Carsten Rother · 2017
Earlier work this paper cites.
Scannet: Richly-annotated 3d reconstructions of indoor scenes
Angela Dai, Angel X Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner · 2017
Earlier work this paper cites.
Geometric loss functions for camera pose regression with deep learning
Alex Kendall and Roberto Cipolla · 2017
Earlier work this paper cites.
Camera relocalization by computing pairwise relative poses using convolutional neural network
Zakaria Laskar, Iaroslav Melekhov, Surya Kalia, and Juho Kannala · 2017
Earlier work this paper cites.
Image-based localization using hourglass networks
Iaroslav Melekhov, Juha Ylioinas, Juho Kannala, and Esa Rahtu · 2017
Earlier work this paper cites.
Deep regression for monocular camera-based 6-dof global localization in outdoor environments
Tayyab Naseer and Wolfram Burgard · 2017
Earlier work this paper cites.
Are large-scale 3d models really necessary for accurate visual localization?
Torsten Sattler, Akihiko Torii, Josef Sivic, Marc Pollefeys, Hajime Taira, Masatoshi Okutomi, and Tomas Pajdla · 2017
Earlier work this paper cites.
Semantic scene completion from a single depth image
Shuran Song, Fisher Yu, Andy Zeng, Angel X Chang, Manolis Savva, and Thomas Funkhouser · 2017
Earlier work this paper cites.
Attention is all you need
A Vaswani · 2017
Earlier work this paper cites.
Image-based localization using lstms for structured feature correlation
Florian Walch, Caner Hazirbas, Laura Leal-Taixe, Torsten Sattler, Sebastian Hilsenbeck, and Daniel Cremers · 2017
Earlier work this paper cites.
Delving deeper into convolutional neural networks for camera relocalization
Jian Wu, Liwei Ma, and Xiaolin Hu · 2017
Earlier work this paper cites.
Relocnet: Continuous metric learning relocalisation using neural nets
Vassileios Balntas, Shuda Li, and Victor Prisacariu · 2018
Earlier work this paper cites.
Graph-cut ransac
Daniel Barath and Jiří Matas · 2018
Earlier work this paper cites.
Learning less is more-6d camera localization via 3d surface regression
Eric Brachmann and Carsten Rother · 2018
Earlier work this paper cites.
Geometry-aware learning of maps for camera localization
Samarth Brahmbhatt, Jinwei Gu, Kihwan Kim, James Hays, and Jan Kautz · 2018
Earlier work this paper cites.
Superpoint: Self-supervised interest point detection and description
Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin · 2018
Earlier work this paper cites.
Rpnet: An end-to-end network for relative camera pose estimation
Sovann En, Alexis Lechervy, and Frédéric Jurie · 2018
Earlier work this paper cites.
Full-frame scene coordinate regression for image-based localization
Xiaotian Li, Juha Ylioinas, and Juho Kannala · 2018
Earlier work this paper cites.
Megadepth: Learning single-view depth prediction from internet photos
Zhengqi Li and Noah Snavely · 2018
Cited alongside, same era.
Improved visual relocalization by discovering anchor points
Soham Saha, Girish Varma, and CV Jawahar · 2018
Cited alongside, same era.
Benchmarking 6dof outdoor visual localization in changing conditions
Torsten Sattler, Will Maddern, Carl Toft, Akihiko Torii, Lars Hammarstrand, Erik Stenborg, Daniel Safari, Masatoshi Okutomi, Marc Pollefeys, Josef Sivic, et al · 2018
Cited alongside, same era.
Inloc: Indoor visual localization with dense matching and view synthesis
Hajime Taira, Masatoshi Okutomi, Torsten Sattler, Mircea Cimpoi, Marc Pollefeys, Josef Sivic, Tomas Pajdla, and Akihiko Torii · 2018
Cited alongside, same era.
Stereo magnification: Learning view synthesis using multiplane images
Tinghui Zhou, Richard Tucker, John Flynn, Graham Fyffe, and Noah Snavely · 2018
Dfnet: Enhance absolute pose regression with direct feature matching
Shuai Chen, Xinghui Li, Zirui Wang, and Victor A Prisacariu · 2022
Later among the works it cites.
Visual localization via few-shot scene region classification
Siyan Dong, Shuzhe Wang, Yixin Zhuang, Juho Kannala, Marc Pollefeys, and Baoquan Chen · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2022
Later among the works it cites.
Leveraging image matching toward end-to-end relative camera pose regression
Fadi Khatib, Yuval Margalit, Meirav Galun, and Ronen Basri · 2022
Later among the works it cites.
Lens: Localization enhanced by nerf synthesis
Arthur Moreau, Nathan Piasco, Dzmitry Tsishkou, Bogdan Stanciulescu, and Arnaud de La Fortelle · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Baseline desensitizing in translation averaging
Bingbing Zhuang, Loong-Fah Cheong, and Gim Hee Lee · 2018
Cited alongside, same era.
Magsac: marginalizing sample consensus
Daniel Barath, Jiri Matas, and Jana Noskova · 2019
Cited alongside, same era.
Camnet: Coarse-to-fine retrieval for camera re-localization
Mingyu Ding, Zhe Wang, Jiankai Sun, Jianping Shi, and Ping Luo · 2019
Cited alongside, same era.
D2-net: A trainable cnn for joint description and detection of local features
Mihai Dusmanu, Ignacio Rocco, Tomas Pajdla, Marc Pollefeys, Josef Sivic, Akihiko Torii, and Torsten Sattler · 2019
Cited alongside, same era.
R2d2: Reliable and repeatable detector and descriptor
Jerome Revaud, Cesar De Souza, Martin Humenberger, and Philippe Weinzaepfel · 2019
Cited alongside, same era.
From coarse to fine: Robust hierarchical localization at large scale
Paul-Edouard Sarlin, Cesar Cadena, Roland Siegwart, and Marcin Dymczyk · 2019
Cited alongside, same era.
Understanding the limitations of cnn-based absolute camera pose regression
Torsten Sattler, Qunjie Zhou, Marc Pollefeys, and Laura Leal-Taixe · 2019
Cited alongside, same era.
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Meshloc: Mesh-based visual localization
Vojtech Panek, Zuzana Kukelova, and Torsten Sattler · 2022
Later among the works it cites.
Croco: Self-supervised pre-training for 3d vision tasks by cross-view completion
Philippe Weinzaepfel, Vincent Leroy, Thomas Lucas, Romain Brégier, Yohann Cabon, Vaibhav Arora, Leonid Antsfeld, Boris Chidlovskii, Gabriela Csurka, and Jérôme Revaud · 2022
Later among the works it cites.
Scenesqueezer: Learning to compress scene for camera relocalization
Luwei Yang, Rakesh Shrestha, Wenbo Li, Shuaicheng Liu, Guofeng Zhang, Zhaopeng Cui, and Ping Tan · 2022
Later among the works it cites.
RelPose: Predicting probabilistic relative rotation for single objects in the wild
Jason Y. Zhang, Deva Ramanan, and Shubham Tulsiani · 2022
Later among the works it cites.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al · 2023
Later among the works it cites.
Accelerated coordinate encoding: Learning to relocalize in minutes using rgb and poses
Eric Brachmann, Tommaso Cavallari, and Victor Adrian Prisacariu · 2023
Later among the works it cites.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2023
Later among the works it cites.
Lazy visual localization via motion averaging
Siyan Dong, Shaohui Liu, Hengkai Guo, Baoquan Chen, and Marc Pollefeys · 2023
Later among the works it cites.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Lightglue: Local feature matching at light speed
Philipp Lindenberger, Paul-Edouard Sarlin, and Marc Pollefeys · 2023
Later among the works it cites.
Scalable diffusion models with transformers
William Peebles and Saining Xie · 2023
Later among the works it cites.
Coarse-to-fine multi-scene pose regression with transformers
Yoli Shavit, Ron Ferens, and Yosi Keller · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Posediffusion: Solving pose estimation via diffusion-aided bundle adjustment
Jianyuan Wang, Christian Rupprecht, and David Novotny · 2023
Later among the works it cites.
Croco v2: Improved cross-view completion pre-training for stereo matching and optical flow
Philippe Weinzaepfel, Thomas Lucas, Vincent Leroy, Yohann Cabon, Vaibhav Arora, Romain Brégier, Gabriela Csurka, Leonid Antsfeld, Boris Chidlovskii, and Jérôme Revaud · 2023
Later among the works it cites.
Scannet++: A high-fidelity dataset of 3d indoor scenes
Chandan Yeshwanth, Yueh-Cheng Liu, Matthias Nießner, and Angela Dai · 2023
Later among the works it cites.
A laplace-inspired distribution on so(3) for probabilistic rotation estimation
Yingda Yin, Yang Wang, He Wang, and Baoquan Chen · 2023
Later among the works it cites.
Sequential modeling enables scalable learning for large vision models
Yutong Bai, Xinyang Geng, Karttikeya Mangalam, Amir Bar, Alan L Yuille, Trevor Darrell, Jitendra Malik, and Alexei A Efros · 2024
Closest in time.
Roma: Robust dense feature matching
Johan Edstedt, Qiyu Sun, Georg Bökman, Mårten Wadenbäck, and Michael Felsberg · 2024
Closest in time.
Learning to produce semi-dense correspondences for visual localization
Khang Truong Giang, Soohwan Song, and Sungho Jo · 2024
Closest in time.
Grounding image matching in 3d with mast3r
Vincent Leroy, Yohann Cabon, and Jérôme Revaud · 2024
Closest in time.
Relpose++: Recovering 6d poses from sparse-view observations
Amy Lin, Jason Y Zhang, Deva Ramanan, and Shubham Tulsiani · 2024
Closest in time.
Dl3dv-10k: A large-scale scene dataset for deep learning-based 3d vision
Lu Ling, Yichen Sheng, Zhi Tu, Wentian Zhao, Cheng Xin, Kun Wan, Lantao Yu, Qianyu Guo, Zixun Yu, Yawen Lu, et al · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2024
Closest in time.
Large language models: A survey
Shervin Minaee, Tomas Mikolov, Narjes Nikzad, Meysam Chenaghlu, Richard Socher, Xavier Amatriain, and Jianfeng Gao · 2024
Closest in time.
Sam 2: Segment anything in images and videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloe Rolland, Laura Gustafson, et al · 2024
Closest in time.
Far: Flexible accurate and robust 6dof relative camera pose estimation
Chris Rockwell, Nilesh Kulkarni, Linyi Jin, Jeong Joon Park, Justin Johnson, and David F Fouhey · 2024
Closest in time.
Splatt3r: Zero-shot gaussian splatting from uncalibarated image pairs
Brandon Smart, Chuanxia Zheng, Iro Laina, and Victor Adrian Prisacariu · 2024
Closest in time.
Roformer: Enhanced transformer with rotary position embedding
Jianlin Su, Murtadha Ahmed, Yu Lu, Shengfeng Pan, Wen Bo, and Yunfeng Liu · 2024
Closest in time.
Cambrian-1: A fully open, vision-centric exploration of multimodal llms
Shengbang Tong, Ellis Brown, Penghao Wu, Sanghyun Woo, Manoj Middepogu, Sai Charitha Akula, Jihan Yang, Shusheng Yang, Adithya Iyer, Xichen Pan, et al · 2024
Closest in time.
Panopose: Self-supervised relative pose estimation for panoramic images
Diantao Tu, Hainan Cui, Xianwei Zheng, and Shuhan Shen · 2024
Closest in time.
3d reconstruction with spatial memory
Hengyi Wang and Lourdes Agapito · 2024
Closest in time.
No pose, no problem: Surprisingly simple 3d gaussian splats from sparse unposed images
Botao Ye, Sifei Liu, Haofei Xu, Xueting Li, Marc Pollefeys, Ming-Hsuan Yang, and Songyou Peng · 2024
Closest in time.
Robust incremental structure-from-motion with hybrid features
Shaohui Liu, Yidan Gao, Tianyi Zhang, Rémi Pautrat, Johannes L Schönberger, Viktor Larsson, and Marc Pollefeys · 2025
Closest in time.