Fetching the paper…
Reading the bibliography…
Local feature matching is an essential component in many visual applications.
Y. Tao, D. Papadias, and Q. Shen, “Continuous nearest neighbor search,” in VLDB’02: Proceedings of the 28th International Conference on Very Large Databases . Elsevier, 2002, pp. 287–298
2002
Earlier work this paper cites.
D. G. Lowe, “Distinctive image features from scale-invariant keypoints,” International journal of computer vision , vol. 60, no. 2, pp. 91–110, 2004
2004
Earlier work this paper cites.
H. Bay, T. Tuytelaars, and L. V. Gool, “Surf: Speeded up robust features,” in European conference on computer vision . Springer, 2006, pp. 404–417
2006
Earlier work this paper cites.
E. Rublee, V. Rabaud, K. Konolige, and G. Bradski, “Orb: An efficient alternative to sift or surf,” in 2011 International conference on computer vision . Ieee, 2011, pp. 2564–2571
2011
Earlier work this paper cites.
M. Cuturi, “Sinkhorn distances: Lightspeed computation of optimal transport,” Advances in neural information processing systems , vol. 26, 2013
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. L. Schonberger and J.-M. Frahm, “Structure-from-motion revisited,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 4104–4113
2016
Earlier work this paper cites.
M. Karpushin, G. Valenzise, and F. Dufaux, “Keypoint detection in rgbd images based on an anisotropic scale space,” IEEE Transactions on Multimedia , vol. 18, no. 9, pp. 1762–1771, 2016
2016
Earlier work this paper cites.
Y. Song, X. Chen, X. Wang, Y. Zhang, and J. Li, “6-dof image localization from massive geo-tagged reference images,” IEEE Transactions on Multimedia , vol. 18, no. 8, pp. 1542–1554, 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
R. Mur-Artal and J. D. Tardós, “Orb-slam2: An open-source slam system for monocular, stereo, and rgb-d cameras,” IEEE transactions on robotics , vol. 33, no. 5, pp. 1255–1262, 2017
2017
Earlier work this paper cites.
J. Bian, W.-Y. Lin, Y. Matsushita, S.-K. Yeung, T.-D. Nguyen, and M.-M. Cheng, “Gms: Grid-based motion statistics for fast, ultra-robust feature correspondence,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 4181–4190
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in neural information processing systems , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. Girshick, K. He, B. Hariharan, and S. Belongie, “Feature pyramid networks for object detection,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 2117–2125
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner, “Scannet: Richly-annotated 3d reconstructions of indoor scenes,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 5828–5839
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
V. Balntas, K. Lenc, A. Vedaldi, and K. Mikolajczyk, “Hpatches: A benchmark and evaluation of handcrafted and learned local descriptors,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 5173–5182
2017
Earlier work this paper cites.
T. Qin, P. Li, and S. Shen, “Vins-mono: A robust and versatile monocular visual-inertial state estimator,” IEEE Transactions on Robotics , vol. 34, no. 4, pp. 1004–1020, 2018
2018
Earlier work this paper cites.
D. DeTone, T. Malisiewicz, and A. Rabinovich, “Superpoint: Self-supervised interest point detection and description,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops , June 2018
2018
Cited alongside, same era.
D. DeTone, T. Malisiewicz, and A. Rabinovich, “Superpoint: Self-supervised interest point detection and description,” in Proceedings of the IEEE conference on computer vision and pattern recognition workshops , 2018, pp. 224–236
2018
Cited alongside, same era.
I. Rocco, M. Cimpoi, R. Arandjelović, A. Torii, T. Pajdla, and J. Sivic, “Neighbourhood consensus networks,” Advances in neural information processing systems , vol. 31, 2018
2018
Cited alongside, same era.
K. Fu, Q. Zhao, and I. Y.-H. Gu, “Refinet: A deep segmentation assisted refinement network for salient object detection,” IEEE Transactions on Multimedia , vol. 21, no. 2, pp. 457–469, 2018
2018
Cited alongside, same era.
U. Efe, K. G. Ince, and A. Alatan, “Dfm: A performance baseline for deep feature matching,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 4284–4293
2021
Later among the works it cites.
J. Sun, Z. Shen, Y. Wang, H. Bao, and X. Zhou, “Loftr: Detector-free local feature matching with transformers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 8922–8931
2021
Later among the works it cites.
2021
Later among the works it cites.
H. Cui, D. Tu, F. Tang, P. Xu, H. Liu, and S. Shen, “Vidsfm: Robust and accurate structure-from-motion for monocular videos,” IEEE Transactions on Image Processing , vol. 31, pp. 2449–2462, 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Dusmanu, I. Rocco, T. Pajdla, M. Pollefeys, J. Sivic, A. Torii, and T. Sattler, “D2-net: A trainable cnn for joint description and detection of local features,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019
2019
Cited alongside, same era.
J. Revaud, P. Weinzaepfel, C. De Souza, N. Pion, G. Csurka, Y. Cabon, and M. Humenberger, “R2d2: repeatable and reliable detector and descriptor,” in NeurIPS , 2019
2019
Cited alongside, same era.
J. Zhang, D. Sun, Z. Luo, A. Yao, L. Zhou, T. Shen, Y. Chen, L. Quan, and H. Liao, “Learning two-view correspondences and geometry using order-aware network,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 5845–5854
2019
Cited alongside, same era.
K. Sun, W. Tao, and Y. Qian, “Guide to match: multi-layer feature matching with a hybrid gaussian mixture model,” IEEE Transactions on Multimedia , vol. 22, no. 9, pp. 2246–2261, 2019
2019
Cited alongside, same era.
M. J. Tyszkiewicz, P. Fua, and E. Trulls, “Disk: Learning local features with policy gradient,” in NeurIPS , 2020
2020
Cited alongside, same era.
P.-E. Sarlin, D. DeTone, T. Malisiewicz, and A. Rabinovich, “Superglue: Learning feature matching with graph neural networks,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 4938–4947
2020
Cited alongside, same era.
Z. Luo, L. Zhou, X. Bai, H. Chen, J. Zhang, Y. Yao, S. Li, T. Fang, and L. Quan, “Aslfeat: Learning local features of accurate shape and localization,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 6589–6598
2020
Cited alongside, same era.
I. Rocco, R. Arandjelović, and J. Sivic, “Efficient neighbourhood consensus networks via submanifold sparse convolutions,” in European Conference on Computer Vision . Springer, 2020, pp. 605–621
2020
Cited alongside, same era.
X. Zhao, X. Wu, J. Miao, W. Chen, P. C. Chen, and Z. Li, “Alike: Accurate and lightweight keypoint detection and descriptor extraction,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
B. Fan, Y. Yang, W. Feng, F. Wu, J. Lu, and H. Liu, “Seeing through darkness: Visual localization at night via weakly supervised learning of domain invariant features,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
Y. Shi, J.-X. Cai, Y. Shavit, T.-J. Mu, W. Feng, and K. Zhang, “Clustergnn: Cluster-based coarse-to-fine graph neural network for efficient feature matching,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 12 517–12 526
2022
Later among the works it cites.
J. Chen, S. Chen, X. Chen, Y. Dai, and Y. Yang, “Csr-net: Learning adaptive context structure representation for robust feature correspondence,” IEEE Transactions on Image Processing , vol. 31, pp. 3197–3210, 2022
2022
Later among the works it cites.
J. Ma, Y. Wang, A. Fan, G. Xiao, and R. Chen, “Correspondence attention transformer: A context-sensitive network for two-view correspondence learning,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
K. Truong Giang, S. Song, and S. Jo, “Topicfm: Robust and interpretable feature matching with topic-assisted,” arXiv e-prints , pp. arXiv–2207, 2022
2022
Later among the works it cites.
H. Chen, Z. Luo, L. Zhou, Y. Tian, M. Zhen, T. Fang, D. McKinnon, Y. Tsin, and L. Quan, “Aspanformer: Detector-free image matching with adaptive span transformer,” in European Conference on Computer Vision . Springer, 2022, pp. 20–36
2022
Later among the works it cites.
W. Li, H. Liu, R. Ding, M. Liu, P. Wang, and W. Yang, “Exploiting temporal contexts with strided transformer for 3d human pose estimation,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
S. Jiayao, S. Zhou, Y. Cui, and Z. Fang, “Real-time 3d single object tracking with transformer,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
J. Pei, T. Cheng, H. Tang, and C. Chen, “Transformer-based efficient salient instance segmentation networks with orientative query,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Cai, L. Li, D. Wang, X. Li, and X. Liu, “Htmatch: An efficient hybrid transformer based graph neural network for local feature matching,” Signal Processing , p. 108859, 2022
2022
Later among the works it cites.
Z. Li and N. Snavely, “Megadepth: Learning single-view depth prediction from internet photos,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 2041–2050
2050
Closest in time.