Fetching the paper…
Reading the bibliography…
In this work, we aim at an important but less explored problem of a simple yet effective backbone specific for cross-view geo-localization task.
Bansal M, Sawhney HS, Cheng H, et al (2011) Geo-localization of street views with aerial image databases. In: Proceedings of the 19th ACM international conference on Multimedia, pp 1125–1128
2011
Earlier work this paper cites.
Senlet T, Elgammal A (2012) Satellite image based precise robot localization on sidewalks. In: 2012 IEEE International Conference on Robotics and Automation, IEEE, pp 2647–2653
2012
Earlier work this paper cites.
Viswanathan A, Pires BR, Huber D (2014) Vision based robot localization by ground to satellite matching in gps-denied situations. In: 2014 IEEE/RSJ International Conference on Intelligent Robots and Systems, IEEE, pp 192–198
2014
Earlier work this paper cites.
Schroff F, Kalenichenko D, Philbin J (2015) Facenet: A unified embedding for face recognition and clustering. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 815–823
2015
Earlier work this paper cites.
Workman S, Jacobs N (2015) On the location dependence of convolutional neural network features. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, pp 70–78
2015
Earlier work this paper cites.
Workman S, Souvenir R, Jacobs N (2015) Wide-area image geolocalization with aerial reference imagery. In: Proceedings of the IEEE International Conference on Computer Vision, pp 3961–3969
2015
Earlier work this paper cites.
He K, Zhang X, Ren S, et al (2016) Deep residual learning for image recognition. In: CVPR, pp 770–778
2016
Earlier work this paper cites.
Varior RR, Haloi M, Wang G (2016) Gated siamese convolutional neural network architecture for human re-identification. In: European conference on computer vision, Springer, pp 791–808
2016
Earlier work this paper cites.
Vo NN, Hays J (2016) Localizing and orienting street views using overhead imagery. In: European conference on computer vision, Springer, pp 494–509
2016
Earlier work this paper cites.
Chen W, Chen X, Zhang J, et al (2017) Beyond triplet loss: a deep quadruplet network for person re-identification. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 403–412
2017
Earlier work this paper cites.
Hermans A, Beyer L, Leibe B (2017) In defense of the triplet loss for person re-identification. arXiv preprint arXiv:170307737
2017
Earlier work this paper cites.
Selvaraju RR, Cogswell M, Das A, et al (2017) Grad-cam: Visual explanations from deep networks via gradient-based localization. In: Proceedings of the IEEE international conference on computer vision, pp 618–626
2017
Earlier work this paper cites.
Vaswani A, Shazeer N, Parmar N, et al (2017) Attention is all you need. In: NeurIPS, pp 5998–6008
2017
Earlier work this paper cites.
Wang J, Zhou F, Wen S, et al (2017) Deep metric learning with angular loss. In: Proceedings of the IEEE international conference on computer vision, pp 2593–2601
2017
Cited alongside, same era.
Wu CY, Manmatha R, Smola AJ, et al (2017) Sampling matters in deep embedding learning. In: Proceedings of the IEEE International Conference on Computer Vision, pp 2840–2848
2017
Cited alongside, same era.
Zhai M, Bessinger Z, Workman S, et al (2017) Predicting ground-level scene layout from aerial imagery. In: CVPR, pp 867–875
2017
Cited alongside, same era.
Radenović F, Iscen A, Tolias G, et al (2018) Revisiting oxford and paris: Large-scale image retrieval benchmarking. In: CVPR, pp 5706–5715
2018
Cited alongside, same era.
Cai S, Guo Y, Khan S, et al (2019) Ground-to-aerial image geo-localization with a hard exemplar reweighting triplet loss. In: ICCV, pp 8391–8400
El-Nouby A, Neverova N, Laptev I, et al (2021) Training vision transformers for image retrieval. arXiv preprint arXiv:210205644
2021
Later among the works it cites.
Han K, Xiao A, Wu E, et al (2021) Transformer in transformer. Advances in Neural Information Processing Systems 34
2021
Later among the works it cites.
Liu Z, Lin Y, Cao Y, et al (2021) Swin transformer: Hierarchical vision transformer using shifted windows. In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp 10,012–10,022
2021
Later among the works it cites.
Toker A, Zhou Q, Maximov M, et al (2021) Coming down to earth: Satellite-to-street view synthesis for geo-localization. In: CVPR, pp 6488–6497
2021
Later among the works it cites.
Tolstikhin IO, Houlsby N, Kolesnikov A, et al (2021) Mlp-mixer: An all-mlp architecture for vision. Advances in Neural Information Processing Systems 34
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Liu L, Li H (2019) Lending orientation to neural networks for cross-view geo-localization. In: CVPR, pp 5624–5633
2019
Cited alongside, same era.
Regmi K, Shah M (2019) Bridging the domain gap for ground-to-aerial image matching. In: ICCV, pp 470–479
2019
Cited alongside, same era.
Shi Y, Liu L, Yu X, et al (2019) Spatial-aware feature aggregation for image based cross-view geo-localization. NeurIPS 32:10,090–10,100
2019
Cited alongside, same era.
Sun B, Chen C, Zhu Y, et al (2019) Geocapsnet: Ground to aerial view image geo-localization using capsule network. In: ICME, IEEE, pp 742–747
2019
Cited alongside, same era.
Foret P, Kleiner A, Mobahi H, et al (2020) Sharpness-aware minimization for efficiently improving generalization. arXiv preprint arXiv:201001412
2020
Cited alongside, same era.
Hu S, Lee GH (2020) Image-based geo-localization using satellite imagery. International Journal of Computer Vision 128(5):1205–1219
2020
Cited alongside, same era.
Zheng Z, Wei Y, Yang Y (2020) University-1652: A multi-view multi-source benchmark for drone-based geo-localization. In: ACMMM, pp 1395–1403
2020
Cited alongside, same era.
2021
Later among the works it cites.
Touvron H, Cord M, Douze M, et al (2021) Training data-efficient image transformers & distillation through attention. In: ICML, PMLR, pp 10,347–10,357
2021
Later among the works it cites.
Wang F, Liu H (2021) Understanding the behaviour of contrastive loss. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp 2495–2504
2021
Later among the works it cites.
Xiao T, Dollar P, Singh M, et al (2021) Early convolutions help transformers see better. NeurIPS 34
2021
Later among the works it cites.
Yang H, Lu X, Zhu Y (2021) Cross-view geo-localization with layer-to-layer transformer. NeurIPS 34
2021
Later among the works it cites.
Yuan L, Chen Y, Wang T, et al (2021) Tokens-to-token vit: Training vision transformers from scratch on imagenet. In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp 558–567
2021
Later among the works it cites.
Zhu S, Shah M, Chen C (2022) Transgeo: Transformer is all you need for cross-view image geo-localization. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 1162–1171
2022
Later among the works it cites.
Senlet T, Elgammal A (2011) A framework for global vehicle localization using stereo images and satellite and road maps. In: 2011 IEEE International Conference on Computer Vision Workshops (ICCV Workshops), IEEE, pp 2034–2041
2041
Closest in time.