Fetching the paper…
Reading the bibliography…
Self-supervised learning through masked autoencoders (MAEs) has recently attracted great attention for remote sensing (RS) image representation learning, and thus embodies a significant potential for content-based image retrieval (CBIR) from ever-growing RS image archives.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in IEEE Conference on Computer Vision and Pattern Recognition , 2009, pp. 248–255
2009
Earlier work this paper cites.
B. Thomee, D. A. Shamma, G. Friedland, B. Elizalde, K. Ni, D. Poland, D. Borth, and L.-J. Li, “YFCC100M: The new data in multimedia research,” Commun. ACM , vol. 59, no. 2, p. 64–73, jan 2016
2016
Earlier work this paper cites.
Y. Li, Y. Zhang, X. Huang, and J. Ma, “Learning source-invariant deep hashing convolutional neural networks for cross-source remote sensing image retrieval,” IEEE Trans. Geosci. Remote Sens. , vol. 56, no. 11, pp. 6521–6536, 2018
2018
Earlier work this paper cites.
B. Chaudhuri, B. Demir, S. Chaudhuri, and L. Bruzzone, “Multilabel remote sensing image retrieval using a semisupervised graph-theoretic method,” IEEE Trans. Geosci. Remote Sens. , vol. 56, no. 2, pp. 1144–1158, 2018
2018
Earlier work this paper cites.
P. Bachman, H. R. Devon, and W. Buchwalter, “Learning representations by maximizing mutual information across views,” Intl. Conf. Neural Inf. Process. Syst. (NeurIPS) , pp. 15 535–15 545, 2019
2019
Earlier work this paper cites.
B. Poole, S. Ozair, A. van den Oord, A. A. Alemi, and G. Tucker, “On variational bounds of mutual information,” International Conference on Machine Learning , pp. 5171–5180, 2019
2019
Earlier work this paper cites.
W. Xiong, Z. Xiong, Y. Zhang, Y. Cui, and X. Gu, “A deep cross-modality hashing network for sar and optical remote sensing images retrieval,” IEEE J. Sel. Top. Appl. Earth Obs. Remote Sens. , vol. 13, pp. 5284–5296, 2020
2020
Earlier work this paper cites.
U. Chaudhuri, B. Banerjee, A. Bhattacharya, and M. Datcu, “Cmir-net : A deep learning based model for cross-modal retrieval in remote sensing,” Pattern Recognition Letters , vol. 131, pp. 456–462, 2020
2020
Earlier work this paper cites.
G. Sumbul, J. Kang, and B. Demir, “Deep learning for image search and retrieval in large remote sensing archives,” in Deep Learning for the Earth Sciences: A comprehensive approach to remote sensing, climate science and geosciences . Hoboken, NJ, USA: Wiley, 2021, ch. 11, pp. 150–160
2021
Earlier work this paper cites.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, and N. Houlsby, “An image is worth 16x16 words: Transformers for image recognition at scale,” Intl. Conf. Learn. Represent. (ICLR) , 2021
2021
Earlier work this paper cites.
G. Sumbul, A. d. Wall, T. Kreuziger, F. Marcelino, H. Costa, P. Benevides, M. Caetano, B. Demir, and V. Markl, “BigEarthNet-MM: A Large-Scale, Multimodal, Multilabel Benchmark Archive for Remote Sensing Image Classification and Retrieval,” IEEE Geosci. Remote Sens. Magazine , vol. 9, no. 3, pp. 174–180, 2021
2021
Earlier work this paper cites.
T. Ridnik, E. Ben-Baruch, A. Noy, and L. Zelnik, “ImageNet-21K pretraining for the masses,” in Proceedings of the Neural Information Processing Systems Track on Datasets and Benchmarks , vol. 1, 2021
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever, “Learning transferable visual models from natural language supervision,” in Intl. Conf. Mach. Learn. , vol. 139, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
G. Sumbul, A. de Wall, T. Kreuziger, F. Marcelino, H. Costa, P. Benevides, M. Caetano, B. Demir, and V. Markl, “BigEarthNet-MM: A large scale multi-modal multi-label benchmark archive for remote sensing image classification and retrieval,” IEEE Geosci. Remote Sens. Magazine , vol. 9, no. 3, pp. 174–180, 2021
2021
Earlier work this paper cites.
G. Sumbul and B. Demir, “A novel graph-theoretic deep representation learning method for multi-label remote sensing image retrieval,” in IEEE International Geoscience and Remote Sensing Symposium , 2021, pp. 266–269
2021
Earlier work this paper cites.
H. Touvron, M. Cord, M. Douze, F. Massa, A. Sablayrolles, and H. Jegou, “Training data-efficient image transformers & distillation through attention,” Intl. Conf. Mach. Learn. , pp. 10 347–10 357, 2021
2021
Cited alongside, same era.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, and N. Houlsby, “An image is worth 16x16 words: Transformers for image recognition at scale,” Intl. Conf. Learn. Represent. (ICLR) , 2021
2021
Cited alongside, same era.
X. Tang, Y. Yang, J. Ma, Y.-M. Cheung, C. Liu, F. Liu, X. Zhang, and L. Jiao, “Meta-hashing for remote sensing image retrieval,” IEEE Trans. Geosci. Remote Sens. , vol. 60, pp. 1–19, 2022
2022
Cited alongside, same era.
G. Sumbul and B. Demir, “Plasticity-stability preserving multi-task learning for remote sensing image retrieval,” IEEE Trans. Geosci. Remote Sens. , vol. 60, no. 5620116, pp. 1–16, 2022
2022
Cited alongside, same era.
Y. Su, G. Zhang, S. Mei, J. Lian, Y. Wang, and S. Wan, “Reconstruction-assisted and distance-optimized adversarial training: A defense framework for remote sensing scene classification,” IEEE Trans. Geosci. Remote Sens. , vol. 61, pp. 1–13, 2023
2023
Later among the works it cites.
Y. Wang, N. A. A. Braham, Z. Xiong, C. Liu, C. M. Albrecht, and X. X. Zhu, “Ssl4eo-s12: A large-scale multimodal, multitemporal dataset for self-supervised learning in earth observation [software and data sets],” IEEE Geosci. Remote Sens. Magazine , vol. 11, no. 3, pp. 98–106, 2023
2023
Later among the works it cites.
D. Muhtar, X. Zhang, P. Xiao, Z. Li, and F. Gu, “Cmid: A unified self-supervised learning framework for remote sensing image understanding,” IEEE Trans. Geosci. Remote Sens. , vol. 61, pp. 1–17, 2023
2023
Later among the works it cites.
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. Sumbul, M. Müller, and B. Demir, “A novel self-supervised cross-modal image retrieval method in remote sensing,” in IEEE Intl. Conf. Image Process. , 2022, pp. 2426–2430
2022
Cited alongside, same era.
H. Jung, Y. Oh, S. Jeong, C. Lee, and T. Jeon, “Contrastive self-supervised learning with smoothed representation for remote sensing,” IEEE Geosci. Remote Sens. Lett. , vol. 19, pp. 1–5, 2022
2022
Cited alongside, same era.
Z. Xie, Z. Zhang, Y. Cao, Y. Lin, J. Bao, Z. Yao, Q. Dai, and H. Hu, “Simmim: a simple framework for masked image modeling,” in IEEE/CVF Conf. Comput. Vis. Pattern Recog. (CVPR) , 2022, pp. 9643–9653
2022
Cited alongside, same era.
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. Girshick, “Masked autoencoders are scalable vision learners,” in IEEE/CVF Conf. Comput. Vis. Pattern Recog. (CVPR) , 2022, pp. 15 979–15 988
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Y. Sun, S. Feng, Y. Ye, X. Li, J. Kang, Z. Huang, and C. Luo, “Multisensor fusion and explicit semantic preserving-based deep hashing for cross-modal remote sensing image retrieval,” IEEE Trans. Geosci. Remote Sens. , vol. 60, pp. 1–14, 2022
2022
Cited alongside, same era.
U. Chaudhuri, R. Bose, B. Banerjee, A. Bhattacharya, and M. Datcu, “Zero-shot cross-modal retrieval for remote sensing images with minimal supervision,” IEEE Trans. Geosci. Remote Sens. , vol. 60, pp. 1–15, 2022
2022
Cited alongside, same era.
Y. Gao, X. Sun, and C. Liu, “A general self-supervised framework for remote sensing image classification,” Remote Sens. , vol. 14, no. 19, 2022
2022
Cited alongside, same era.
Later among the works it cites.
C. J. Reed, R. Gupta, S. Li, S. Brockman, C. Funk, B. Clipp, K. Keutzer, S. Candido, M. Uyttendaele, and T. Darrell, “Scale-mae: A scale-aware masked autoencoder for multiscale geospatial representation learning,” Intl. Conf. Comput. Vis. (ICCV) , 2023
2023
Later among the works it cites.
D. Wang, Q. Zhang, Y. Xu, J. Zhang, B. Du, D. Tao, and L. Zhang, “Advancing plain vision transformer toward remote sensing foundation model,” IEEE Trans. Geosci. Remote Sens. , vol. 61, no. 5607315, pp. 1–15, 2023
2023
Later among the works it cites.
X. Sun, P. Wang, W. Lu, Z. Zhu, X. Lu, Q. He, J. Li, X. Rong, Z. Yang, H. Chang, Q. He, G. Yang, R. Wang, J. Lu, and K. Fu, “Ringmo: A remote sensing foundation model with masked image modeling,” IEEE Trans. Geosci. Remote Sens. , vol. 61, no. 5612822, pp. 1–22, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
G. Sumbul and B. Demir, “Generative reasoning integrated label noise robust deep image representation learning,” IEEE Trans. Image Process. , vol. 32, pp. 4529–4542, 2023
2023
Later among the works it cites.
G. Kwon, Z. Cai, A. Ravichandran, E. Bas, R. Bhotika, and S. Soatto, “Masked vision and language modeling for multi-modal representation learning,” International Conference on Learning Representations , 2023
2023
Later among the works it cites.
2024
Closest in time.
B. Han, S. Zhang, X. Shi, and M. Reichstein, “Bridging remote sensors with multisensor geospatial foundation models,” in IEEE/CVF Conf. Comput. Vis. Pattern Recog. (CVPR) , 2024, pp. 27 852–27 862
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.