Fetching the paper…
Reading the bibliography…
Change captioning has become essential for accurately describing changes in multi-temporal remote sensing data, providing an intuitive way to monitor Earth's dynamics through natural language.
Y. Qiu, S. Yamamoto, K. Nakashima, R. Suzuki, K. Iwata, H. Kataoka, and Y. Satoh, “Describing and localizing multiple changes with transformers,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 1971–1980
1980
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting of the Association for Computational Linguistics , 2002, pp. 311–318
2002
Earlier work this paper cites.
C.-Y. Lin, “Rouge: A package for automatic evaluation of summaries,” in Text summarization branches out , 2004, pp. 74–81
2004
Earlier work this paper cites.
S. Banerjee and A. Lavie, “Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,” in Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and/or summarization , 2005, pp. 65–72
2005
Earlier work this paper cites.
C. Song, B. Huang, L. Ke, and K. S. Richards, “Remote sensing of alpine lake water environment changes on the tibetan plateau and surroundings: A review,” ISPRS Journal of Photogrammetry and Remote Sensing , vol. 92, pp. 26–37, 2014
2014
Earlier work this paper cites.
E. A. Wentz, S. Anderson, M. Fragkias, M. Netzband, V. Mesev, S. W. Myint, D. Quattrochi, A. Rahman, and K. C. Seto, “Supporting global environmental change research: A review of trends and knowledge gaps in urban remote sensing,” Remote Sensing , vol. 6, no. 5, pp. 3879–3905, 2014
2014
Earlier work this paper cites.
D. P. Kingma, “Adam: A method for stochastic optimization,” arXiv preprint arXiv:1412.6980 , 2014
2014
Earlier work this paper cites.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan, “Show and tell: A neural image caption generator,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 3156–3164
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh, “Vqa: Visual question answering,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2425–2433
2015
Earlier work this paper cites.
F. Yan and K. Mikolajczyk, “Deep correlation for matching images and text,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 3441–3450
2015
Earlier work this paper cites.
R. Vedantam, C. Lawrence Zitnick, and D. Parikh, “Cider: Consensus-based image description evaluation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 4566–4575
2015
Earlier work this paper cites.
S. Reed, Z. Akata, X. Yan, L. Logeswaran, B. Schiele, and H. Lee, “Generative adversarial text to image synthesis,” in International conference on machine learning . PMLR, 2016, pp. 1060–1069
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
L. Chen, H. Zhang, J. Xiao, L. Nie, J. Shao, W. Liu, and T.-S. Chua, “Sca-cnn: Spatial and channel-wise attention in convolutional networks for image captioning,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 5659–5667
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 7132–7141
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
D. H. Park, T. Darrell, and A. Rohrbach, “Robust change captioning,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 4624–4633
2019
Earlier work this paper cites.
S. Lobry, D. Marcos, J. Murray, and D. Tuia, “Rsvqa: Visual question answering for remote sensing data,” IEEE Transactions on Geoscience and Remote Sensing , vol. 58, no. 12, pp. 8555–8566, 2020
2020
Earlier work this paper cites.
H. Chen and Z. Shi, “A spatial-temporal attention-based method and a new dataset for remote sensing image change detection,” Remote Sensing , vol. 12, no. 10, p. 1662, 2020
2020
Cited alongside, same era.
R. Zhao, Z. Shi, and Z. Zou, “High-resolution remote sensing image captioning based on structured attention,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–14, 2021
2021
Cited alongside, same era.
Q. Cheng, Y. Zhou, P. Fu, Y. Xu, and L. Zhang, “A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,” IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing , vol. 14, pp. 4284–4297, 2021
2021
Cited alongside, same era.
M. Hosseinzadeh and Y. Wang, “Image change captioning by learning from an auxiliary task,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 2725–2734
2021
Cited alongside, same era.
K. Kuckreja, M. S. Danish, M. Naseer, A. Das, S. Khan, and F. S. Khan, “Geochat: Grounded large vision-language model for remote sensing,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 27 831–27 840
2024
Later among the works it cites.
Z. Zhang, T. Zhao, Y. Guo, and J. Yin, “Rs5m and georsclip: A large scale vision-language dataset and a large vision-language model for remote sensing,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
W. Zhang, M. Cai, T. Zhang, Y. Zhuang, and X. Mao, “Earthgpt: A universal multi-modal large language model for multi-sensor image comprehension in remote sensing domain,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Q. Cheng, H. Huang, Y. Xu, Y. Zhou, H. Li, and Z. Wang, “Nwpu-captions dataset and mlca-net for remote sensing image captioning,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–19, 2022
2022
Cited alongside, same era.
Z. Yuan, W. Zhang, C. Tian, X. Rong, Z. Zhang, H. Wang, K. Fu, and X. Sun, “Remote sensing cross-modal text-image retrieval based on global and local information,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–16, 2022
2022
Cited alongside, same era.
X. Zhang, W. Yu, and M.-O. Pun, “Multilevel deformable attention-aggregated networks for change detection in bitemporal remote sensing imagery,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–18, 2022
2022
Cited alongside, same era.
C. Liu, R. Zhao, H. Chen, Z. Zou, and Z. Shi, “Remote sensing image change captioning with dual-branch transformers: A new method and a large scale dataset,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–20, 2022
2022
Cited alongside, same era.
G. Hoxha, S. Chouaf, F. Melgani, and Y. Smara, “Change captioning: A new paradigm for multitemporal remote sensing image analysis,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–14, 2022
2022
Cited alongside, same era.
Z. Guo, T.-J. J. Wang, and J. Laaksonen, “Clip4idc: Clip for image difference captioning,” AACL-IJCNLP 2022 , p. 33, 2022
2022
Cited alongside, same era.
Y. Zhan, Z. Xiong, and Y. Yuan, “Rsvg: Exploring data and models for visual grounding on remote sensing data,” IEEE Transactions on Geoscience and Remote Sensing , vol. 61, pp. 1–13, 2023
2023
Cited alongside, same era.
Y. Xu, W. Yu, P. Ghamisi, M. Kopp, and S. Hochreiter, “Txt2img-mhn: Remote sensing image generation from text using modern hopfield networks,” IEEE Transactions on Image Processing , 2023
2023
Cited alongside, same era.
Later among the works it cites.
F. Liu, D. Chen, Z. Guan, X. Zhou, J. Zhu, Q. Ye, L. Fu, and J. Zhou, “Remoteclip: A vision language foundation model for remote sensing,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Later among the works it cites.
S. Dong, L. Wang, B. Du, and X. Meng, “Changeclip: Remote sensing change detection with multimodal vision-language representation learning,” ISPRS Journal of Photogrammetry and Remote Sensing , vol. 208, pp. 53–69, 2024
2024
Later among the works it cites.
Y. Wang and P. Ghamisi, “Rsadapter: Adapting multimodal models for remote sensing visual question answering,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Later among the works it cites.
L. Blickensdörfer, K. Oehmichen, D. Pflugmacher, B. Kleinschmit, and P. Hostert, “National tree species mapping using sentinel-1/2 time series and german national forest inventory data,” Remote Sensing of Environment , vol. 304, p. 114069, 2024
2024
Later among the works it cites.
P. Liu, C. Ren, Z. Wang, M. Jia, W. Yu, H. Ren, and C. Xia, “Evaluating the potential of sentinel-2 time series imagery and machine learning for tree species classification in a mountainous forest,” Remote Sensing , vol. 16, no. 2, p. 293, 2024
2024
Later among the works it cites.
F. Valerio, S. Godinho, A. T. Marques, T. Crispim-Mendes, R. Pita, and J. P. Silva, “Gee_xtract: High-quality remote sensing data preparation and extraction for multiple spatio-temporal ecological scaling,” Ecological Informatics , vol. 80, p. 102502, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
W. Yu, X. Zhang, S. Das, X. Xiang Zhu, and P. Ghamisi, “Maskcd: A remote sensing change detection network based on mask classification,” IEEE Transactions on Geoscience and Remote Sensing , vol. 62, pp. 1–16, 2024
2024
Later among the works it cites.
C. Liu, K. Chen, B. Chen, H. Zhang, Z. Zou, and Z. Shi, “Rscama: Remote sensing image change captioning with state space model,” IEEE Geoscience and Remote Sensing Letters , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Liu, K. Chen, H. Zhang, Z. Qi, Z. Zou, and Z. Shi, “Change-agent: Towards interactive comprehensive remote sensing change interpretation and analysis,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Zhu, L. Li, K. Chen, C. Liu, F. Zhou, and Z. Shi, “Semantic-cc: Boosting remote sensing image change captioning via foundational knowledge and semantic guidance,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Later among the works it cites.
2025
Closest in time.
H. Kim, J. Kim, H. Lee, H. Park, and G. Kim, “Agnostic change captioning with cycle consistency,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 2095–2104
2095
Closest in time.