Fetching the paper…
Reading the bibliography…
Given a language expression, referring remote sensing image segmentation (RRSIS) aims to identify ground objects and assign pixel-wise labels within the imagery.
E. Loper and S. Bird, “Nltk: The natural language toolkit,” arXiv preprint cs/0205028 , 2002
2002
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in 2009 IEEE conference on computer vision and pattern recognition . Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
R. Hu, M. Rohrbach, and T. Darrell, “Segmentation from natural language expressions,” in Proceedings of the European Conference on Computer Vision (ECCV) . Springer, 2016, pp. 108–124
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
S. Lei, Z. Shi, and Z. Zou, “Super-resolution for remote sensing images via local–global combined network,” IEEE Geoscience and Remote Sensing Letters , vol. 14, no. 8, pp. 1243–1247, 2017
2017
Earlier work this paper cites.
C. Liu, Z. Lin, X. Shen, J. Yang, X. Lu, and A. Yuille, “Recurrent multimodal interaction for referring image segmentation,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 1271–1280
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
E. Margffoy-Tuay, J. C. Pérez, E. Botero, and P. Arbeláez, “Dynamic multimodal instance segmentation guided by natural language queries,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 630–645
2018
Earlier work this paper cites.
H. Shi, H. Li, F. Meng, and Q. Wu, “Key-word-aware network for referring expression image segmentation,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 38–54
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
R. Li, K. Li, Y.-C. Kuo, M. Shu, X. Qi, X. Shen, and J. Jia, “Referring image segmentation via recurrent refinement networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 5745–5753
2018
Earlier work this paper cites.
R. Li, K. Li, Y.-C. Kuo, M. Shu, X. Qi, X. Shen, and J. Jia, “Referring image segmentation via recurrent refinement networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 5745–5753
2018
Earlier work this paper cites.
X. Xiao, L. Wang, K. Ding, S. Xiang, and C. Pan, “Deep hierarchical encoder-decoder network for image captioning,” IEEE Transactions on Multimedia , vol. 21, no. 11, pp. 2942–2956, 2019
2019
Earlier work this paper cites.
X. Xiao, L. Wang, S. Xiang, and C. Pan, “What and where the themes dominate in image,” in The Thirty-Third AAAI Conference on Artificial Intelligence, (AAAI) . AAAI Press, 2019, pp. 9021–9029
2019
Earlier work this paper cites.
X. Xiao, L. Wang, K. Ding, S. Xiang, and C. Pan, “Dense semantic embedding network for image captioning,” Pattern Recognit. , vol. 90, pp. 285–296, 2019
2019
Earlier work this paper cites.
L. Ye, M. Rochan, Z. Liu, and Y. Wang, “Cross-modal self-attention network for referring image segmentation,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 10 502–10 511
2019
Cited alongside, same era.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga et al. , “Pytorch: An imperative style, high-performance deep learning library,” Advances in neural information processing systems , vol. 32, 2019
2019
Cited alongside, same era.
Z. Hu, G. Feng, J. Sun, L. Zhang, and H. Lu, “Bi-directional relationship inferring network for referring image segmentation,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 4424–4433
2020
Cited alongside, same era.
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz et al. , “Transformers: State-of-the-art natural language processing,” in Proceedings of the 2020 conference on empirical methods in natural language processing: system demonstrations , 2020, pp. 38–45
C. Zhao, B. Qin, S. Feng, W. Zhu, W. Sun, W. Li, and X. Jia, “Hyperspectral image classification with multi-attention transformer and adaptive superpixel segmentation-based active learning,” IEEE Transactions on Image Processing , vol. 32, pp. 3606–3621, 2023
2023
Later among the works it cites.
Z. Zheng, Y. Zhong, J. Wang, A. Ma, and L. Zhang, “Farseg++: Foreground-aware relation network for geospatial object segmentation in high spatial resolution remote sensing imagery,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2023
2023
Later among the works it cites.
C. Liu, H. Ding, Y. Zhang, and X. Jiang, “Multi-modal mutual attention and iterative interaction for referring image segmentation,” IEEE Transactions on Image Processing , vol. 32, pp. 3054–3065, 2023
2023
Later among the works it cites.
Y. Zhan, Z. Xiong, and Y. Yuan, “Rsvg: Exploring data and models for visual grounding on remote sensing data,” IEEE Transactions on Geoscience and Remote Sensing , vol. 61, pp. 1–13, 2023
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
T. Hui, S. Liu, S. Huang, G. Li, S. Yu, F. Zhang, and J. Han, “Linguistic structure guided context modeling for referring image segmentation,” in Proceedings of the European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 59–75
2020
Cited alongside, same era.
S. Huang, T. Hui, S. Liu, G. Li, Y. Wei, J. Han, L. Liu, and B. Li, “Referring image segmentation via cross-modal progressive comprehension,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 10 488–10 497
2020
Cited alongside, same era.
X. Yuan, J. Shi, and L. Gu, “A review of deep learning methods for semantic segmentation of remote sensing imagery,” Expert Systems with Applications , vol. 169, p. 114417, 2021
2021
Cited alongside, same era.
S. Lei and Z. Shi, “Hybrid-scale self-similarity exploitation for remote sensing image super-resolution,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–10, 2021
2021
Cited alongside, same era.
S. Lei, Z. Shi, and W. Mo, “Transformer-based multi-stage enhancement for remote sensing image super-resolution,” IEEE Transactions on Geoscience and Remote Sensing , 2021
2021
Cited alongside, same era.
Y. Jing, T. Kong, W. Wang, L. Wang, L. Li, and T. Tan, “Locate then segment: A strong pipeline for referring image segmentation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 9858–9867
2021
Cited alongside, same era.
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo, “Swin transformer: Hierarchical vision transformer using shifted windows,” in Proceedings of the IEEE/CVF international conference on computer vision , 2021, pp. 10 012–10 022
2021
Cited alongside, same era.
S. Liu, T. Hui, S. Huang, Y. Wei, B. Li, and G. Li, “Cross-modal progressive comprehension for referring segmentation,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 44, no. 9, pp. 4761–4775, 2021
2021
Cited alongside, same era.
Later among the works it cites.
Y. Cho, H. Yu, and S.-J. Kang, “Cross-aware early fusion with stage-divided vision and language transformer encoders for referring image segmentation,” IEEE Transactions on Multimedia , 2023
2023
Later among the works it cites.
Z. Yuan, L. Mou, Y. Hua, and X. X. Zhu, “Rrsis: Referring remote sensing image segmentation,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Closest in time.
C. Qiu, X. Zhang, X. Tong, N. Guan, X. Yi, K. Yang, J. Zhu, and A. Yu, “Few-shot remote sensing image scene classification: Recent advances, new baselines, and future trends,” ISPRS Journal of Photogrammetry and Remote Sensing , vol. 209, pp. 368–382, 2024
2024
Closest in time.
F. Tian, S. Lei, Y. Zhou, J. Cheng, G. Liang, Z. Zou, H.-C. Li, and Z. Shi, “Hirenet: Hierarchical-relation network for few-shot remote sensing image scene classification,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Closest in time.
B. Qin, S. Feng, C. Zhao, B. Xi, W. Li, and R. Tao, “Fdgnet: Frequency disentanglement and data geometry for domain generalization in cross-scene hyperspectral image classification,” IEEE Transactions on Neural Networks and Learning Systems , 2024
2024
Closest in time.
S. Gui, S. Song, R. Qin, and Y. Tang, “Remote sensing object detection in the deep learning era—a review,” Remote Sensing , vol. 16, no. 2, p. 327, 2024
2024
Closest in time.
J. Zhu, J. Zhang, H. Chen, Y. Xie, H. Gu, and H. Lian, “A cross-view intelligent person search method based on multi-feature constraints,” International Journal of Digital Earth , vol. 17, no. 1, p. 2346259, 2024
2024
Closest in time.
M. Wang, L. Gao, L. Ren, X. Sun, and J. Chanussot, “Hyperspectral simultaneous anomaly detection and denoising: Insights from integrative perspective,” IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing , 2024
2024
Closest in time.
S. Cao, D. Feng, S. Liu, W. Xu, H. Chen, Y. Xie, H. Zhang, S. Pirasteh, and J. Zhu, “Bemrf-net: Boundary enhancement and multiscale refinement fusion for building extraction from remote sensing imagery,” IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing , 2024
2024
Closest in time.
Y.-C. Li, S. Lei, N. Liu, H.-C. Li, and Q. Du, “Ida-siamnet: Interactive-and dynamic-aware siamese network for building change detection,” IEEE Transactions on Geoscience and Remote Sensing , 2024
2024
Closest in time.
Y. Xie, N. Zhan, J. Zhu, B. Xu, H. Chen, W. Mao, X. Luo, and Y. Hu, “Landslide extraction from aerial imagery considering context association characteristics,” International Journal of Applied Earth Observation and Geoinformation , vol. 131, p. 103950, 2024
2024
Closest in time.
S. Liu, Y. Ma, X. Zhang, H. Wang, J. Ji, X. Sun, and R. Ji, “Rotated multi-scale interaction network for referring remote sensing image segmentation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 26 658–26 668
2024
Closest in time.
K. Kuckreja, M. S. Danish, M. Naseer, A. Das, S. Khan, and F. S. Khan, “Geochat: Grounded large vision-language model for remote sensing,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 27 831–27 840
2024
Closest in time.