Fetching the paper…
Reading the bibliography…
Referring Remote Sensing Image Segmentation (RRSIS) aims to segment target objects in remote sensing (RS) images based on textual descriptions.
Unsupervised cross-lingual representation learning at scale
Conneau, A. 2019 · 1911
Earlier work this paper cites.
V-net: Fully convolutional neural networks for volumetric medical image segmentation
Milletari, F.; Navab, N.; and Ahmadi, S.-A. 2016 · 2016
Earlier work this paper cites.
Referring image segmentation via recurrent refinement networks
Li, R.; Li, K.; Kuo, Y.-C.; Shu, M.; Qi, X.; Shen, X.; and Jia, J. 2018 · 2018
Earlier work this paper cites.
Decoupled Weight Decay Regularization
Loshchilov, I.; and Hutter, F. 2018 · 2018
Earlier work this paper cites.
Cross-modal self-attention network for referring image segmentation
Ye, L.; Rochan, M.; Liu, Z.; and Wang, Y. 2019 · 2019
Earlier work this paper cites.
Bi-directional relationship inferring network for referring image segmentation
Hu, Z.; Feng, G.; Sun, J.; Zhang, L.; and Lu, H. 2020 · 2020
Earlier work this paper cites.
Referring image segmentation via cross-modal progressive comprehension
Huang, S.; Hui, T.; Liu, S.; Li, G.; Wei, Y.; Han, J.; Liu, L.; and Li, B. 2020 · 2020
Earlier work this paper cites.
Linguistic structure guided context modeling for referring image segmentation
Hui, T.; Liu, S.; Huang, S.; Li, G.; Yu, S.; Zhang, F.; and Han, J. 2020 · 2020
Earlier work this paper cites.
Cross-modal progressive comprehension for referring segmentation
Liu, S.; Hui, T.; Huang, S.; Wei, Y.; Li, B.; and Li, G. 2021 · 2021
Earlier work this paper cites.
Visual grounding in remote sensing images
Sun, Y.; Feng, S.; Li, X.; Ye, Y.; Kang, J.; and Huang, X. 2022 · 2022
Earlier work this paper cites.
Cris: Clip-driven referring image segmentation
Wang, Z.; Lu, Y.; Li, Q.; Tao, X.; Guo, Y.; Gong, M.; and Liu, T. 2022 · 2022
Earlier work this paper cites.
Lavt: Language-aware vision transformer for referring image segmentation
Yang, Z.; Wang, J.; Tang, Y.; Chen, K.; Zhao, H.; and Torr, P. H. 2022 · 2022
Earlier work this paper cites.
Cheng, Y.; Li, L.; Xu, Y.; Li, X.; Yang, Z.; Wang, W.; and Yang, Y. 2023 · 2023
Earlier work this paper cites.
Beyond one-to-one: Rethinking the referring image segmentation
Hu, Y.; Wang, Q.; Shao, W.; Xie, E.; Li, Z.; Han, J.; and Luo, P. 2023 · 2023
Earlier work this paper cites.
Segment anything
Kirillov, A.; Mintun, E.; Ravi, N.; Mao, H.; Rolland, C.; Gustafson, L.; Xiao, T.; Whitehead, S.; Berg, A. C.; Lo, W.-Y.; et al. 2023 · 2023
Cited alongside, same era.
Learning to learn better for video object segmentation
Lan, M.; Zhang, J.; Zhang, L.; and Tao, D. 2023 · 2023
Cited alongside, same era.
Hiera: A hierarchical vision transformer without the bells-and-whistles
Ryali, C.; Hu, Y.-T.; Bolya, D.; Wei, C.; Fan, H.; Huang, P.-Y.; Aggarwal, V.; Chowdhury, A.; Poursaeed, O.; Hoffman, J.; et al. 2023 · 2023
Cited alongside, same era.
Image as a foreign language: Beit pretraining for vision and vision-language tasks
Wang, W.; Bao, H.; Dong, L.; Bjorck, J.; Peng, Z.; Liu, Q.; Aggarwal, K.; Mohammed, O. K.; Singhal, S.; Som, S.; et al. 2023 · 2023
Cited alongside, same era.
Medical sam adapter: Adapting segment anything model for medical image segmentation
Wu, J.; Ji, W.; Liu, Y.; Fu, H.; Xu, M.; Xu, Y.; and Jin, Y. 2023 · 2023
Cited alongside, same era.
Glamm: Pixel grounding large multimodal model
Rasheed, H.; Maaz, M.; Shaji, S.; Shaker, A.; Khan, S.; Cholakkal, H.; Anwer, R. M.; Xing, E.; Yang, M.-H.; and Khan, F. S. 2024 · 2024
Later among the works it cites.
Sam 2: Segment anything in images and videos
Ravi, N.; Gabeur, V.; Hu, Y.-T.; Hu, R.; Ryali, C.; Ma, T.; Khedr, H.; Rädle, R.; Rolland, C.; Gustafson, L.; et al. 2024 · 2024
Later among the works it cites.
Efficientsam: Leveraged masked image pretraining for efficient segment anything
Xiong, Y.; Varadarajan, B.; Wu, L.; Xiang, X.; Xiao, F.; Zhu, C.; Dai, X.; Wang, D.; Sun, F.; Iandola, F.; et al. 2024 · 2024
Later among the works it cites.
Visa: Reasoning video object segmentation via large language models
Yan, C.; Wang, H.; Yan, S.; Jiang, X.; Hu, Y.; Kang, G.; Xie, W.; and Gavves, E. 2024 · 2024
Later among the works it cites.
A Simple Baseline with Single-encoder for Referring Image Segmentation
Yu, S.; Jung, I.; Han, B.; Kim, T.; Kim, Y.; Wee, D.; and Son, J. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xu, J.; Xu, L.; Yang, Y.; Li, X.; Wang, F.; Xie, Y.; Huang, Y.-J.; and Li, Y. 2023 · 2023
Cited alongside, same era.
Rsvg: Exploring data and models for visual grounding on remote sensing data
Zhan, Y.; Xiong, Z.; and Yuan, Y. 2023 · 2023
Cited alongside, same era.
Faster segment anything: Towards lightweight sam for mobile applications
Zhang, C.; Han, D.; Qiao, Y.; Kim, J. U.; Bae, S.-H.; Lee, S.; and Hong, C. S. 2023 · 2023
Cited alongside, same era.
Chen, T.; Lu, A.; Zhu, L.; Ding, C.; Yu, C.; Ji, D.; Li, Z.; Sun, L.; Mao, P.; and Zang, Y. 2024 · 2024
Cited alongside, same era.
Unleashing the potential of SAM for medical adaptation via hierarchical decoding
Cheng, Z.; Wei, Q.; Zhu, H.; Wang, Y.; Qu, L.; Shao, W.; and Zhou, Y. 2024 · 2024
Cited alongside, same era.
Segment anything in high quality
Ke, L.; Ye, M.; Danelljan, M.; Tai, Y.-W.; Tang, C.-K.; Yu, F.; et al. 2024 · 2024
Cited alongside, same era.
Lisa: Reasoning segmentation via large language model
Lai, X.; Tian, Z.; Chen, Y.; Li, Y.; Yuan, Y.; Liu, S.; and Jia, J. 2024 · 2024
Cited alongside, same era.
Later among the works it cites.
Rrsis: Referring remote sensing image segmentation
Yuan, Z.; Mou, L.; Hua, Y.; and Zhu, X. X. 2024 · 2024
Later among the works it cites.
Surgicalsam: Efficient class promptable surgical instrument segmentation
Yue, W.; Zhang, J.; Hu, K.; Xia, Y.; Luo, J.; and Wang, Z. 2024 · 2024
Later among the works it cites.
Evf-sam: Early vision-language fusion for text-prompted segment anything model
Zhang, Y.; Cheng, T.; Hu, R.; Liu, L.; Liu, H.; Ran, L.; Chen, X.; Liu, W.; and Wang, X. 2024 · 2024
Later among the works it cites.
Convolution meets lora: Parameter efficient finetuning for segment anything model
Zhong, Z.; Tang, Z.; He, T.; Fang, H.; and Yuan, C. 2024 · 2024
Later among the works it cites.
Rong, F.; Lan, M.; Zhang, Q.; and Zhang, L. 2025 · 2025
Closest in time.
PCP: A Prompt-Based Cartographic-Level Polygonal Vector Extraction Framework for Remote Sensing Images
Wang, C.; Xi, Z.; Liu, D.; Feng, Y.; Deng, Y.; Li, K.; Chen, J.; Chen, J.; and Meng, Y. 2025b · 2025
Closest in time.
RS-OOD: A Vision-Language Augmented Framework for Out-of-Distribution Detection in Remote Sensing
Yingrui Ji, J. C. A. Y. C. W. K. L. Y. Z., Jiansheng Chen. 2025 · 2025
Closest in time.
UniUIR: Considering Underwater Image Restoration as An All-in-One Learner
Zhang, X.; Zhang, H.; Wang, G.; Zhang, Q.; Zhang, L.; and Du, B. 2025 · 2025
Closest in time.