Fetching the paper…
Reading the bibliography…
Localizing desired objects from remote sensing images is of great use in practical applications.
“Long short-term memory,”
Sepp Hochreiter and Jürgen Schmidhuber, · 1997
Earlier work this paper cites.
“Color-and texture-based image segmentation using em and its application to content-based image retrieval,”
Serge Belongie, Chad Carson, Hayit Greenspan, and Jitendra Malik, · 1998
Earlier work this paper cites.
“The watershed transform: Definitions, algorithms and parallelization strategies,”
Jos BTM Roerdink and Arnold Meijster, · 2000
Earlier work this paper cites.
“Unsupervised segmentation of color-texture regions in images and video,”
Yining Deng and Bangalore S Manjunath, · 2001
Earlier work this paper cites.
“Fast approximate energy minimization via graph cuts,”
Yuri Boykov, Olga Veksler, and Ramin Zabih, · 2001
Earlier work this paper cites.
“Evaluating bag-of-visual-words representations in scene classification,”
Jun Yang, Yu-Gang Jiang, Alexander G Hauptmann, and Chong-Wah Ngo, · 2007
Earlier work this paper cites.
“Bag-of-visual-words and spatial extensions for land-use classification,”
Yi Yang and Shawn Newsam, · 2010
Earlier work this paper cites.
“Automatic target detection in high-resolution remote sensing images using spatial sparse coding bag-of-words model,”
Hao Sun, Xian Sun, Hongqi Wang, Yu Li, and Xiangjuan Li, · 2011
Earlier work this paper cites.
“Slic superpixels compared to state-of-the-art superpixel methods,”
Radhakrishna Achanta, Appu Shaji, Kevin Smith, Aurelien Lucchi, Pascal Fua, and Sabine Süsstrunk, · 2012
Earlier work this paper cites.
“Very deep convolutional networks for large-scale image recognition,”
Karen Simonyan and Andrew Zisserman, · 2014
Earlier work this paper cites.
“Fast r-cnn,”
Ross Girshick, · 2015
Earlier work this paper cites.
“Fully convolutional networks for semantic segmentation,”
Jonathan Long, Evan Shelhamer, and Trevor Darrell, · 2015
Earlier work this paper cites.
“U-net: Convolutional networks for biomedical image segmentation,”
Olaf Ronneberger, Philipp Fischer, and Thomas Brox, · 2015
Earlier work this paper cites.
“Batch normalization: Accelerating deep network training by reducing internal covariate shift,”
Sergey Ioffe and Christian Szegedy, · 2015
Earlier work this paper cites.
“Deep residual learning for image recognition,”
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun, · 2016
Earlier work this paper cites.
“You only look once: Unified, real-time object detection,”
Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi, · 2016
Earlier work this paper cites.
“Segmentation from natural language expressions,”
Ronghang Hu, Marcus Rohrbach, and Trevor Darrell, · 2016
Earlier work this paper cites.
“Deep semantic understanding of high resolution remote sensing image,”
Bo Qu, Xuelong Li, Dacheng Tao, and Xiaoqiang Lu, · 2016
Earlier work this paper cites.
“Gaussian error linear units (gelus),”
Dan Hendrycks and Kevin Gimpel, · 2016
Earlier work this paper cites.
“Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,”
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L Yuille, · 2017
Earlier work this paper cites.
“Recurrent multimodal interaction for referring image segmentation,”
Chenxi Liu, Zhe Lin, Xiaohui Shen, Jimei Yang, Xin Lu, and Alan Yuille, · 2017
Earlier work this paper cites.
“Exploring models and data for remote sensing image caption generation,”
Xiaoqiang Lu, Binqiang Wang, Xiangtao Zheng, and Xuelong Li, · 2017
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin, · 2017
Cited alongside, same era.
“Referring image segmentation via recurrent refinement networks,”
Ruiyu Li, Kaican Li, Yi-Chun Kuo, Michelle Shu, Xiaojuan Qi, Xiaoyong Shen, and Jiaya Jia, · 2018
Cited alongside, same era.
“Dynamic multimodal instance segmentation guided by natural language queries,”
Edgar Margffoy-Tuay, Juan C Pérez, Emilio Botero, and Pablo Arbeláez, · 2018
Cited alongside, same era.
“Key-word-aware network for referring expression image segmentation,”
Hengcan Shi, Hongliang Li, Fanman Meng, and Qingbo Wu, · 2018
Cited alongside, same era.
“Encoder fusion network with co-attention embedding for referring image segmentation,”
Guang Feng, Zhiwei Hu, Lihe Zhang, and Huchuan Lu, · 2021
Later among the works it cites.
“Learning transferable visual models from natural language supervision,”
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al., · 2021
Later among the works it cites.
“RSVQA meets BigEarthNet: A new, large-scale, visual question answering dataset for remote sensing,”
Sylvain Lobry, Begüm Demir, and Devis Tuia, · 2021
Later among the works it cites.
“Mutual attention inception network for remote sensing visual question answering,”
Xiangtao Zheng, Binqiang Wang, Xingqian Du, and Xiaoqiang Lu, · 2021
Later among the works it cites.
“High-resolution remote sensing image captioning based on structured attention,”
Rui Zhao, Zhenwei Shi, and Zhengxia Zou, · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova, · 2018
Cited alongside, same era.
“Decoupled weight decay regularization,”
Ilya Loshchilov and Frank Hutter, · 2018
Cited alongside, same era.
“Cross-modal self-attention network for referring image segmentation,”
Linwei Ye, Mrigank Rochan, Zhi Liu, and Yang Wang, · 2019
Cited alongside, same era.
“Skyscapes fine-grained semantic understanding of aerial scenes,”
Seyed Majid Azimi, Corentin Henry, Lars Sommer, Arne Schumann, and Eleonora Vig, · 2019
Cited alongside, same era.
“BigEarthNet: A large-scale benchmark archive for remote sensing image understanding,”
Gencer Sumbul, Marcela Charfuelan, Begüm Demir, and Volker Markl, · 2019
Cited alongside, same era.
“Lam: Remote sensing image captioning with label-attention mechanism,”
Zhengyuan Zhang, Wenhui Diao, Wenkai Zhang, Menglong Yan, Xin Gao, and Xian Sun, · 2019
Cited alongside, same era.
“Phrasecut: Language-based image segmentation in the wild,”
Chenyun Wu, Zhe Lin, Scott Cohen, Trung Bui, and Subhransu Maji, · 2020
Cited alongside, same era.
“Recurrent attention and semantic gate for remote sensing image captioning,”
Yunpeng Li, Xiangrong Zhang, Jing Gu, Chen Li, Xin Wang, Xu Tang, and Licheng Jiao, · 2021
Later among the works it cites.
“A convnet for the 2020s,”
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie, · 2022
Later among the works it cites.
“LAVT: Language-aware vision transformer for referring image segmentation,”
Zhao Yang, Jiaqi Wang, Yansong Tang, Kai Chen, Hengshuang Zhao, and Philip HS Torr, · 2022
Later among the works it cites.
“CRIS: Clip-driven referring image segmentation,”
Zhaoqing Wang, Yu Lu, Qiang Li, Xunqiang Tao, Yandong Guo, Mingming Gong, and Tongliang Liu, · 2022
Later among the works it cites.
“Visual grounding in remote sensing images,”
Yuxi Sun, Shanshan Feng, Xutao Li, Yunming Ye, Jian Kang, and Xu Huang, · 2022
Later among the works it cites.
“Bi-modal transformer-based approach for visual question answering in remote sensing imagery,”
Yakoub Bazi, Mohamad Mahmoud Al Rahhal, Mohamed Lamine Mekhalfi, Mansour Abdulaziz Al Zuair, and Farid Melgani, · 2022
Later among the works it cites.
“Prompt-RSVQA: Prompting visual context to a language model for remote sensing visual question answering,”
Christel Chappuis, Valérie Zermatten, Sylvain Lobry, Bertrand Le Saux, and Devis Tuia, · 2022
Later among the works it cites.
“Change detection meets visual question answering,”
Zhenghang Yuan, Lichao Mou, Zhitong Xiong, and Xiao Xiang Zhu, · 2022
Later among the works it cites.
“Exploring transformer and multilabel classification for remote sensing image captioning,”
Hitesh Kandala, Sudipan Saha, Biplab Banerjee, and Xiao Xiang Zhu, · 2022
Later among the works it cites.
“Remote-sensing image captioning based on multilayer aggregated transformer,”
Chenyang Liu, Rui Zhao, and Zhenwei Shi, · 2022
Later among the works it cites.
“Polyformer: Referring image segmentation as sequential polygon generation,”
Jiang Liu, Hui Ding, Zhaowei Cai, Yuting Zhang, Ravi Kumar Satzoda, Vijay Mahadevan, and R Manmatha, · 2023
Closest in time.
“Multi-modal mutual attention and iterative interaction for referring image segmentation,”
Chang Liu, Henghui Ding, Yulun Zhang, and Xudong Jiang, · 2023
Closest in time.
“RSVG: Exploring data and models for visual grounding on remote sensing data,”
Yang Zhan, Zhitong Xiong, and Yuan Yuan, · 2023
Closest in time.
“A spatial hierarchical reasoning network for remote sensing visual question answering,”
Zixiao Zhang, Licheng Jiao, Lingling Li, Xu Liu, Puhua Chen, Fang Liu, Yuxuan Li, and Zhicheng Guo, · 2023
Closest in time.
“Multi-source interactive stair attention for remote sensing image captioning,”
Xiangrong Zhang, Yunpeng Li, Xin Wang, Feixiang Liu, Zhaoji Wu, Xina Cheng, and Licheng Jiao, · 2023
Closest in time.