Fetching the paper…
Reading the bibliography…
Aerial Image Segmentation is a top-down perspective semantic segmentation and has several challenging characteristics such as strong imbalance in the foreground-background distribution, complex background, intra-class heterogeneity, inter-class homogeneity, and tiny objects.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
On the use of imagery for climate change engagement
Saffron J O’neill, Maxwell Boykoff, Simon Niemeyer, and Sophie A Day · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
Jonathan Long, Evan Shelhamer, and Trevor Darrell · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Earlier work this paper cites.
Semantic segmentation with boundary neural fields
Gedas Bertasius, Jianbo Shi, and Lorenzo Torresani · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Gaussian error linear units (gelus)
Dan Hendrycks and Kevin Gimpel · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna · 2016
Earlier work this paper cites.
Segnet: A deep convolutional encoder-decoder architecture for image segmentation
Vijay Badrinarayanan, Alex Kendall, and Roberto Cipolla · 2017
Earlier work this paper cites.
Linknet: Exploiting encoder representations for efficient semantic segmentation
Abhishek Chaurasia and Eugenio Culurciello · 2017
Earlier work this paper cites.
Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L Yuille · 2017
Earlier work this paper cites.
Rethinking atrous convolution for semantic image segmentation
Liang-Chieh Chen, George Papandreou, Florian Schroff, and Hartwig Adam · 2017
Earlier work this paper cites.
Deformable convolutional networks
Jifeng Dai, Haozhi Qi, Yuwen Xiong, Yi Li, Guodong Zhang, Han Hu, and Yichen Wei · 2017
Earlier work this paper cites.
Segmentation-aware convolutional networks using local attention masks
Adam W Harley, Konstantinos G Derpanis, and Iasonas Kokkinos · 2017
Earlier work this paper cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Andrew G Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam · 2017
Earlier work this paper cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke, and Alexander Alemi · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Pyramid scene parsing network
Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, and Jiaya Jia · 2017
Earlier work this paper cites.
Encoder-decoder with atrous separable convolution for semantic image segmentation
Liang-Chieh Chen, Yukun Zhu, George Papandreou, Florian Schroff, and Hartwig Adam · 2018
Earlier work this paper cites.
Squeeze-and-excitation networks
Jie Hu, Li Shen, and Gang Sun · 2018
Earlier work this paper cites.
Deep contextual recurrent residual networks for scene labeling
T Hoang Ngan Le, Chi Nhan Duong, Ligong Han, Khoa Luu, Kha Gia Quach, and Marios Savvides · 2018
Earlier work this paper cites.
Pyramid attention network for semantic segmentation
Hanchao Li, Pengfei Xiong, Jie An, and Lingxue Wang · 2018
Earlier work this paper cites.
Land cover mapping at very high resolution with rotation equivariant cnns: Towards small yet accurate models
Diego Marcos, Michele Volpi, Benjamin Kellenberger, and Devis Tuia · 2018
Earlier work this paper cites.
Assisting flood disaster response with earth observation data and products: A critical assessment
Guy JP Schumann, G Robert Brakenridge, Albert J Kettner, Rashid Kashif, and Emily Niebuhr · 2018
Earlier work this paper cites.
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
Earlier work this paper cites.
Unified perceptual parsing for scene understanding
Tete Xiao, Yingcheng Liu, Bolei Zhou, Yuning Jiang, and Jian Sun · 2018
Earlier work this paper cites.
Denseaspp for semantic segmentation in street scenes
Maoke Yang, Kun Yu, Chi Zhang, Zhiwei Li, and Kuiyuan Yang · 2018
Earlier work this paper cites.
Context encoding for semantic segmentation
Hang Zhang, Kristin Dana, Jianping Shi, Zhongyue Zhang, Xiaogang Wang, Ambrish Tyagi, and Amit Agrawal · 2018
Earlier work this paper cites.
Psanet: Point-wise spatial attention network for scene parsing
Hengshuang Zhao, Yi Zhang, Shu Liu, Jianping Shi, Chen Change Loy, Dahua Lin, and Jiaya Jia · 2018
Earlier work this paper cites.
Unet++: A nested u-net architecture for medical image segmentation
Zongwei Zhou, Md Mahfuzur Rahman Siddiquee, Nima Tajbakhsh, and Jianming Liang · 2018
Earlier work this paper cites.
Boundary-aware feature propagation for scene segmentation
Henghui Ding, Xudong Jiang, Ai Qun Liu, Nadia Magnenat Thalmann, and Gang Wang · 2019
Earlier work this paper cites.
Dual attention network for scene segmentation
Jun Fu, Jing Liu, Haijie Tian, Yong Li, Yongjun Bao, Zhiwei Fang, and Hanqing Lu · 2019
Earlier work this paper cites.
Improving public data for building segmentation from convolutional neural networks (cnns) for fused airborne lidar and image data using active contours
David Griffiths and Jan Boehm · 2019
Cited alongside, same era.
Dynamic multi-scale filters for semantic segmentation
Junjun He, Zhongying Deng, and Yu Qiao · 2019
Cited alongside, same era.
Adaptive pyramid context network for semantic segmentation
Junjun He, Zhongying Deng, Lei Zhou, Yali Wang, and Yu Qiao · 2019
Cited alongside, same era.
Ccnet: Criss-cross attention for semantic segmentation
Zilong Huang, Xinggang Wang, Lichao Huang, Chang Huang, Yunchao Wei, and Wenyu Liu · 2019
Cited alongside, same era.
Expectation-maximization attention networks for semantic segmentation
Xia Li, Zhisheng Zhong, Jianlong Wu, Yibo Yang, Zhouchen Lin, and Hong Liu · 2019
Cited alongside, same era.
Deep high-resolution representation learning for human pose estimation
Abcnet: Attentive bilateral contextual network for efficient semantic segmentation of fine-resolution remotely sensed imagery
Rui Li, Shunyi Zheng, Ce Zhang, Chenxi Duan, Libo Wang, and Peter M Atkinson · 2021
Later among the works it cites.
Pointflow: Flowing semantics through points for aerial image segmentation
Xiangtai Li, Hao He, Xia Li, Duo Li, Guangliang Cheng, Jianping Shi, Lubin Weng, Yunhai Tong, and Zhouchen Lin · 2021
Later among the works it cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Later among the works it cites.
On creating benchmark dataset for aerial image interpretation: Reviews, guidances, and million-aid
Yang Long, Gui-Song Xia, Shengyang Li, Wen Yang, Michael Ying Yang, Xiao Xiang Zhu, Liangpei Zhang, and Deren Li · 2021
Later among the works it cites.
Image segmentation using deep learning: A survey
Shervin Minaee, Yuri Y Boykov, Fatih Porikli, Antonio J Plaza, Nasser Kehtarnavaz, and Demetri Terzopoulos · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ke Sun, Bin Xiao, Dong Liu, and Jingdong Wang · 2019
Cited alongside, same era.
isaid: A large-scale dataset for instance segmentation in aerial images
Syed Waqas Zamir, Aditya Arora, Akshita Gupta, Salman Khan, Guolei Sun, Fahad Shahbaz Khan, Fan Zhu, Ling Shao, Gui-Song Xia, and Xiang Bai · 2019
Cited alongside, same era.
Cross-modal self-attention network for referring image segmentation
Linwei Ye, Mrigank Rochan, Zhi Liu, and Yang Wang · 2019
Cited alongside, same era.
Pixel-level remote sensing image recognition based on bidirectional word vectors
Hongfeng You, Shengwei Tian, Long Yu, and Yalong Lv · 2019
Cited alongside, same era.
Segmentation transformer: Object-contextual representations for semantic segmentation
Yuhui Yuan, Xiaokang Chen, Xilin Chen, and Jingdong Wang · 2019
Cited alongside, same era.
Evaluation of semantic segmentation methods for deforestation detection in the amazon
RB Andrade, GAOP Costa, GLA Mota, MX Ortega, RQ Feitosa, PJ Soto, and Christian Heipke · 2020
Cited alongside, same era.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Cited alongside, same era.
Hybrid multiple attention network for semantic segmentation in aerial images
Ruigang Niu, Xian Sun, Yu Tian, Wenhui Diao, Kaiqiang Chen, and Kun Fu · 2021
Later among the works it cites.
Segmenter: Transformer for semantic segmentation
Robin Strudel, Ricardo Garcia, Ivan Laptev, and Cordelia Schmid · 2021
Later among the works it cites.
Sparse r-cnn: End-to-end object detection with learnable proposals
Peize Sun, Rufeng Zhang, Yi Jiang, Tao Kong, Chenfeng Xu, Wei Zhan, Masayoshi Tomizuka, Lei Li, Zehuan Yuan, Changhu Wang, et al · 2021
Later among the works it cites.
Training data-efficient image transformers & distillation through attention
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Hervé Jégou · 2021
Later among the works it cites.
AEI: Actors-Environment Interaction with Adaptive Attention for Temporal Action Proposals Generation
Khoa Vo, Hyekang Joo, Kashu Yamazaki, Sang Truong, Kris Kitani, Minh-Triet Tran, and Ngan Le · 2021
Later among the works it cites.
Loveda: A remote sensing land-cover dataset for domain adaptive semantic segmentation
Junjue Wang, Zhuo Zheng, Ailong Ma, Xiaoyan Lu, and Yanfei Zhong · 2021
Later among the works it cites.
Hierarchical human semantic parsing with comprehensive part-relation modeling
Wenguan Wang, Tianfei Zhou, Siyuan Qi, Jianbing Shen, and Song-Chun Zhu · 2021
Later among the works it cites.
Segformer: Simple and efficient design for semantic segmentation with transformers
Enze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar, Jose M Alvarez, and Ping Luo · 2021
Later among the works it cites.
Rest: An efficient transformer for visual recognition
Qinglong Zhang and Yu-Bin Yang · 2021
Later among the works it cites.
Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers
Sixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu, Zekun Luo, Yabiao Wang, Yanwei Fu, Jianfeng Feng, Tao Xiang, Philip HS Torr, et al · 2021
Later among the works it cites.
Masked-attention mask transformer for universal image segmentation
Bowen Cheng, Ishan Misra, Alexander G Schwing, Alexander Kirillov, and Rohit Girdhar · 2022
Later among the works it cites.
Swin transformer embedding unet for remote sensing image semantic segmentation
Xin He, Yong Zhou, Jiaqi Zhao, Di Zhang, Rui Yao, and Yong Xue · 2022
Later among the works it cites.
Dam-al: Dilated attention mechanism with attention loss for 3d infant brain image segmentation
Dinh-Hieu Hoang, Gia-Han Diep, Minh-Triet Tran, and Ngan T H Le · 2022
Later among the works it cites.
Bsnet: Dynamic hybrid gradient convolution based boundary-sensitive network for remote sensing image segmentation
Jianlong Hou, Zhi Guo, Youming Wu, Wenhui Diao, and Tao Xu · 2022
Later among the works it cites.
Dn-detr: Accelerate detr training by introducing query denoising
Feng Li, Hao Zhang, Shilong Liu, Jian Guo, Lionel M Ni, and Lei Zhang · 2022
Later among the works it cites.
Dab-detr: Dynamic anchor boxes are better queries for detr
Shilong Liu, Feng Li, Hao Zhang, Xiao Yang, Xianbiao Qi, Hang Su, Jun Zhu, and Lei Zhang · 2022
Later among the works it cites.
Factseg: Foreground activation-driven small object semantic segmentation in large-scale remote sensing imagery
Ailong Ma, Junjue Wang, Yanfei Zhong, and Zhuo Zheng · 2022
Later among the works it cites.
Deep learning-based change detection in remote sensing images: A review
Ayesha Shafique, Guo Cao, Zia Khan, Muhammad Asad, and Muhammad Aslam · 2022
Later among the works it cites.
Ringmo: A remote sensing foundation model with masked image modeling
Xian Sun, Peijin Wang, Wanxuan Lu, Zicong Zhu, Xiaonan Lu, Qibin He, Junxi Li, Xuee Rong, Zhujun Yang, Hao Chang, et al · 2022
Later among the works it cites.
Aisformer: Amodal instance segmentation with transformer
Minh Tran, Khoa Vo, Kashu Yamazaki, Arthur Fernandes, Michael Kidd, and Ngan Le · 2022
Later among the works it cites.
Aoe-net: Entities interactions modeling with adaptive attention mechanism for temporal action proposals generation
Khoa Vo, Sang Truong, Kashu Yamazaki, Bhiksha Raj, Minh-Triet Tran, and Ngan Le · 2022
Later among the works it cites.
An empirical study of remote sensing pretraining
Di Wang, Jing Zhang, Bo Du, Gui-Song Xia, and Dacheng Tao · 2022
Later among the works it cites.
Advancing plain vision transformer towards remote sensing foundation model
Di Wang, Qiming Zhang, Yufei Xu, Jing Zhang, Bo Du, Dacheng Tao, and Liangpei Zhang · 2022
Later among the works it cites.
A novel transformer based semantic segmentation scheme for fine-resolution remote sensing images
Libo Wang, Rui Li, Chenxi Duan, Ce Zhang, Xiaoliang Meng, and Shenghui Fang · 2022
Later among the works it cites.
Unetformer: A unet-like transformer for efficient semantic segmentation of remote sensing urban scene imagery
Libo Wang, Rui Li, Ce Zhang, Shenghui Fang, Chenxi Duan, Xiaoliang Meng, and Peter M Atkinson · 2022
Later among the works it cites.
Aanet: an attention-based alignment semantic segmentation network for high spatial resolution remote sensing images
Gunagkuo Xue, Yikun Liu, Yuwen Huang, Mingsong Li, and Gongping Yang · 2022
Later among the works it cites.
Vlcap: Vision-language with contrastive learning for coherent video paragraph captioning
Kashu Yamazaki, Sang Truong, Khoa Vo, Michael Kidd, Chase Rainwater, Khoa Luu, and Ngan Le · 2022
Later among the works it cites.
Openearthmap: A benchmark dataset for global high-resolution land cover mapping
Junshi Xia, Naoto Yokoya, Bruno Adriano, and Clifford Broni-Bediako · 2023
Closest in time.
Rssformer: Foreground saliency enhancement for remote sensing land-cover segmentation
Rongtao Xu, Changwei Wang, Jiguang Zhang, Shibiao Xu, Weiliang Meng, and Xiaopeng Zhang · 2023
Closest in time.
Vltint: Visual-linguistic transformer-in-transformer for coherent video paragraph captioning
Kashu Yamazaki, Khoa Vo, Sang Truong, Bhiksha Raj, and Ngan Le · 2023
Closest in time.