Fetching the paper…
Reading the bibliography…
Current methods for disaster scene interpretation in remote sensing images (RSIs) mostly focus on isolated tasks such as segmentation, detection, or visual question-answering (VQA).
P. Maes, “Modeling adaptive autonomous agents,” Artificial life , vol. 1, no. 1_2, pp. 135–162, 1993
1993
Earlier work this paper cites.
2005
Earlier work this paper cites.
D. Brunner, G. Lemoine, and L. Bruzzone, “Earthquake damage assessment of buildings using vhr optical and sar imagery,” IEEE Transactions on Geoscience and Remote Sensing , vol. 48, no. 5, pp. 2403–2420, 2010
2010
Earlier work this paper cites.
2010
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Computer Vision–ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part V 13 . Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh, “Vqa: Visual question answering,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2425–2433
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
N. Kussul, M. Lavreniuk, S. Skakun, and A. Shelestov, “Deep learning classification of land cover and crop types using remote sensing data,” IEEE Geoscience and Remote Sensing Letters , vol. 14, no. 5, pp. 778–782, 2017
2017
Earlier work this paper cites.
K. Nogueira, O. A. Penatti, and J. A. Dos Santos, “Towards better exploiting convolutional neural networks for remote sensing scene classification,” Pattern Recognition , vol. 61, pp. 539–556, 2017
2017
Earlier work this paper cites.
G. Cheng, J. Han, and X. Lu, “Remote sensing image scene classification: Benchmark and state of the art,” Proceedings of the IEEE , vol. 105, no. 10, pp. 1865–1883, 2017
2017
Earlier work this paper cites.
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia, “Pyramid scene parsing network,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 2881–2890
2017
Earlier work this paper cites.
B. Huang, B. Zhao, and Y. Song, “Urban land-use mapping using a deep convolutional neural network with high spatial resolution multispectral remote sensing imagery,” Remote Sensing of Environment , vol. 214, pp. 73–86, 2018
2018
Earlier work this paper cites.
R. Gupta, B. Goodman, N. Patel, R. Hosfelt, S. Sajeev, E. Heim, J. Doshi, K. Lucas, H. Choset, and M. Gaston, “Creating xbd: A dataset for assessing building damage from satellite imagery,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition workshops , 2019, pp. 10–17
2019
Earlier work this paper cites.
S. Shao, Z. Li, T. Zhang, C. Peng, G. Yu, X. Zhang, J. Li, and J. Sun, “Objects365: A large-scale, high-quality dataset for object detection,” in Proceedings of the IEEE/CVF international conference on computer vision , 2019, pp. 8430–8439
2019
Earlier work this paper cites.
W. Zhang, P. Tang, and L. Zhao, “Remote sensing image scene classification using cnn-capsnet,” Remote Sensing , vol. 11, no. 5, p. 494, 2019
2019
Earlier work this paper cites.
K. Marino, M. Rastegari, A. Farhadi, and R. Mottaghi, “Ok-vqa: A visual question answering benchmark requiring external knowledge,” in Proceedings of the IEEE/cvf conference on computer vision and pattern recognition , 2019, pp. 3195–3204
2019
Earlier work this paper cites.
C. Kyrkou and T. Theocharides, “Emergencynet: Efficient aerial image classification for drone-based emergency monitoring using atrous convolutional feature fusion,” IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing , vol. 13, pp. 1687–1699, 2020
2020
Earlier work this paper cites.
Y. Bai, J. Hu, J. Su, X. Liu, H. Liu, X. He, S. Meng, E. Mas, and S. Koshimura, “Pyramid pooling module-based semi-siamese network: A benchmark model for assessing building damage from xbd satellite imagery datasets,” Remote. Sens. , vol. 12, p. 4055, 2020. [Online]. Available: https://api.semanticscholar.org/CorpusID:230526541
2020
Earlier work this paper cites.
S. Lobry, D. Marcos, J. Murray, and D. Tuia, “Rsvqa: Visual question answering for remote sensing data,” IEEE Transactions on Geoscience and Remote Sensing , vol. 58, no. 12, p. 8555–8566, Dec. 2020. [Online]. Available: http://dx.doi.org/10.1109/TGRS.2020.2988782
2020
Earlier work this paper cites.
A. Kuznetsova, H. Rom, N. Alldrin, J. Uijlings, I. Krasin, J. Pont-Tuset, S. Kamali, S. Popov, M. Malloci, A. Kolesnikov et al. , “The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale,” International journal of computer vision , vol. 128, no. 7, pp. 1956–1981, 2020
2020
Earlier work this paper cites.
L. Zhou, H. Palangi, L. Zhang, H. Hu, J. Corso, and J. Gao, “Unified vision-language pre-training for image captioning and vqa,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 07, 2020, pp. 13 041–13 049
2020
Cited alongside, same era.
Y. Shen, S. Zhu, T. Yang, C. Chen, D. Pan, J. Chen, L. Xiao, and Q. Du, “Bdanet: Multiscale convolutional neural network with cross-directional attention for building damage assessment from satellite images,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–14, 2021
2021
Cited alongside, same era.
M. Rahnemoonfar, T. Chowdhury, A. Sarkar, D. Varshney, M. Yari, and R. R. Murphy, “Floodnet: A high resolution aerial imagery dataset for post flood scene understanding,” IEEE Access , vol. 9, pp. 89 644–89 654, 2021
2021
Cited alongside, same era.
X. Yuan, J. Shi, and L. Gu, “A review of deep learning methods for semantic segmentation of remote sensing imagery,” Expert Systems with Applications , vol. 169, p. 114417, 2021
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
M. Ding, Z. Yang, W. Hong, W. Zheng, C. Zhou, D. Yin, J. Lin, X. Zou, Z. Shao, H. Yang et al. , “Cogview: Mastering text-to-image generation via transformers,” Advances in Neural Information Processing Systems , vol. 34, pp. 19 822–19 835, 2021
2021
Cited alongside, same era.
2022
Cited alongside, same era.
C. Chappuis, V. Zermatten, S. Lobry, B. Le Saux, and D. Tuia, “Prompt–rsvqa: Prompting visual context to a language model for remote sensing visual question answering,” in 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) , 2022, pp. 1371–1380
2022
Cited alongside, same era.
2022
Cited alongside, same era.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen, “LoRA: Low-rank adaptation of large language models,” in International Conference on Learning Representations , 2022. [Online]. Available: https://openreview.net/forum?id=nZeVKeeFYf9
2022
Cited alongside, same era.
M. Rahnemoonfar, T. Chowdhury, and R. Murphy, “Rescuenet: a high resolution uav semantic segmentation dataset for natural disaster damage assessment,” Scientific data , vol. 10, no. 1, p. 913, 2023
2023
Cited alongside, same era.
A. Sarkar, T. Chowdhury, R. R. Murphy, A. Gangopadhyay, and M. Rahnemoonfar, “Sam-vqa: Supervised attention-based visual question answering model for post-disaster damage assessment on remote sensing imagery,” IEEE Transactions on Geoscience and Remote Sensing , vol. 61, pp. 1–16, 2023. [Online]. Available: https://api.semanticscholar.org/CorpusID:258726537
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Later among the works it cites.
2023
Later among the works it cites.
2024
Closest in time.
L. Wang, C. Ma, X. Feng, Z. Zhang, H. Yang, J. Zhang, Z. Chen, J. Tang, X. Chen, Y. Lin et al. , “A survey on large language model based autonomous agents,” Frontiers of Computer Science , vol. 18, no. 6, p. 186345, 2024
2024
Closest in time.
D. Zhao, B. Yuan, Z. Chen, T. Li, Z. Liu, W. Li, and Y. Gao, “Panoptic perception: A novel task and fine-grained dataset for universal remote sensing image interpretation,” IEEE Transactions on Geoscience and Remote Sensing , vol. 62, pp. 1–14, 2024
2024
Closest in time.
C. Feng, D. Danier, F. Zhang, and D. Bull, “Rankdvqa: Deep vqa based on ranking-inspired hybrid training,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2024, pp. 1648–1658
2024
Closest in time.
D. Zhao, J. Lu, and B. Yuan, “See, perceive, and answer: A unified benchmark for high-resolution postdisaster evaluation in remote sensing images,” IEEE Transactions on Geoscience and Remote Sensing , vol. 62, pp. 1–14, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
C. Qian, W. Liu, H. Liu, N. Chen, Y. Dang, J. Li, C. Yang, W. Chen, Y. Su, X. Cong et al. , “Chatdev: Communicative agents for software development,” in Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2024, pp. 15 174–15 186
2024
Closest in time.
Z. Wang, S. Cai, G. Chen, A. Liu, X. S. Ma, and Y. Liang, “Describe, explain, plan and select: interactive planning with llms enables open-world multi-task agents,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
T. Schick, J. Dwivedi-Yu, R. Dessì, R. Raileanu, M. Lomeli, E. Hambro, L. Zettlemoyer, N. Cancedda, and T. Scialom, “Toolformer: Language models can teach themselves to use tools,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Y. Shen, K. Song, X. Tan, D. Li, W. Lu, and Y. Zhuang, “Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
X. Zhu, J. Liang, and A. Hauptmann, “Msnet: A multilevel instance segmentation network for natural disaster damage assessment in aerial videos,” in Proceedings of the IEEE/CVF winter conference on applications of computer vision , 2021, pp. 2023–2032
2032
Closest in time.
H. Chen and Z. Shi, “A spatial-temporal attention-based method and a new dataset for remote sensing image change detection,” Remote Sensing , vol. 12, no. 10, 2020. [Online]. Available: https://www.mdpi.com/2072-4292/12/10/1662
2072
Closest in time.