Fetching the paper…
Reading the bibliography…
Large vision and language assistants have enabled new capabilities for interpreting natural images.
Bag-of-visual-words and spatial extensions for land-use classification
Yi Yang and Shawn Newsam · 2010
Earlier work this paper cites.
Urban change detection based on remote sensing and gis study of salem revenue division, salem district, tamil nadu, india
Shanmugam Tamilenthi and Rajagopalan Baskaran · 2013
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Remote sensing change detection methods to track deforestation and growth in threatened rainforests in madre de dios, peru
Jacob Shermeyer and Barry Haack · 2015
Earlier work this paper cites.
Remote sensing change detection for ecological monitoring in united states protected areas
Katherine S Willis · 2015
Earlier work this paper cites.
Deep semantic understanding of high resolution remote sensing image
Bo Qu, Xuelong Li, Dacheng Tao, and Xiaoqiang Lu · 2016
Earlier work this paper cites.
Remote sensing image scene classification: Benchmark and state of the art
Gong Cheng, Junwei Han, and Xiaoqiang Lu · 2017
Earlier work this paper cites.
Damage detection from aerial images via convolutional neural networks
Aito Fujita, Ken Sakurada, Tomoyuki Imaizumi, Riho Ito, Shuhei Hikosaka, and Ryosuke Nakamura · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Exploring models and data for remote sensing image caption generation
Xiaoqiang Lu, Binqiang Wang, Xiangtao Zheng, and Xuelong Li · 2017
Earlier work this paper cites.
Current remote sensing approaches to monitoring forest degradation in support of countries measurement, reporting and verification (mrv) systems for redd+
Anthea L Mitchell, Ake Rosenqvist, and Brice Mora · 2017
Earlier work this paper cites.
Aid: A benchmark data set for performance evaluation of aerial scene classification
Gui-Song Xia, Jingwen Hu, Fan Hu, Baoguang Shi, Xiang Bai, Yanfei Zhong, Liangpei Zhang, and Xiaoqiang Lu · 2017
Earlier work this paper cites.
Functional map of the world
Gordon Christie, Neil Fendley, James Wilson, and Ryan Mukherjee · 2018
Earlier work this paper cites.
Fully convolutional siamese networks for change detection
Rodrigo Caye Daudt, Bertr Le Saux, and Alexandre Boulch · 2018
Earlier work this paper cites.
Dota: A large-scale dataset for object detection in aerial images
Gui-Song Xia, Xiang Bai, Jian Ding, Zhen Zhu, Serge Belongie, Jiebo Luo, Mihai Datcu, Marcello Pelillo, and Liangpei Zhang · 2018
Earlier work this paper cites.
Multitask learning for large-scale semantic change detection
Rodrigo Caye Daudt, Bertrand Le Saux, Alexandre Boulch, and Yann Gousseau · 2019
Earlier work this paper cites.
Creating xbd: A dataset for assessing building damage from satellite imagery
Ritwik Gupta, Bryce Goodman, Nirav Patel, Ricky Hosfelt, Sandra Sajeev, Eric Heim, Jigar Doshi, Keane Lucas, Howie Choset, and Matthew Gaston · 2019
Earlier work this paper cites.
Sound active attention framework for remote sensing image captioning
Xiaoqiang Lu, Binqiang Wang, and Xiangtao Zheng · 2019
Earlier work this paper cites.
Detecting ecological changes with a remote sensing based ecological index (rsei) produced time series and change vector analysis
Hanqiu Xu, Yifan Wang, Huade Guan, Tingting Shi, and Xisheng Hu · 2019
Earlier work this paper cites.
Near real-time wildfire progression monitoring with sentinel-1 sar time series and deep learning
Yifang Ban, Puzhao Zhang, Andrea Nascetti, Alexandre R Bevington, and Michael A Wulder · 2020
Earlier work this paper cites.
Change detection of deforestation in the brazilian amazon using landsat data and convolutional neural networks
Pablo Pozzobon De Bem, Osmar Abílio de Carvalho Junior, Renato Fontes Guimarães, and Roberto Arnaldo Trancoso Gomes · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Earlier work this paper cites.
Improved mapping and change detection of the start of the crop growing season in the us corn belt from long-term avhrr ndvi
Hyeon-Ju Gim, Chang-Hoi Ho, Sujong Jeong, Jinwon Kim, Song Feng, and Michael J Hayes · 2020
Earlier work this paper cites.
Object detection in optical remote sensing images: A survey and a new benchmark
Ke Li, Gang Wan, Gong Cheng, Liqiu Meng, and Junwei Han · 2020
Earlier work this paper cites.
Rsvqa: Visual question answering for remote sensing data
Sylvain Lobry, Diego Marcos, Jesse Murray, and Devis Tuia · 2020
Earlier work this paper cites.
Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters
Jeff Rasley, Samyam Rajbhandari, Olatunji Ruwase, and Yuxiong He · 2020
Earlier work this paper cites.
A systematic review and assessment of algorithms to detect, characterize, and monitor urban land change
Meredith Reba and Karen C Seto · 2020
Cited alongside, same era.
Geography-aware self-supervised learning
Kumar Ayush, Burak Uzkent, Chenlin Meng, Kumar Tanmay, Marshall Burke, David Lobell, and Stefano Ermon · 2021
Cited alongside, same era.
A review on change detection method and accuracy assessment for land use land cover
Ali Hassan Chughtai, Habibullah Abbasi, and Ismail Rakip Karas · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Cited alongside, same era.
Seasonal contrast: Unsupervised pre-training from uncurated remote sensing data
Oscar Manas, Alexandre Lacoste, Xavier Giró-i Nieto, David Vazquez, and Pau Rodriguez · 2021
Cited alongside, same era.
Geochat: Grounded large vision-language model for remote sensing
Kartik Kuckreja, Muhammad Sohail Danish, Muzammal Naseer, Abhijit Das, Salman Khan, and Fahad Shahbaz Khan · 2023
Later among the works it cites.
Llama-vid: An image is worth 2 tokens in large language models
Yanwei Li, Chengyao Wang, and Jiaya Jia · 2023
Later among the works it cites.
Video-llava: Learning united visual representation by alignment before projection
Bin Lin, Bin Zhu, Yang Ye, Munan Ning, Peng Jin, and Li Yuan · 2023
Later among the works it cites.
Video-chatgpt: Towards detailed video understanding via large vision and language models
Muhammad Maaz, Hanoona Rasheed, Salman Khan, and Fahad Shahbaz Khan · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Application of deep learning architectures for satellite image time series prediction: A review
Waytehad Rose Moskolaï, Wahabou Abdou, Albert Dipanda, and Kolyang · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Floodnet: A high resolution aerial imagery dataset for post flood scene understanding
Maryam Rahnemoonfar, Tashnim Chowdhury, Argho Sarkar, Debvrat Varshney, Masoud Yari, and Robin Roberson Murphy · 2021
Cited alongside, same era.
S2looking: A satellite side-looking dataset for building change detection
Li Shen, Yao Lu, Hao Chen, Hao Wei, Donghai Xie, Jiabao Yue, Rui Chen, Shouye Lv, and Bitao Jiang · 2021
Cited alongside, same era.
Qfabric: Multi-task change detection dataset
Sagar Verma, Akash Panigrahi, and Siddharth Gupta · 2021
Cited alongside, same era.
Asymmetric siamese networks for semantic change detection in aerial images
Kunping Yang, Gui-Song Xia, Zicheng Liu, Bo Du, Wen Yang, Marcello Pelillo, and Liangpei Zhang · 2021
Cited alongside, same era.
Llava-phi: Efficient multi-modal assistant with small language model
Yichen Zhu, Minjie Zhu, Ning Liu, Zhicai Ou, Xiaofeng Mou, and Jian Tang · 2021
Cited alongside, same era.
Hanoona Rasheed, Muhammad Maaz, Sahal Shaji, Abdelrahman Shaker, Salman Khan, Hisham Cholakkal, Rao M Anwer, Erix Xing, Ming-Hsuan Yang, and Fahad S Khan · 2023
Later among the works it cites.
Generative multimodal models are in-context learners
Quan Sun, Yufeng Cui, Xiaosong Zhang, Fan Zhang, Qiying Yu, Zhengxiong Luo, Yueze Wang, Yongming Rao, Jingjing Liu, Tiejun Huang, et al · 2023
Later among the works it cites.
Video understanding with large language models: A survey
Yunlong Tang, Jing Bi, Siting Xu, Luchuan Song, Susan Liang, Teng Wang, Daoan Zhang, Jie An, Jingyang Lin, Rongyi Zhu, et al · 2023
Later among the works it cites.
Temporal-agnostic change region proposal for semantic change detection
Shiqi Tian, Xicheng Tan, Ailong Ma, Zhuo Zheng, Liangpei Zhang, and Yanfei Zhong · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Change detection methods for remote sensing in the last decade: A comprehensive review
Guangliang Cheng, Yunmeng Huang, Xiangtai Li, Shuchang Lyu, Zhaoyang Xu, Hongbo Zhao, Qi Zhao, and Shiming Xiang · 2024
Closest in time.
A survey of safety and trustworthiness of large language models through the lens of verification and validation
Xiaowei Huang, Wenjie Ruan, Wei Huang, Gaojie Jin, Yi Dong, Changshun Wu, Saddek Bensalem, Ronghui Mu, Yi Qi, Xingyu Zhao, et al · 2024
Closest in time.
Crop monitoring by multimodal remote sensing: A review
Priyabrata Karmakar, Shyh Wei Teng, Manzur Murshed, Shaoning Pang, Yanyu Li, and Hao Lin · 2024
Closest in time.
Moe-llava: Mixture of experts for large vision-language models
Bin Lin, Zhenyu Tang, Yang Ye, Jiaxi Cui, Bin Zhu, Peng Jin, Junwu Zhang, Munan Ning, and Li Yuan · 2024
Closest in time.
Junwei Luo, Zhen Pang, Yongjun Zhang, Tingzhu Wang, Linlin Wang, Bo Dang, Jiangwei Lao, Jian Wang, Jingdong Chen, Yihua Tan, et al · 2024
Closest in time.
Deep learning for satellite image time series analysis: A review
Lynn Miller, Charlotte Pelletier, and Geoffrey I Webb · 2024
Closest in time.
Lhrs-bot: Empowering remote sensing with vgi-enhanced large multimodal language model
Dilxat Muhtar, Zhenshi Li, Feng Gu, Xueliang Zhang, and Pengfeng Xiao · 2024
Closest in time.
Cdchat: A large multimodal model for remote sensing change description
Mubashir Noman, Noor Ahsan, Muzammal Naseer, Hisham Cholakkal, Rao Muhammad Anwer, Salman Khan, and Fahad Shahbaz Khan · 2024
Closest in time.
H2rsvlm: Towards helpful and honest remote sensing large vision language model
Chao Pang, Jiang Wu, Jiayu Li, Yi Liu, Jiaxing Sun, Weijia Li, Xingxing Weng, Shuai Wang, Litong Feng, Gui-Song Xia, et al · 2024
Closest in time.
A comprehensive survey of hallucination mitigation techniques in large language models
SM Tonmoy, SM Zaman, Vinija Jain, Anku Rani, Vipula Rawte, Aman Chadha, and Amitava Das · 2024
Closest in time.
Rs-gpt4v: A unified multimodal instruction-following dataset for remote sensing image understanding
Linrui Xu, Ling Zhao, Wang Guo, Qiujun Li, Kewang Long, Kaiqi Zou, Yuhan Wang, and Haifeng Li · 2024
Closest in time.
Chatearthnet: A global-scale image-text dataset empowering vision-language geo-foundation models
Zhenghang Yuan, Zhitong Xiong, Lichao Mou, and Xiao Xiang Zhu · 2024
Closest in time.
Yang Zhan, Zhitong Xiong, and Yuan Yuan · 2024
Closest in time.
Earthgpt: A universal multi-modal large language model for multi-sensor image comprehension in remote sensing domain
Wei Zhang, Miaoxin Cai, Tong Zhang, Yin Zhuang, and Xuerui Mao · 2024
Closest in time.
Weihong Zhong, Xiaocheng Feng, Liang Zhao, Qiming Li, Lei Huang, Yuxuan Gu, Weitao Ma, Yuan Xu, and Bing Qin · 2024
Closest in time.
Tinyllava: A framework of small-scale large multimodal models
Baichuan Zhou, Ying Hu, Xi Weng, Junlong Jia, Jie Luo, Xien Liu, Ji Wu, and Lei Huang · 2024
Closest in time.