Fetching the paper…
Reading the bibliography…
Remote sensing imagery, despite its broad applications in helping achieve Sustainable Development Goals and tackle climate change, has not yet benefited from the recent advancements of versatile, task-agnostic vision language models (VLMs).
Bag-of-visual-words and spatial extensions for land-use classification
Yang, Y.; and Newsam, S. 2010 · 2010
Earlier work this paper cites.
Combining satellite imagery and machine learning to predict poverty
Jean, N.; Burke, M.; Xie, S. M.; Davis, W. M.; Lobell, D.; and Ermon, S. 2016 · 2016
Earlier work this paper cites.
Learning visual features from large weakly supervised data
Joulin, A.; Van Der Maaten, L.; Jabri, A.; and Vasilache, N. 2016 · 2016
Earlier work this paper cites.
Deep semantic understanding of high resolution remote sensing image
Qu, B.; Li, X.; Tao, D.; and Lu, X. 2016 · 2016
Earlier work this paper cites.
Remote sensing image scene classification: Benchmark and state of the art
Cheng, G.; Han, J.; and Lu, X. 2017 · 2017
Earlier work this paper cites.
Exploring Models and Data for Remote Sensing Image Caption Generation
Lu, X.; Wang, B.; Zheng, X.; and Li, X. 2017 · 2017
Earlier work this paper cites.
AID: A benchmark data set for performance evaluation of aerial scene classification
Xia, G.-S.; Hu, J.; Hu, F.; Shi, B.; Bai, X.; Zhong, Y.; Zhang, L.; and Lu, X. 2017 · 2017
Earlier work this paper cites.
Deep Gaussian Process for Crop Yield Prediction Based on Remote Sensing Data
You, J.; Li, X.; Low, M.; Lobell, D.; and Ermon, S. 2017 · 2017
Earlier work this paper cites.
Functional map of the world
Christie, G.; Fendley, N.; Wilson, J.; and Mukherjee, R. 2018 · 2018
Earlier work this paper cites.
DeepSolar: A Machine Learning Framework to Efficiently Construct a Solar Deployment Database in the United States
Yu, J.; Wang, Z.; Majumdar, A.; and Rajagopal, R. 2018 · 2018
Earlier work this paper cites.
PatternNet: A benchmark dataset for performance evaluation of remote sensing image retrieval
Zhou, W.; Newsam, S.; Li, C.; and Shao, Z. 2018 · 2018
Earlier work this paper cites.
Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification
Helber, P.; Bischke, B.; Dengel, A.; and Borth, D. 2019 · 2019
Earlier work this paper cites.
Tile2vec: Unsupervised representation learning for spatially distributed data
Jean, N.; Wang, S.; Samar, A.; Azzari, G.; Lobell, D.; and Ermon, S. 2019 · 2019
Earlier work this paper cites.
Bigearthnet: A large-scale benchmark archive for remote sensing image understanding
Sumbul, G.; Charfuelan, M.; Demir, B.; and Markl, V. 2019 · 2019
Earlier work this paper cites.
Learning visual representations with caption annotations
Sariyildiz, M. B.; Perez, J.; and Larlus, D. 2020 · 2020
Earlier work this paper cites.
Geography-aware self-supervised learning
Ayush, K.; Uzkent, B.; Meng, C.; Tanmay, K.; Burke, M.; Lobell, D.; and Ermon, S. 2021 · 2021
Cited alongside, same era.
Virtex: Learning visual representations from textual annotations
Desai, K.; and Johnson, J. 2021 · 2021
Cited alongside, same era.
GridTracer: Automatic Mapping of Power Grids Using Deep Learning and Overhead Imagery
Huang, B.; Yang, J.; Streltsov, A.; Bradbury, K.; Collins, L. M.; and Malof, J. M. 2021 · 2021
Cited alongside, same era.
Scaling up visual and vision-language representation learning with noisy text supervision
Jia, C.; Yang, Y.; Xia, Y.; Chen, Y.-T.; Parekh, Z.; Pham, H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T. 2021 · 2021
Cited alongside, same era.
A global inventory of photovoltaic solar energy generating units
Kruitwagen, L.; Story, K. T.; Friedrich, J.; Byers, L.; Skillman, S.; and Hepburn, C. 2021 · 2021
Cited alongside, same era.
Satmae: Pre-training transformers for temporal and multi-spectral satellite imagery
Cong, Y.; Khanna, S.; Meng, C.; Liu, P.; Rozi, E.; He, Y.; Burke, M.; Lobell, D.; and Ermon, S. 2022 · 2022
Later among the works it cites.
Opensentinelmap: A large-scale land use dataset using openstreetmap and sentinel-2 imagery
Johnson, N.; Treible, W.; and Crispell, D. 2022 · 2022
Later among the works it cites.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Li, J.; Li, D.; Xiong, C.; and Hoi, S. 2022 · 2022
Later among the works it cites.
Scale-mae: A scale-aware masked autoencoder for multiscale geospatial representation learning
Reed, C. J.; Gupta, R.; Li, S.; Brockman, S.; Funk, C.; Clipp, B.; Candido, S.; Uyttendaele, M.; and Darrell, T. 2022 · 2022
Later among the works it cites.
ReforesTree: A Dataset for Estimating Tropical Forest Carbon Stock with Deep Learning and Aerial Imagery
Reiersen, G.; Dao, D.; Lütjens, B.; Klemmer, K.; Amara, K.; Steinegger, A.; Zhang, C.; and Zhu, X. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lee, J.; Brooks, N. R.; Tajwar, F.; Burke, M.; Ermon, S.; Lobell, D.; Biswas, D.; and Luby, S. P. 2021 · 2021
Cited alongside, same era.
On creating benchmark dataset for aerial image interpretation: Reviews, guidances, and million-aid
Long, Y.; Xia, G.-S.; Li, S.; Yang, W.; Yang, M. Y.; Zhu, X. X.; Zhang, L.; and Li, D. 2021 · 2021
Cited alongside, same era.
Seasonal contrast: Unsupervised pre-training from uncurated remote sensing data
Manas, O.; Lacoste, A.; Giró-i Nieto, X.; Vazquez, D.; and Rodriguez, P. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Cited alongside, same era.
Deforestation Detection with Fully Convolutional Networks in the Amazon Forest from Landsat-8 and Sentinel-2 Images
Torres, D. L.; Turnes, J. N.; Vega, P. J. S.; Feitosa, R. Q.; Silva, D. E.; Junior, J. M.; and de Almeida, C. A. 2021 · 2021
Cited alongside, same era.
LoveDA: A remote sensing land-cover dataset for domain adaptive semantic segmentation
Wang, J.; Zheng, Z.; Ma, A.; Lu, X.; and Zhong, Y. 2021 · 2021
Cited alongside, same era.
Flamingo: a visual language model for few-shot learning
Alayrac, J.-B.; Donahue, J.; Luc, P.; Miech, A.; Barr, I.; Hasson, Y.; Lenc, K.; Mensch, A.; Millican, K.; Reynolds, M.; et al. 2022 · 2022
Cited alongside, same era.
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Schuhmann, C.; Beaumont, R.; Vencu, R.; Gordon, C.; Wightman, R.; Cherti, M.; Coombes, T.; Katta, A.; Mullis, C.; Wortsman, M.; et al. 2022 · 2022
Later among the works it cites.
EarthNets: Empowering AI in Earth Observation
Xiong, Z.; Zhang, F.; Wang, Y.; Shi, Y.; and Zhu, X. X. 2022 · 2022
Later among the works it cites.
Coca: Contrastive captioners are image-text foundation models
Yu, J.; Wang, Z.; Vasudevan, V.; Yeung, L.; Seyedhosseini, M.; and Wu, Y. 2022 · 2022
Later among the works it cites.
When and Why Vision-Language Models Behave like Bags-Of-Words, and What to Do About It?
Yuksekgonul, M.; Bianchi, F.; Kalluri, P.; Jurafsky, D.; and Zou, J. 2022 · 2022
Later among the works it cites.
Satlas: A Large-Scale Dataset for Remote Sensing Image Understanding
Bastani, F.; Wolters, P.; Gupta, R.; Ferdinando, J.; and Kembhavi, A. 2023 · 2023
Closest in time.
A billion-scale foundation model for remote sensing images
Cha, K.; Seo, J.; and Lee, T. 2023 · 2023
Closest in time.
SATIN: A Multi-Task Metadataset for Classifying Satellite Imagery using Vision-Language Models
Jonathan Roberts, K. H.; and Albanie, S. 2023 · 2023
Closest in time.
RemoteCLIP: A Vision Language Foundation Model for Remote Sensing
Liu, F.; Chen, D.; Guan, Z.; Zhou, X.; Zhu, J.; and Zhou, J. 2023 · 2023
Closest in time.
Gfm: Building geospatial foundation models via continual pretraining
Mendieta, M.; Han, B.; Shi, X.; Zhu, Y.; Chen, C.; and Li, M. 2023 · 2023
Closest in time.