Fetching the paper…
Reading the bibliography…
Urban socioeconomic indicator prediction aims to infer various metrics related to sustainable development in diverse urban landscapes using data-driven methods.
Generating interpretable poverty maps using object detection in satellite images
Ayush, K.; Uzkent, B.; Burke, M.; Lobell, D.; and Ermon, S. 2020 · 2002
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Integrating aerial and street view images for urban land use classification
Cao, R.; Zhu, J.; Tu, W.; Li, Q.; Cao, J.; Liu, B.; Zhang, Q.; and Qiu, G. 2018 · 2018
Earlier work this paper cites.
Perceiving commerial activeness over satellite images
He, Z.; Yang, S.; Zhang, W.; and Zhang, J. 2018 · 2018
Earlier work this paper cites.
Geoman: Multi-level attention networks for geo-sensory time series prediction
Liang, Y.; Ke, S.; Zhang, J.; Yi, X.; and Zheng, Y. 2018 · 2018
Earlier work this paper cites.
Take a Look Around: Using Street View and Satellite Images to Estimate House Prices
Law, S.; Paige, B.; and Russell, C. 2019 · 2019
Earlier work this paper cites.
ViLBERT: pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Lu, J.; Batra, D.; Parikh, D.; and Lee, S. 2019 · 2019
Earlier work this paper cites.
Semantic segmentation of crop type in Africa: A novel dataset and analysis of deep learning methods
M Rustowicz, R.; Cheong, R.; Wang, L.; Ermon, S.; Burke, M.; and Lobell, D. 2019 · 2019
Earlier work this paper cites.
LXMERT: Learning Cross-Modality Encoder Representations from Transformers
Tan, H. H.; and Bansal, M. 2019 · 2019
Earlier work this paper cites.
Integrating Remote Sensing and Street View Images to Quantify Urban Forest Ecosystem Services
Barbierato, E.; Bernetti, I.; Capecchi, I.; and Saragosa, C. 2020 · 2020
Earlier work this paper cites.
Uniter: Universal image-text representation learning
Chen, Y.-C.; Li, L.; Yu, L.; El Kholy, A.; Ahmed, F.; Gan, Z.; Cheng, Y.; and Liu, J. 2020 · 2020
Earlier work this paper cites.
Learning to score economic development from satellite imagery
Han, S.; Ahn, D.; Park, S.; Yang, J.; Lee, S.; Kim, J.; Yang, H.; Park, S.; and Cha, M. 2020 · 2020
Earlier work this paper cites.
A survey on contrastive self-supervised learning
Jaiswal, A.; Babu, A. R.; Zadeh, M. Z.; Banerjee, D.; and Makedon, F. 2020 · 2020
Earlier work this paper cites.
Vilbertscore: Evaluating image caption using vision-and-language bert
Lee, H.; Yoon, S.; Dernoncourt, F.; Kim, D. S.; Bui, T.; and Jung, K. 2020 · 2020
Earlier work this paper cites.
Unicoder-vl: A universal encoder for vision and language by cross-modal pre-training
Li, G.; Duan, N.; Fang, Y.; Gong, M.; and Jiang, D. 2020 · 2020
Earlier work this paper cites.
Self-attention for raw optical satellite time series classification
Rußwurm, M.; and Körner, M. 2020 · 2020
Earlier work this paper cites.
Urban2vec: Incorporating street view imagery and pois for multi-modal urban neighborhood embedding
Wang, Z.; Li, H.; and Rajagopal, R. 2020 · 2020
Earlier work this paper cites.
BERTScore: Evaluating Text Generation with BERT
Zhang, T.; Kishore, V.; Wu, F.; Weinberger, K. Q.; and Artzi, Y. 2020 · 2020
Earlier work this paper cites.
Long time series nighttime light dataset of China (2000–2020)
Zhong, X.; Yan, Q.; and Li, G. 2022 · 2020
Earlier work this paper cites.
Efficient poverty mapping from high resolution remote sensing images
Ayush, K.; Uzkent, B.; Tanmay, K.; Burke, M.; Lobell, D.; and Ermon, S. 2021 · 2021
Cited alongside, same era.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; Uszkoreit, J.; and Houlsby, N. 2021 · 2021
Cited alongside, same era.
CLIPScore: A Reference-free Evaluation Metric for Image Captioning
Hessel, J.; Holtzman, A.; Forbes, M.; Bras, R. L.; and Choi, Y. 2021 · 2021
Cited alongside, same era.
Scaling up visual and vision-language representation learning with noisy text supervision
Jia, C.; Yang, Y.; Xia, Y.; Chen, Y.-T.; Parekh, Z.; Pham, H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T. 2021 · 2021
Cited alongside, same era.
Vilt: Vision-and-language transformer without convolution or region supervision
Kim, W.; Son, B.; and Kim, I. 2021 · 2021
Cited alongside, same era.
A challenger to gpt-4v? early explorations of gemini in visual expertise
Fu, C.; Zhang, R.; Lin, H.; Wang, Z.; Gao, T.; Luo, Y.; Huang, Y.; Zhang, Z.; Qiu, L.; Ye, G.; et al. 2023 · 2023
Later among the works it cites.
Llama-adapter v2: Parameter-efficient visual instruction model
Gao, P.; Han, J.; Zhang, R.; Lin, Z.; Geng, S.; Zhou, A.; Zhang, W.; Lu, P.; He, C.; Yue, X.; et al. 2023 · 2023
Later among the works it cites.
InfoMetIC: An Informative Metric for Reference-free Image Caption Evaluation
Hu, A.; Chen, S.; Zhang, L.; and Jin, Q. 2023 · 2023
Later among the works it cites.
Comprehensive urban space representation with varying numbers of street-level images
Huang, Y.; Zhang, F.; Gao, Y.; Tu, W.; Duarte, F.; Ratti, C.; Guo, D.; and Liu, Y. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Li, Y.; Liang, F.; Zhao, L.; Cui, Y.; Ouyang, W.; Shao, J.; Yu, F.; and Yan, J. 2021 · 2021
Cited alongside, same era.
Fine-grained urban flow prediction
Liang, Y.; Ouyang, K.; Sun, J.; Wang, Y.; Zhang, J.; Zheng, Y.; Rosenblum, D.; and Zimmermann, R. 2021 · 2021
Cited alongside, same era.
Fully convolutional recurrent networks for multidate crop recognition from multitemporal image sequences
Martinez, J. A. C.; La Rosa, L. E. C.; Feitosa, R. Q.; Sanches, I. D.; and Happ, P. N. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Cited alongside, same era.
Multimodal few-shot learning with frozen language models
Tsimpoukelli, M.; Menick, J. L.; Cabi, S.; Eslami, S.; Vinyals, O.; and Hill, F. 2021 · 2021
Cited alongside, same era.
Sustainbench: Benchmarks for monitoring the sustainable development goals with machine learning
Yeh, C.; Meng, C.; Wang, S.; Driscoll, A.; Rozi, E.; Liu, P.; Lee, J.; Burke, M.; Lobell, D. B.; and Ermon, S. 2021 · 2021
Cited alongside, same era.
Multi-view joint graph representation learning for urban region embedding
Zhang, M.; Li, T.; Li, Y.; and Hui, P. 2021 · 2021
Cited alongside, same era.
Li, J.; Li, D.; Savarese, S.; and Hoi, S. 2023 · 2023
Later among the works it cites.
Knowledge-infused contrastive learning for urban imagery-based socioeconomic prediction
Liu, Y.; Zhang, X.; Ding, J.; Xi, Y.; and Li, Y. 2023d · 2023
Later among the works it cites.
Geollm: Extracting geospatial knowledge from large language models
Manvi, R.; Khanna, S.; Mai, G.; Burke, M.; Lobell, D.; and Ermon, S. 2023 · 2023
Later among the works it cites.
Sdxl: Improving latent diffusion models for high-resolution image synthesis
Podell, D.; English, Z.; Lacey, K.; Blattmann, A.; Dockhorn, T.; Müller, J.; Penna, J.; and Rombach, R. 2023 · 2023
Later among the works it cites.
A survey of hallucination in large foundation models
Rawte, V.; Sheth, A.; and Das, A. 2023 · 2023
Later among the works it cites.
GeoCLIP: Clip-Inspired Alignment between Locations and Images for Effective Worldwide Geo-localization
Vivanco, V.; Nayak, G. K.; and Shah, M. 2023 · 2023
Later among the works it cites.
The dawn of lmms: Preliminary explorations with gpt-4v (ision)
Yang, Z.; Li, L.; Lin, K.; Wang, J.; Lin, C.-C.; Liu, Z.; and Wang, L. 2023 · 2023
Later among the works it cites.
mplug-owl: Modularization empowers large language models with multimodality
Ye, Q.; Xu, H.; Xu, G.; Ye, J.; Yan, M.; Zhou, Y.; Wang, J.; Hu, A.; Shi, P.; Shi, Y.; et al. 2023 · 2023
Later among the works it cites.
Trajectory-User Linking via Hierarchical Spatio-Temporal Attention Networks
Chen, W.; Huang, C.; Yu, Y.; Jiang, Y.; and Dong, J. 2024 · 2024
Closest in time.
GitHub - CSAILVision/semantic-segmentation-pytorch: Pytorch implementation for Semantic Segmentation/Scene Parsing on MIT ADE20K dataset — github.com
CSAILVision. 2024 · 2024
Closest in time.
LLM-enhanced Reranking in Recommender Systems
Gao, J.; Chen, B.; Zhao, X.; Liu, W.; Li, X.; Wang, Y.; Zhang, Z.; Wang, W.; Ye, Y.; Lin, S.; et al. 2024 · 2024
Closest in time.
Mapbox - Location Data & Maps for Developers
Mapbox. 2024 · 2024
Closest in time.
Urbanclip: Learning text-enhanced urban region profiling with contrastive language-image pretraining from the web
Yan, Y.; Wen, H.; Zhong, S.; Chen, W.; Chen, H.; Wen, Q.; Zimmermann, R.; and Liang, Y. 2024 · 2024
Closest in time.
Towards Urban General Intelligence: A Review and Outlook of Urban Foundation Models
Zhang, W.; Han, J.; Xu, Z.; Ni, H.; Liu, H.; and Xiong, H. 2024 · 2024
Closest in time.
UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
Zhong, S.; Hao, X.; Yan, Y.; Zhang, Y.; Song, Y.; and Liang, Y. 2024 · 2024
Closest in time.
Deep Learning for Cross-Domain Data Fusion in Urban Computing: Taxonomy, Advances, and Outlook
Zou, X.; Yan, Y.; Hao, X.; Hu, Y.; Wen, H.; Liu, E.; Zhang, J.; Li, Y.; Li, T.; Zheng, Y.; et al. 2024 · 2024
Closest in time.