Fetching the paper…
Reading the bibliography…
Understanding urban socioeconomic conditions through visual data is a challenging yet essential task for sustainable urban development and policy planning.
ODIAC fossil fuel CO2 emissions dataset (ODIAC2022), 2015
Tomohiro Oda and Shamil Maksyutov · 2015
Earlier work this paper cites.
https://www.bts.gov/ , 2017
National Household Travel Survey · 2017
Earlier work this paper cites.
Worldpop, open data for spatial demography
Andrew J Tatem · 2017
Earlier work this paper cites.
Lianjia housing price data
Lianjia Real Estate Platform · 2020
Earlier work this paper cites.
Global maps of travel time to healthcare facilities
Daniel J. Weiss, Andrew Nelson, Carlos A. Vargas-Ruiz, et al · 2020
Earlier work this paper cites.
Housing data
Zillow · 2020
Earlier work this paper cites.
Crimes - 2001 to present
Chicago · 2021
Earlier work this paper cites.
Citywide crime statistics - incident level data
New York · 2021
Earlier work this paper cites.
Police department incident reports - historical 2003 to present
San Francisco · 2021
Earlier work this paper cites.
Predicting multi-level socioeconomic indicators from structural urban imagery
Tong Li, Shiduo Xin, Yanxin Xi, Sasu Tarkoma, Pan Hui, and Yong Li · 2022
Earlier work this paper cites.
GHS-BUILT-H R2022A - GHS building height, derived from AW3D30, SRTM30, and Sentinel2 composite (2018) - OBSOLETE RELEASE, 2022
Martino Pesaresi and Panagiotis Politis · 2022
Earlier work this paper cites.
Global gridded gdp data set consistent with the shared socioeconomic pathways
Tingting Wang and Fubao Sun · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, and et al · 2023
Earlier work this paper cites.
Urban visual intelligence: Uncovering hidden city profiles with street view images
Zhuangyuan Fan, Fan Zhang, Becky P. Y. Loo, and Carlo Ratti · 2023
Earlier work this paper cites.
Healthy cities: A comprehensive dataset for environmental determinants of health in england cities
Zhenyu Han, Tong Xia, Yanxin Xi, and Yong Li · 2023
Earlier work this paper cites.
Swe-bench: Can language models resolve real-world github issues?
Carlos E Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan · 2023
Earlier work this paper cites.
Seed-bench: Benchmarking multimodal llms with generative comprehension
Bohao Li, Rui Wang, Guangzhi Wang, Yuying Ge, Yixiao Ge, and Ying Shan · 2023
Earlier work this paper cites.
Toolllm: Facilitating large language models to master 16000+ real-world apis
Yujia Qin, Shihao Liang, Yining Ye, Kunlun Zhu, Lan Yan, Yaxi Lu, Yankai Lin, Xin Cong, Xiangru Tang, Bill Qian, et al · 2023
Cited alongside, same era.
Scibench: Evaluating college-level scientific problem-solving abilities of large language models
Xiaoxuan Wang, Ziniu Hu, Pan Lu, Yanqiao Zhu, Jieyu Zhang, Satyen Subramaniam, Arjun R Loomba, Shichang Zhang, Yizhou Sun, and Wei Wang · 2023
Cited alongside, same era.
Agieval: A human-centric benchmark for evaluating foundation models
Wanjun Zhong, Ruixiang Cui, Yiduo Guo, Yaobo Liang, Shuai Lu, Yanlin Wang, Amin Saied, Weizhu Chen, and Nan Duan · 2023
Cited alongside, same era.
Hierarchical knowledge graph learning enabled socioeconomic indicator prediction in location-based social network
Zhilun Zhou, Yu Liu, Jingtao Ding, Depeng Jin, and Yong Li · 2023
Cited alongside, same era.
Musecl: Predicting urban socioeconomic indicators via multi-semantic contrastive learning
Xixian Yong and Xiao Zhou · 2024
Later among the works it cites.
Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Xiang Yue, Yuansheng Ni, Kai Zhang, Tianyu Zheng, Ruoqi Liu, Ge Zhang, Samuel Stevens, Dongfu Jiang, Weiming Ren, Yuxuan Sun, et al · 2024
Later among the works it cites.
Uv-sam: Adapting segment anything model for urban village identification
Xin Zhang, Yu Liu, Yuming Lin, Qingmin Liao, and Yong Li · 2024
Later among the works it cites.
Phi-4-mini technical report: Compact yet powerful multimodal language models via mixture-of-loras
Abdelrahman Abouelenin, Atabak Ashfaq, Adam Atkinson, Hany Awadalla, Nguyen Bach, Jianmin Bao, Alon Benhaim, Martin Cai, Vishrav Chaudhary, Congcong Chen, et al · 2025
Closest in time.
Amazon nova
Amazon · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhiqiang Shen Aidar Myrzakhan, Sondos Mahmoud Bsharat · 2024
Cited alongside, same era.
Infiagent-dabench: Evaluating agents on data analysis tasks
Xueyu Hu, Ziyu Zhao, Shuang Wei, Ziwei Chai, Qianli Ma, Guoyin Wang, Xuwu Wang, Jing Su, Jingjing Xu, Ming Zhu, et al · 2024
Cited alongside, same era.
Citypulse: fine-grained assessment of urban change with street view time series
Tianyuan Huang, Zejia Wu, Jiajun Wu, Jackelyn Hwang, and Ram Rajagopal · 2024
Cited alongside, same era.
Livecodebench: Holistic and contamination free evaluation of large language models for code
Naman Jain, King Han, Alex Gu, Wen-Ding Li, Fanjia Yan, Tianjun Zhang, Sida Wang, Armando Solar-Lezama, Koushik Sen, and Ion Stoica · 2024
Cited alongside, same era.
Prismatic vlms: investigating the design space of visually-conditioned language models
Siddharth Karamcheti, Suraj Nair, Ashwin Balakrishna, Percy Liang, Thomas Kollar, and Dorsa Sadigh · 2024
Cited alongside, same era.
Can multiple-choice questions really be useful in detecting the abilities of llms?
Wangyue Li, Liangzhi Li, Tong Xiang, Xiao Liu, Wei Deng, and Noa Garcia · 2024
Cited alongside, same era.
Long-term detection and monitory of chinese urban village using satellite imagery
Yuming Lin, Xin Zhang, Yu Liu, Zhenyu Han, Qingmin Liao, and Yong Li · 2024
Cited alongside, same era.
Transforming our world: the 2030 agenda for sustainable development
United Nations · 2024
Cited alongside, same era.
Shuai Bai, Keqin Chen, and Xuejing et al Liu · 2025
Closest in time.
Gemini2.0
Google DeepMind · 2025
Closest in time.
Urbanvlp: Multi-granularity vision-language pretraining for urban socioeconomic indicator prediction
Xixuan Hao, Wei Chen, Yibo Yan, Siru Zhong, Kun Wang, Qingsong Wen, and Yuxuan Liang · 2025
Closest in time.
Urban sensing in the era of large language models
Ce Hou, Fan Zhang, Yong Li, Haifeng Li, Gengchen Mai, Yuhao Kang, Ling Yao, Wenhao Yu, Yao Yao, Song Gao, et al · 2025
Closest in time.
Minimax-01: Scaling foundation models with lightning attention
Aonian Li, Bangwei Gong, Bo Yang, Boji Shan, Chang Liu, Cheng Zhu, Chunhao Zhang, Congchao Guo, Da Chen, Dong Li, et al · 2025
Closest in time.
Cityrise: Reasoning urban socio-economic status in vision-language models via reinforcement learning
Tianhui Liu, Hetian Pang, Xin Zhang, Jie Feng, Yong Li, and Pan Hui · 2025
Closest in time.
Local data for better health
PLACES · 2025
Closest in time.
Flexireg: Flexible urban region representation learning
Fengze Sun, Yanchuan Chang, Egemen Tanin, Shanika Karunasekera, and Jianzhong Qi · 2025
Closest in time.
Cross-platform complementarity: Assessing the data quality and availability of google street view and baidu street view
Lei Wang, Martin Kada, Tianlin Zhang, and Jie He · 2025
Closest in time.
Perceiving urban inequality from imagery using visual language models with chain-of-thought reasoning
Yunke Zhang, Ruolong Ma, Xin Zhang, and Yong Li · 2025
Closest in time.
Urbench: A comprehensive benchmark for evaluating large multimodal models in multi-view urban scenarios
Baichuan Zhou, Haote Yang, Dairong Chen, Junyan Ye, Tianyi Bai, Jinhua Yu, Songyang Zhang, Dahua Lin, Conghui He, and Weijia Li · 2025
Closest in time.