Fetching the paper…
Reading the bibliography…
As large language models (LLMs) continue to advance and gain widespread use, establishing systematic and reliable evaluation methodologies for LLMs and vision-language models (VLMs) has become essential to ensure their real-world effectiveness and reliability.
The image of the city
Kevin Lynch. 1964 · 1964
Earlier work this paper cites.
Congested traffic states in empirical observations and microscopic simulations
Martin Treiber, Ansgar Hennecke, and Dirk Helbing. 2000 · 2000
Earlier work this paper cites.
General lane-changing model MOBIL for car-following models
Arne Kesting, Martin Treiber, and Dirk Helbing. 2007 · 2007
Earlier work this paper cites.
Smart cities of the future
Michael Batty, Kay W Axhausen, Fosca Giannotti, Alexei Pozdnoukhov, Armando Bazzani, Monica Wachowicz, Georgios Ouzounis, and Yuval Portugali. 2012 · 2012
Earlier work this paper cites.
The max-pressure controller for arbitrary networks of signalized intersections
Pravin Varaiya. 2013 · 2013
Earlier work this paper cites.
Urban computing: concepts, methodologies, and applications
Yu Zheng, Licia Capra, Ouri Wolfson, and Hai Yang. 2014 · 2014
Earlier work this paper cites.
Combining satellite imagery and machine learning to predict poverty
Neal Jean, Marshall Burke, Michael Xie, W Matthew Davis, David B Lobell, and Stefano Ermon. 2016 · 2016
Earlier work this paper cites.
Participatory cultural mapping based on collective behavior data in location-based social networks
Dingqi Yang, Daqing Zhang, and Bingqing Qu. 2016 · 2016
Earlier work this paper cites.
The cognitive map in humans: spatial navigation and beyond
Russell A Epstein, Eva Zita Patai, Joshua B Julian, and Hugo J Spiers. 2017 · 2017
Earlier work this paper cites.
WorldPop, open data for spatial demography
Andrew J Tatem. 2017 · 2017
Earlier work this paper cites.
Microscopic Traffic Simulation using SUMO, In The 21st IEEE International Conference on Intelligent Transportation Systems
Pablo Alvarez Lopez, Michael Behrisch, Laura Bieker-Walz, and et al. 2018 · 2018
Earlier work this paper cites.
Learning to navigate in cities without a map
Piotr Mirowski, Matt Grimes, Mateusz Malinowski, and et al. 2018 · 2018
Earlier work this paper cites.
Virtualhome: Simulating household activities via programs. In
Xavier Puig, Kevin Ra, Marko Boben, Jiaman Li, Tingwu Wang, Sanja Fidler, and Antonio Torralba. 2018 · 2018
Earlier work this paper cites.
Touchdown: Natural language navigation and spatial reasoning in visual street environments. In
Howard Chen, Alane Suhr, Dipendra Misra, Noah Snavely, and Yoav Artzi. 2019 · 2019
Earlier work this paper cites.
A survey on traffic signal control methods
Hua Wei, Guanjie Zheng, Vikash Gayah, and Zhenhui Li. 2019 · 2019
Earlier work this paper cites.
Cityflow: A multi-agent reinforcement learning environment for large scale city traffic scenario. In
Huichu Zhang, Siyuan Feng, Chang Liu, and et al. 2019 · 2019
Earlier work this paper cites.
Where to go next: Modeling long-and short-term user preferences for point-of-interest recommendation. In
Ke Sun, Tieyun Qian, Tong Chen, Yile Liang, Quoc Viet Hung Nguyen, and Hongzhi Yin. 2020 · 2020
Earlier work this paper cites.
Using publicly available satellite imagery and deep learning to understand economic well-being in Africa
Christopher Yeh, Anthony Perez, Anne Driscoll, George Azzari, Zhongyi Tang, David Lobell, Stefano Ermon, and Marshall Burke. 2020 · 2020
Earlier work this paper cites.
Intelligent driving intelligence test for autonomous vehicles with naturalistic and adversarial environment
Shuo Feng, Xintao Yan, Haowei Sun, Yiheng Feng, and Henry X Liu. 2021 · 2021
Earlier work this paper cites.
Geographic question answering: challenges, uniqueness, classification, and future directions
Gengchen Mai, Krzysztof Janowicz, Rui Zhu, Ling Cai, and Ni Lao. 2021 · 2021
Earlier work this paper cites.
Scott Reed, Konrad Zolna, Emilio Parisotto, and et al. 2022 · 2022
Earlier work this paper cites.
Challenging big-bench tasks and whether chain-of-thought can solve them
Mirac Suzgun, Nathan Scales, Nathanael Schärli, Sebastian Gehrmann, Yi Tay, Hyung Won Chung, Aakanksha Chowdhery, Quoc V Le, Ed H Chi, Denny Zhou, et al · 2022
Earlier work this paper cites.
Advancing plain vision transformer toward remote sensing foundation model
Di Wang, Qiming Zhang, Yufei Xu, Jing Zhang, Bo Du, Dacheng Tao, and Liangpei Zhang. 2022 · 2022
Cited alongside, same era.
Emergent abilities of large language models
Jason Wei, Yi Tay, Rishi Bommasani, and et al. 2022 · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Are large language models geospatially knowledgeable?. In
Prabin Bhandari, Antonios Anastasopoulos, and Dieter Pfoser. 2023 · 2023
Cited alongside, same era.
Learning A Foundation Language Model for Geoscience Knowledge Understanding and Utilization
Cheng Deng, Tianhang Zhang, Zhongmou He, Qiyuan Chen, Yuanyuan Shi, Le Zhou, Luoyi Fu, Weinan Zhang, Xinbing Wang, Chenghu Zhou, et al · 2023
Where would i go next? large language models as human mobility predictors
Xinglei Wang, Meng Fang, Zichao Zeng, and Tao Cheng. 2023a · 2023
Later among the works it cites.
Learning interactive real-world simulators
Mengjiao Yang, Yilun Du, Kamyar Ghasemipour, Jonathan Tompson, Dale Schuurmans, and Pieter Abbeel. 2023 · 2023
Later among the works it cites.
Deep learning for trajectory data management and mining: A survey and beyond
Wei Chen, Yuxuan Liang, Yuanshao Zhu, Yanchuan Chang, Kang Luo, Haomin Wen, Lei Li, Yanwei Yu, Qingsong Wen, Chao Chen, et al · 2024
Closest in time.
Understanding World or Predicting Future? A Comprehensive Survey of World Models
Jingtao Ding, Yunke Zhang, Yu Shang, Yuheng Zhang, Zefang Zong, Jie Feng, Yuan Yuan, Hongyuan Su, Nian Li, Nicholas Sukiennik, et al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Urban visual intelligence: Uncovering hidden city profiles with street view images
Zhuangyuan Fan, Fan Zhang, Becky PY Loo, and Carlo Ratti. 2023 · 2023
Cited alongside, same era.
Language models represent space and time
Wes Gurnee and Max Tegmark. 2023 · 2023
Cited alongside, same era.
Pigeon: Predicting image geolocations
Lukas Haas, Silas Alberti, and Michal Skreta. 2023 · 2023
Cited alongside, same era.
Metagpt: Meta programming for multi-agent collaborative framework
Sirui Hong, Xiawu Zheng, Jonathan Chen, Yuheng Cheng, Jinlin Wang, Ceyao Zhang, Zili Wang, Steven Ka Shing Yau, Zijuan Lin, Liyang Zhou, et al · 2023
Cited alongside, same era.
"Cities: Skylines II"
Paradox Interactive. 2023 · 2023
Cited alongside, same era.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, and et al. 2023 · 2023
Cited alongside, same era.
Geochat: Grounded large vision-language model for remote sensing
Kartik Kuckreja, Muhammad Sohail Danish, Muzammal Naseer, Abhijit Das, Salman Khan, and Fahad Shahbaz Khan. 2023 · 2023
Cited alongside, same era.
Vlmevalkit: An open-source toolkit for evaluating large multi-modality models. In
Haodong Duan, Junming Yang, Yuxuan Qiao, Xinyu Fang, Lin Chen, Yuan Liu, Xiaoyi Dong, Yuhang Zang, Pan Zhang, Jiaqi Wang, et al · 2024
Closest in time.
CityGPT: Empowering Urban Spatial Cognition of Large Language Models
Jie Feng, Tianhui Liu, Yuwei Du, Siqi Guo, Yuming Lin, and Yong Li. 2024 · 2024
Closest in time.
A population-to-individual tuning framework for adapting pretrained LM to on-device user intent prediction. In
Jiahui Gong, Jingtao Ding, Fanjin Meng, Guilong Chen, Hong Chen, Shen Zhao, Haisheng Lu, and Yong Li. 2024 · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, and et al. 2024 · 2024
Closest in time.
RemoteCLIP: A Vision Language Foundation Model for Remote Sensing
Fan Liu, Delong Chen, Zhangqingyun Guan, Xiaocong Zhou, Jiale Zhu, Qiaolin Ye, Liyong Fu, and Jun Zhou. 2024a · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2024b · 2024
Closest in time.
Large language models are geographically biased
Rohin Manvi, Samar Khanna, Marshall Burke, David Lobell, and Stefano Ermon. 2024 · 2024
Closest in time.
MiniCPM-Llama3-V 2.5
OpenBMB. 2024 · 2024
Closest in time.
Velma: Verbalization embodiment of llm agents for vision and language navigation in street view. In
Raphael Schumann, Wanrong Zhu, Weixi Feng, Tsu-Jui Fu, Stefan Riezler, and William Yang Wang. 2024 · 2024
Closest in time.
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Zhihong Shao, Damai Dai, Daya Guo, and Bo Liu. 2024 · 2024
Closest in time.
Geoclip: Clip-inspired alignment between locations and images for effective worldwide geo-localization
Vicente Vivanco Cepeda, Gaurav Kumar Nayak, and Mubarak Shah. 2024 · 2024
Closest in time.
A survey on large language model based autonomous agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, et al · 2024
Closest in time.
Can Language Models Serve as Text-Based World Simulators?
Ruoyao Wang, Graham Todd, Ziang Xiao, Xingdi Yuan, Marc-Alexandre Côté, Peter Clark, and Peter Jansen. 2024b · 2024
Closest in time.
Language models meet world models: Embodied experiences enhance language models
Jiannan Xiang, Tianhua Tao, Yi Gu, Tianmin Shu, Zirui Wang, Zichao Yang, and Zhiting Hu. 2024 · 2024
Closest in time.
UrbanCLIP: Learning Text-enhanced Urban Region Profiling with Contrastive Language-Image Pretraining from the Web. In
Yibo Yan, Haomin Wen, Siru Zhong, and et al. 2024 · 2024
Closest in time.
V-IRL: Grounding Virtual Intelligence in Real Life
Jihan Yang, Runyu Ding, Ellis Brown, Xiaojuan Qi, and Saining Xie. 2024 · 2024
Closest in time.
How to Enable LLM with 3D Capacity? A Survey of Spatial Reasoning in LLM
Jirong Zha, Yuxuan Fan, Xiao Yang, Chen Gao, and Xinlei Chen. 2025 · 2025
Closest in time.