Fetching the paper…
Reading the bibliography…
The emergence of large language models such as ChatGPT, Gemini, and others highlights the importance of evaluating their diverse capabilities, ranging from natural language understanding to code generation.
“Geographic Information Analysis”
David O’Sullivan and David. Unwin · 2010
Earlier work this paper cites.
“Geographic Information Analysis”
David O’Sullivan and David. Unwin · 2010
Earlier work this paper cites.
“Problematizing spatial literacy within the school curriculum”
David Lane, Raymond Lynch and Oliver McGarr · 2019
Earlier work this paper cites.
“Problematizing spatial literacy within the school curriculum”
David Lane, Raymond Lynch and Oliver McGarr · 2019
Earlier work this paper cites.
“ChatGPT usage and limitations”, 2022
Amos Azaria · 2022
Earlier work this paper cites.
“Getting to Know ArcGIS Desktop 10.8”
Michael Law and Amy Collins · 2022
Earlier work this paper cites.
“Chain-of-thought prompting elicits reasoning in large language models”
Jason Wei et al · 2022
Earlier work this paper cites.
“Emergent abilities of large language models”
Jason Wei et al · 2022
Earlier work this paper cites.
“Large language models are zero-shot reasoners”
Takeshi Kojima et al · 2022
Earlier work this paper cites.
“Complexity-based prompting for multi-step reasoning”
Yao Fu et al · 2022
Earlier work this paper cites.
“ChatGPT usage and limitations”, 2022
Amos Azaria · 2022
Earlier work this paper cites.
“Complexity-based prompting for multi-step reasoning”
Yao Fu et al · 2022
Earlier work this paper cites.
“Large language models are zero-shot reasoners”
Takeshi Kojima et al · 2022
Earlier work this paper cites.
“Getting to Know ArcGIS Desktop 10.8”
Michael Law and Amy Collins · 2022
Earlier work this paper cites.
“Chain-of-thought prompting elicits reasoning in large language models”
Jason Wei et al · 2022
Earlier work this paper cites.
“Emergent abilities of large language models”
Jason Wei et al · 2022
Earlier work this paper cites.
“ChatGPT: Jack of all trades, master of none”
Jan Kocoń et al · 2023
Earlier work this paper cites.
“A survey of large language models”
Wayne Zhao et al · 2023
Earlier work this paper cites.
“Gemini: a family of highly capable multimodal models”
Gemini Team et al · 2023
Earlier work this paper cites.
“GPT4GEO: How a Language Model Sees the World’s Geography” Oral Presentation
Jonathan Roberts et al · 2023
Earlier work this paper cites.
“Autonomous GIS: the next-generation AI-powered GIS”
Zhe Li and He Ning · 2023
Earlier work this paper cites.
“ChatGPT as a mapping assistant: A novel method to enrich maps with generative AI and content derived from street-level photographs”
Levente Juhász, Peter Mooney, Hartwig. Hochmair and Boyuan Guan · 2023
Earlier work this paper cites.
“Superclue: A comprehensive chinese large language model benchmark”
Liang Xu et al · 2023
Earlier work this paper cites.
“Battle of the wordsmiths: Comparing chatgpt, gpt-4, claude, and bard”
Ali Borji and Mehrdad Mohammadian · 2023
Earlier work this paper cites.
“GPT-4 vs. GPT-3.5: A concise showdown”
Anis Koubaa · 2023
Earlier work this paper cites.
“Towards understanding the geospatial skills of chatgpt: Taking a geographic information systems (gis) exam”
Peter Mooney, Wencong Cui, Boyuan Guan and Levente Juhász · 2023
Earlier work this paper cites.
“Mapping with ChatGPT”
Ran Tao and Jinwen Xu · 2023
Earlier work this paper cites.
“Evaluating spatial understanding of large language models”
Yutaro Yamada et al · 2023
Earlier work this paper cites.
“Exploring and Improving the Spatial Reasoning Abilities of Large Language Models”
Manasi Sharma · 2023
Earlier work this paper cites.
“Geo-knowledge-guided GPT models improve the extraction of location descriptions from disaster-related social media messages”
Yingjie Hu et al · 2023
Earlier work this paper cites.
“GeoQAMap - Geographic Question Answering with Maps Leveraging LLM and Open Knowledge Base”
Yu Feng, Linfang Ding and Guohui Xiao · 2023
Earlier work this paper cites.
“Self-Consistency Improves Chain of Thought Reasoning in Language Models”
Xuezhi Wang et al · 2023
Earlier work this paper cites.
“Decomposed Prompting: A Modular Approach for Solving Complex Tasks” In-Person Poster Presentation
Tushar Khot et al · 2023
Earlier work this paper cites.
“Battle of the wordsmiths: Comparing chatgpt, gpt-4, claude, and bard”
Ali Borji and Mehrdad Mohammadian · 2023
Cited alongside, same era.
“GeoQAMap - Geographic Question Answering with Maps Leveraging LLM and Open Knowledge Base”
Yu Feng, Linfang Ding and Guohui Xiao · 2023
Cited alongside, same era.
“Geo-knowledge-guided GPT models improve the extraction of location descriptions from disaster-related social media messages”
Yingjie Hu et al · 2023
Cited alongside, same era.
“ChatGPT as a mapping assistant: A novel method to enrich maps with generative AI and content derived from street-level photographs”
Levente Juhász, Peter Mooney, Hartwig. Hochmair and Boyuan Guan · 2023
Cited alongside, same era.
“Decomposed Prompting: A Modular Approach for Solving Complex Tasks” In-Person Poster Presentation
Tushar Khot et al · 2023
Cited alongside, same era.
“CityBench: Evaluating the Capabilities of Large Language Model as World Model”
Jie Feng et al · 2024
Closest in time.
“Can Large Language Models Generate Geospatial Code?”
Shuyang Hou et al · 2024
Closest in time.
“Evaluation of Code LLMs on Geospatial Code Generation”
Piotr Gramacki, Bruno Martins and Piotr Szymański · 2024
Closest in time.
“One-Shot Learning as Instruction Data Prospector for Large Language Models”
Yunshui Li et al · 2024
Closest in time.
“GPT-4o System Card” This report outlines the safety work carried out prior to releasing GPT-4o, including external red teaming and frontier risk evaluations., 2024
OpenAI · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“ChatGPT: Jack of all trades, master of none”
Jan Kocoń et al · 2023
Cited alongside, same era.
“GPT-4 vs. GPT-3.5: A concise showdown”
Anis Koubaa · 2023
Cited alongside, same era.
“Autonomous GIS: the next-generation AI-powered GIS”
Zhe Li and He Ning · 2023
Cited alongside, same era.
“Towards understanding the geospatial skills of chatgpt: Taking a geographic information systems (gis) exam”
Peter Mooney, Wencong Cui, Boyuan Guan and Levente Juhász · 2023
Cited alongside, same era.
“GPT4GEO: How a Language Model Sees the World’s Geography” Oral Presentation
Jonathan Roberts et al · 2023
Cited alongside, same era.
“Exploring and Improving the Spatial Reasoning Abilities of Large Language Models”
Manasi Sharma · 2023
Cited alongside, same era.
“Mapping with ChatGPT”
Ran Tao and Jinwen Xu · 2023
Cited alongside, same era.
Team GLM et al · 2024
Closest in time.
“GeoGPT: An assistant for understanding and processing geospatial tasks”
Yifan Zhang, Cheng Wei, Zhengting He and Wenhao Yu · 2024
Closest in time.
“MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation”
Jiaqi Chen et al · 2024
Closest in time.
“Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning” Poster presentation
Mohamed Aghzal, Erion Plaku and Ziyu Yao · 2024
Closest in time.
“MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation”
Jiaqi Chen et al · 2024
Closest in time.
“CityBench: Evaluating the Capabilities of Large Language Model as World Model”
Jie Feng et al · 2024
Closest in time.
“Where to move next: Zero-shot generalization of llms for next poi recommendation”
Shanshan Feng et al · 2024
Closest in time.
“ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools”
Team GLM et al · 2024
Closest in time.
“Evaluation of Code LLMs on Geospatial Code Generation”
Piotr Gramacki, Bruno Martins and Piotr Szymański · 2024
Closest in time.
“Correctness Comparison of ChatGPT-4, Gemini, Claude-3, and Copilot for Spatial Tasks”
H.. Hochmair, L. Juhász and T. Kemp · 2024
Closest in time.
“Can Large Language Models Generate Geospatial Code?”
Shuyang Hou et al · 2024
Closest in time.
“C-eval: A multi-level multi-discipline chinese evaluation suite for foundation models”
Yuzhen Huang et al · 2024
Closest in time.
“GPT-3.5, GPT-4, Bard, and Claude’s Performance on the Chinese Reading Comprehension Test.”
Bor-Chen Kuo, Pei-Chen Wu and Chen-Huei Liao · 2024
Closest in time.
“CMMLU: Measuring massive multitask language understanding in Chinese”
Haonan Li et al · 2024
Closest in time.
“STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis”
Wenbin Li et al · 2024
Closest in time.
“One-Shot Learning as Instruction Data Prospector for Large Language Models”
Yunshui Li et al · 2024
Closest in time.
“WILDBENCH: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild”
Bill Lin et al · 2024
Closest in time.
“Measuring Geographic Diversity of Foundation Models with a Natural Language–based Geo-guessing Experiment on GPT-4”
Zilong Liu, Krzysztof Janowicz, Kitty Currier and Meilin Shi · 2024
Closest in time.
“GeoLLM: Extracting Geospatial Knowledge from Large Language Models”
Rohin Manvi et al · 2024
Closest in time.
“GPT-4o System Card” This report outlines the safety work carried out prior to releasing GPT-4o, including external red teaming and frontier risk evaluations., 2024
OpenAI · 2024
Closest in time.
“GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning”
Zhisheng Tang and Mayank Kejriwal · 2024
Closest in time.
“Map Reading and Analysis with GPT-4(vision)”
Jinwen Xu and Ran Tao · 2024
Closest in time.
“GeoLocator: A Location-Integrated Large Multimodal Model (LMM) for Inferring Geo-Privacy”
Yifan Yang et al · 2024
Closest in time.
“GeoGPT: An assistant for understanding and processing geospatial tasks”
Yifan Zhang, Cheng Wei, Zhengting He and Wenhao Yu · 2024
Closest in time.
“NATURAL PLAN: Benchmarking LLMs on Natural Language Planning”
Huaixiu Zheng et al · 2024
Closest in time.
“AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models”
Wanjun Zhong et al · 2024
Closest in time.
“A comprehensive survey on pretrained foundation models: a history from BERT to ChatGPT”
C. Zhou, Q. Li and C. Li · 2024
Closest in time.