Fetching the paper…

CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks · Around