1 paper
Zihan Gao, Yifei Xu, Jacob Thebault-Spieker
Large language models (LLMs) have been widely evaluated on macro-scale geographic tasks, such as global factual recall, event summarization, and regional reasoning. Yet, their abil…