The increased use of Large Language Models (LLMs) in geography raises substantial questions about the safety of integrating these tools across a wide range of processes and analyses, given our very limited understanding of their inner workings. In this extended abstract, we examine how LLMs process relative geographic space using activation patching, an emerging tool for mechanistic interpretability.
Paper
References (9)
03Negative results for saes on downstream tasks and deprioritising sae research2025 · AI Alignment Forum
04On the biology of a large language model2025 · Transformer Circuits Thread
06Models: Metrics and MethodsarXiv
072025. Geospatial Mechanistic Interpretabil-ity of Large Language ModelsarXiv
082024. Computing Geographically: Bridging Giscience and Geography
092024. Towards Best Practices of Activation Patching in Language