2026
LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning
AAAI 2026technical
Large language models (LLMs) have been widely evaluated on macro-scale geographic tasks, such as global factual recall, event summarization, and regional reasoning. Yet, their ability to handle hyper-local knowledge remains poorly understood. This gap is increasingly consequential as real-world appl