Do Open-Weight LLMs Reason From the Spatial Context They Are Given?
A new paper asks a question geographic evaluation has mostly skipped: once a language model is handed the right local data, does it reason from it, or fall back on what it already believes about the place? Sixteen open-weight configurations, three cities, ten seeds, and one planted false premise per case.

