Is Your Chatbot a Tourist or a Townie? Quantifying Geographic and Localness Disparities in LLM Representations of Place CSCW022
Zihan Gao, Jacob Thebault-Spieker
Abstract
People are increasingly using Large Language Models (LLMs) for a sense of “localness,” yet their ability to accurately and equitably represent local knowledge remains unexamined. To investigate this, we conducted a large-scale evaluation using a benchmark of over 12,000 question-answer pairs spanning structured census data, local news, and social media. Our results show that performance is strongly shaped by data modality: structured tasks expose deep limitations in numerical reasoning and calibration, while open-ended prompts reveal a clear performance hierarchy favoring informal user-generated content over professionally edited prose. Our primary finding is the existence of deep, context-dependent disparities that affect communities differently. We uncover a dual geographic bias: in formal news contexts, models exhibit a strong “urban advantage,” leaving rural areas systematically underrepresented with lower semantic depth. Conversely, in social media data, models suffer an “urban penalty,” struggling to navigate the conversational complexity and slang of high-density areas. This indicates that while rural locales face a “poverty of data,” highly documented urban centers face a “poverty of precision.” We also identify a domain bias: models are more adept at handling concrete, physical questions but consistently struggle to capture the nuanced relational and cognitive dimensions of a community. This work provides the first systematic audit of localness disparities in LLMs, revealing how they reflect and risk amplifying real-world inequities. Achieving equitable local representation requires moving beyond passive evaluation to active intervention. We call for a concerted effort from the CSCW community to build richer and more ethical datasets, design interfaces that prioritize user verification over blind trust, and architect AI systems for deeper and more just engagement with place.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 1b5dd449-2f92-4cb0-903a-640de6a02edbRelated papers
- No Filter: Cultural and Socioeconomic Diversity in Contrastive Vision-Language ModelsAngéline Pouget, Lucas Beyer, Emanuele Bugliarello, Xiao Wang et al.NeurIPS 2024 · 17 citations
- Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMsGuy Mor-Lan, Omer Goldman, Matan Eyal, Adi Mayrav Gilady et al.ACL 2026 · 2 citations
- AccessEval: Benchmarking Disability Bias in Large Language ModelsSrikant Panda, Amit Agarwal, Hitesh Laxmichand PatelEMNLP 2025 · 2 citations
- AI Sees Your Location - But With A Bias Toward The Wealthy WorldJingyuan Huang, Jen-tse Huang, Ziyi Liu, Xiaoyuan Liu et al.EMNLP 2025
- XLQA: A Benchmark for Locale-Aware Multilingual Open-Domain Question AnsweringKeon-Woo Roh, Yeong-Joon Ju, Seong-Whan LeeEMNLP 2025
