Reconciling Geospatial Prediction and Retrieval via Sparse Representations
Yi Li, Yuanlong Chen, Weiming Huang, Xiaoli Li, Gao Cong
Abstract
Urban computing harnesses big data to decode complex urban dynamics and revo-lutionize location-based services. Traditional approaches have treated geospatial prediction tasks (e.g., estimating socio-economic indicators) and retrieval tasks (e.g., querying geographic objects) as isolated challenges, necessitating separate models with distinct training objectives. This fragmentation imposes significant computational burdens and limits cross-task synergy, despite advances in representation learning and multi-task foundation models. We present UrbanSparse, a pioneering framework that unifies geospatial prediction and retrieval through a novel sparse-dense representation architecture. By synergistically combining these tasks, UrbanSparse eliminates redundant systems while amplifying their mutual strengths. Our approach introduces two innovations: (1) Bloom filter-based sparse encodings that compress high-sparsity geographic queries and fine-grained text terms for retrieval effectiveness, and (2) a dense semantic codebook that captures granular urban features to boost prediction accuracy. A two-view contrastive learning mechanism further bridges urban objects, regions, and contexts. Experiments on real-world datasets demonstrate 25.16% gains in prediction accuracy and 20.76% improvements in retrieval precision over state-of-the-art baselines, alongside 65.97% faster training. These advantages position UrbanSparse as a scalable solution for large urban datasets. To our knowledge, this is the first unified framework bridging
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on13
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 1,663 citations
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis et al.EMNLP 2020 · 142 citations
- TimeCMA: Towards LLM-Empowered Multivariate Time Series Forecasting via Cross-Modality AlignmentChenxi Liu, Qianxiong Xu, Hao Miao, Sun Yang et al.AAAI 2025 · 141 citations
- GeoLLM: Extracting Geospatial Knowledge from Large Language ModelsRohin Manvi, Samar Khanna, Gengchen Mai, Marshall Burke et al.ICLR 2024 · 104 citations
- A Unified Replay-Based Continuous Learning Framework for Spatio-Temporal Prediction on Streaming DataHao Miao, Yan Zhao, Chenjuan Guo, Bin Yang et al.ICDE 2024 · 64 citations
Related papers
- UrbanFusion: Stochastic Multimodal Fusion for Contrastive Learning of Robust Spatial RepresentationsDominik J. Mühlematter, Lin Che, Ye Hong, Martin Raubal et al.ICML 2026
- UrbanMoE: A Sparse Multi-Modal Mixture-of-Experts Framework for Multi-Task Urban Region ProfilingPingping Liu, Jiamiao Liu, Zijian Zhang, Hao Miao et al.WWW 2026
- Geolocation Representation from Large Language Models Are Generic Enhancers for Spatio-Temporal LearningJunlin He, Tong Nie, Wei MaAAAI 2025 · 18 citations
- UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain AdaptationSiru Zhong, Xixuan Hao, Yibo Yan, Ying Zhang et al.ACM MM 2024 · 8 citations
- UrbanCLIP: Learning Text-enhanced Urban Region Profiling with Contrastive Language-Image Pretraining from the WebYibo Yan, Haomin Wen, Siru Zhong, Wei Chen et al.WWW 2024 · 124 citations
