Mobility-Embedded POIs: Learning What A Place Is and How It Is Used from Human Movement
Maria Despoina Siampou, Shushman Choudhury, Shang-Ling Hsu, Neha Arora, Cyrus Shahabi
Abstract
Recent progress in geospatial foundation models highlights the importance of learning general-purpose representations for real-world locations, particularly points-of-interest (POIs) where human activity concentrates. Existing approaches, however, focus primarily on place identity derived from static textual metadata, or learn representations tied to trajectory context, which capture movement regularities rather than how places are actually used (i.e., POI's function). We argue that POI function is a missing but essential signal for general POI representations. We introduce Mobility-Embedded POIs (ME-POIs), a framework that augments POI embeddings derived, from language models with large-scale human mobility data to learn POI-centric, context-independent representations grounded in real-world usage. ME-POIs encodes individual visits as temporally contextualized embeddings and aligns them with learnable POI representations via contrastive learning to capture usage patterns across users and time. To address long-tail sparsity, we propose a novel mechanism that propagates temporal visit patterns from nearby, frequently visited POIs across multiple spatial scales. We evaluate ME-POIs on five newly proposed map enrichment tasks, testing its ability to capture both the identity and function of POIs. Across all tasks, augmenting text-based embeddings with ME-POIs consistently outperforms both text-only and mobility-only baselines. Notably, ME-POIs trained on mobility data alone can surpass text-only models on certain tasks, highlighting that POI function is a critical component of accurate and generalizable POI representations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d83c00da-1cb6-4ecd-9ab8-14667b3cce56Builds on12
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- STAN: Spatio-Temporal Attention Network for Next Location RecommendationYingtao Luo, Qiang Liu, Zhaocheng LiuWWW 2021 · 438 citations
- Multi-Scale Representation Learning for Spatial Feature Distributions using Grid CellsGengchen Mai, Krzysztof Janowicz, Bo Yan, Rui Zhu et al.ICLR 2020 · 161 citations
- Large Dual Encoders Are Generalizable RetrieversJianmo Ni, Chen Qu, Jing Lu, Zhuyun Dai et al.EMNLP 2022 · 145 citations
- Graph-Flashback Network for Next Location RecommendationXuan Rao, Lisi Chen, Yong Liu, Shuo Shang et al.KDD 2022 · 144 citations
Related papers
- POI-Enhancer: An LLM-based Semantic Enhancement Framework for POI Representation LearningJiawei Cheng, Jingyuan Wang, Yichuan Zhang, Jiahao Ji et al.AAAI 2025 · 30 citations
- MoRA: Mobility as the Backbone for Geospatial Representation Learning at ScaleYa Wen, Jixuan Cai, Qiyao Ma, Linyan Li et al.ICLR 2026 · 5 citations
- Geography-Aware Large Language Models for Next POI RecommendationWei Liu, Zhao Liu, Muzu Xie, Huaijie Zhu et al.ICDE 2026 · 8 citations
- IM-POI: Bridging ID and Multi-modal Gaps in Next POI RecommendationSiyuan Huang, Jiahui Jin, Xin Lin, Xigang Sun et al.ACM MM 2025 · 2 citations
- MGeo: Multi-Modal Geographic Language Model Pre-TrainingRuixue Ding, Boli Chen, Pengjun Xie, Fei Huang et al.SIGIR 2023 · 29 citations
