Urban Region Embedding via Multi-View Contrastive Prediction
Zechen Li, Weiming Huang, Kai Zhao, Min Yang, Yongshun Gong, Meng Chen
Abstract
Recently, learning urban region representations utilizing multi-modal data (information views) has become increasingly popular, for deep understanding of the distributions of various socioeconomic features in cities. However, previous methods usually blend multi-view information in a posteriors stage, falling short in learning coherent and consistent representations across different views. In this paper, we form a new pipeline to learn consistent representations across varying views, and propose the multi-view Contrastive Prediction model for urban Region embedding (ReCP), which leverages the multiple information views from point-of-interest (POI) and human mobility data. Specifically, ReCP comprises two major modules, namely an intra-view learning module utilizing contrastive learning and feature reconstruction to capture the unique information from each single view, and inter-view learning module that perceives the consistency between the two views using a contrastive prediction learning scheme. We conduct thorough experiments on two downstream tasks to assess the proposed model, i.e., land use clustering and region popularity prediction. The experimental results demonstrate that our model outperforms state-of-the-art baseline methods significantly in urban region representation learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- VecCity: A Taxonomy-guided Library for Map Entity Representation Learning [Experiment, Analysis & Benchmark]Wentao Zhang, Jingyuan Wang, Yifan Yang, Leong Hou UVLDB 2025 · 7 citations
- MoRA: Mobility as the Backbone for Geospatial Representation Learning at ScaleYa Wen, Jixuan Cai, Qiyao Ma, Linyan Li et al.ICLR 2026 · 5 citations
- FlexiReg: Flexible Urban Region Representation LearningFengze Sun, Yanchuan Chang, Egemen Tanin, Shanika Karunasekera et al.KDD 2025 · 3 citations
- UrbanMind: Urban Dynamics Prediction with Multifaceted Spatial-Temporal Large Language ModelsYuhang Liu, Yingxue Zhang, Xin Zhang, Ling Tian et al.KDD 2025 · 3 citations
- Improving Region Representation Learning from Urban Imagery with Noisy Long-Caption SupervisionYimei Zhang, Guojiang Shen, Kaili Ning, Tongwei Ren et al.AAAI 2026 · 3 citations
Builds on6
- Self-supervised Learning from a Multi-view PerspectiveYao-Hung Hubert Tsai, Yue Wu, Ruslan Salakhutdinov, Louis-Philippe MorencyICLR 2021 · 232 citations
- Urban2Vec: Incorporating Street View Imagery and POIs for Multi-Modal Urban Neighborhood EmbeddingZhecheng Wang, Haoyuan Li, Ram RajagopalAAAI 2020 · 113 citations
- Heterogeneous Region Embedding with Prompt LearningSilin Zhou, Dan He, Lisi Chen, Shuo Shang et al.AAAI 2023 · 43 citations
- Urban Region Representation Learning with OpenStreetMap Building FootprintsYi Li, Weiming Huang, Gao Cong, Hao Wang et al.KDD 2023 · 38 citations
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie et al.CVPR 2020
Related papers
- Multi-View Urban Region Embedding via Commonality-Specificity DisentanglementZechen Li, Hongwei Jia, Kai Zhao, Weiming Huang et al.KDD 2026
- Comprehensive Urban Region Representation Learning via Multi-View Joint Learning and Contrastive LearningYingde Lin, Yuanbo Xu, Lu Jiang, Pengyang WangAAAI 2026
- Urban Region Pre-training and Prompting: A Graph-based ApproachJiahui Jin, Yifan Song, Dong Kan, Haojia Zhu et al.KDD 2025 · 1 citation
- Beyond the First Law of Geography: Learning Representations of Satellite Imagery by Leveraging Point-of-InterestsYanxin Xi, Tong Li, Huandong Wang, Yong Li et al.WWW 2022 · 86 citations
- Automated Spatio-Temporal Graph Contrastive LearningQianru Zhang, Chao Huang, Lianghao Xia, Zheng Wang et al.WWW 2023 · 72 citations
