MoST: A Foundation Model for Multi-modality Spatio-temporal Traffic Prediction
Ronghui Xu, Jihao Chen, Jindong Tian, Chenjuan Guo, Bin Yang
摘要
Accurate spatio-temporal traffic prediction is essential for optimizing urban traffic management and resource allocation. To reduce the cost and complexity of cross-city deployment, recent studies have explored spatio-temporal foundation models capable of accurate zero-shot prediction. However, these models are limited to single-modal data, which restricts their capacity to capture the complexity of real-world traffic dynamics. The increasing availability of multi-modality data—such as satellite imagery and points of interest (POI)—offers a promising avenue for enhancing cross-city traffic prediction by providing richer background contexts. Despite this potential, developing foundational models for multi-modality spatio-temporal prediction presents two challenges: the availability and quality of multi-modality data vary significantly across cities, with some cities lacking certain modalities or containing noisy information; and spatial patterns are highly localized and specific to individual regions, which hinders generalization. To address these challenges, we propose MoST, a foundation model for multi-modality spatio-temporal traffic prediction. We introduce a Multi-modality Refinement Module that encodes available modality data and adaptively selects task-relevant modalities while suppressing noisy modalities. Furthermore, we design a Spatio-Temporal Prediction Module that incorporates a spatial expert selection mechanism guided by multi-modality cues. This mechanism dynamically identifies region-specific spatial patterns and assigns appropriate spatial experts to model local dependencies. Finally, we conduct extensive experiments on real-world datasets to validate the superior performance and strong generalization capability of MoST.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- STKOpt: Automated Spatio-Temporal Knowledge Optimization for Traffic PredictionYayao Hong, Liyue Chen, Leye Wang, Xiuhuai Xie 等WWW 2025 · 被引用 6 次
- UniST: A Prompt-Empowered Universal Model for Urban Spatio-Temporal PredictionYuan Yuan, Jingtao Ding, Jie Feng, Depeng Jin 等KDD 2024 · 被引用 75 次
- Leveraging Heterogeneous Experts with Advantageous Pattern Memory Learning for Traffic PredictionYueyang Yao, Xingyuan Dai, Yisheng LvICDE 2025 · 被引用 3 次
- Scalable Pre-Training of Compact Urban Spatio-Temporal Predictive Models on Large-Scale Multi-Domain DataJindong Han, Hao Wang, Hui Xiong, Hao LiuVLDB 2025 · 被引用 2 次
- VisionST: Coordinating Cross-modal Traffic Prediction with Interactive Geo-image EncodingJinwen Chen, Hao Miao, Chenxi Liu, Yan Zhao 等WWW 2026
