Scaling Up Multivariate Time Series Pre-Training with Decoupled Spatial-Temporal Representations
Rui Zha, Le Zhang, Shuangli Li, Jingbo Zhou, Tong Xu, Hui Xiong, Enhong Chen
摘要
Data scale has been acknowledged as a crucial factor for enhancing the generalization and effectiveness of pre-training models. While existing methods of multivariate time series pre-training are primarily limited to a single specific dataset, scaling to a larger scenario that includes multiple diverse datasets (e.g., multi-region data) remains a substantial challenge. In this paper, we present a novel Decoupled Spatial-Temporal Representation Learning (DeSTR) framework to serve as the backbone network for investigating the data scaling capability of multivariate time series pre-training architectures. Specifically, DeSTR utilizes two separate encoders to capture both the temporal dynamics within each time series and the spatial correlations among multiple variables. The obtained representations of distinct modalities are then fed into a Spatial-Guided Temporal Transformer to equip the temporal features with spatial discriminative information. Moreover, we employ masked autoencoding as the foundational pre-training framework and introduce spacetime-agnostic augmentation to improve robustness and facilitate implicit spatiotemporal modeling. Finally, we successfully pre-train a unified time series representation learning framework on real-world datasets from three different cities. Extensive experiments are carried out on various downstream tasks to validate the performance of DeSTR, compared with three categories of state-of-the-art baselines: deep sequential models, spatial-temporal graph neural networks, and time series representation learning methods. The results clearly demonstrate the advantages of scaling multivariate time series pre-training to multiple datasets, highlighting the effectiveness of DeSTR as a general spatiotemporal learner.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Efficient Multivariate Time Series Forecasting via Calibrated Language Models with Privileged Knowledge DistillationChenxi Liu, Hao Miao, Qianxiong Xu, Shaowen Zhou 等ICDE 2025 · 被引用 15 次
- AimTS: Augmented Series and Image Contrastive Learning for Time Series ClassificationYuxuan Chen, Shanshan Huang, Yunyao Cheng, Peng Chen 等ICDE 2025 · 被引用 5 次
- Accurate and Efficient Multivariate Time Series Forecasting via Offline ClusteringYiming Niu, Jinliang Deng, Lulu Zhang, Zimu Zhou 等ICDE 2025 · 被引用 4 次
- DIFFODE: Neural ODE with Differentiable Hidden State for Irregular Time Series AnalysisYudong Zhang, Xu Wang, Xuan Yu, Zhengyang Zhou 等ICDE 2025 · 被引用 3 次
- From Teacher Pathways to Invariant Manifolds: Consensus Subspace Distillation for TSFMsZexing Zhang, Tianyang Lei, Jichao Li, Yang KeweiICML 2026
它引用的顶会 Paper21
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 被引用 5,824 次
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 被引用 3,619 次
相关 Paper
- Pre-training Enhanced Spatial-temporal Graph Neural Network for Multivariate Time Series ForecastingZezhi Shao, Zhao Zhang, Fei Wang, Yongjun XuKDD 2022 · 被引用 260 次
- Towards a General Time Series Forecasting Model with Unified Representation and Adaptive TransferYihang Wang, Yuying Qiu, Peng Chen, Kai Zhao 等ICML 2025
- GTM: A General Time-series Model for Enhanced Representation Learning of Time-Series dataCheng He, Xu Huang, Gangwei Jiang, Zhaoyi Li 等ICLR 2026 · 被引用 4 次
- Language Pre-training Guided Masking Representation Learning for Time Series ClassificationLiaoyuan Tang, Zheng Wang, Jie Wang, Guanxiong He 等AAAI 2025 · 被引用 1 次
- Time Without Time: Pseudo-Temporal Representation for Space-Time Super-ResolutionHee Min Choi, Hyoa Kang, Suji Kim, Dokwan Oh 等CVPR 2026
