SwiftTS: A Swift Selection Framework for Time Series Pre-trained Models via Multi-task Meta-Learning
Tengxue Zhang, Biao Ouyang, Yang Shu, Xinyang Chen, Chenjuan Guo, Bin Yang
Abstract
Pre-trained models exhibit strong generalization to various downstream tasks. However, given the numerous models available in the model hub, identifying the most suitable one by individually fine-tuning is time-consuming. In this paper, we propose SwiftTS, a swift selection framework for time series pre-trained models. To avoid expensive forward propagation through all candidates, SwiftTS adopts a learning-guided approach that leverages historical dataset-model performance pairs across diverse horizons to predict model performance on unseen datasets. It employs a lightweight dual-encoder architecture that embeds time series and candidate models with rich characteristics, computing patchwise compatibility scores between data and model embeddings for efficient selection. To further enhance the generalization across datasets and horizons, we introduce a horizon-adaptive expert composition module that dynamically adjusts expert weights, and the transferable cross-task learning with cross-dataset and cross-horizon task sampling to enhance out-of-distribution (OOD) robustness. Extensive experiments on 14 downstream datasets and 8 pre-trained models demonstrate that SwiftTS achieves state-of-the-art performance in time series pre-trained model selection. The code and datasets are available at https://github.com/decisionintelligence/SwiftTS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- ARROW: An Adaptive Rollout and Routing Method for Global Weather ForecastingJindong Tian, Yifei Ding, Ronghui Xu, Hao Miao et al.ICLR 2026 · 14 citations
- Unlocking the Value of Text: Event-Driven Reasoning and Multi-Level Alignment for Time Series ForecastingSiyuan Wang, Peng Chen, Yihang Wang, Wanghui Qiu et al.ICLR 2026 · 4 citations
Builds on25
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 5,863 citations
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- Spatial-Temporal Synchronous Graph Convolutional Networks: A New Framework for Spatial-Temporal Network Data ForecastingChao Song, Youfang Lin, Shengnan Guo, Huaiyu WanAAAI 2020 · 1,659 citations
Related papers
- Unified Transferability Metrics for Time Series Foundation ModelsWeiyang Zhang, Xinyang Chen, Xiucheng Li, Kehai Chen et al.NeurIPS 2025 · 5 citations
- Neural Architecture and Hyperparameter Selection Through Meta-Learning on Time SeriesErfan Moeini, Christopher Vox, Marie Anastacio, Wadie Skaf et al.AAAI 2026 · 1 citation
- TSPulse: Tiny Pre-Trained Models with Disentangled Representations for Rapid Time-Series AnalysisVijay Ekambaram, Subodh Kumar, Arindam Jati, Sumanta Mukherjee et al.ICLR 2026 · 13 citations
- UniTS: A Unified Multi-Task Time Series ModelShanghua Gao, Teddy Koker, Owen Queen, Tom Hartvigsen et al.NeurIPS 2024 · 159 citations
- FAT: Frequency-Aware Pretraining for Enhanced Time-Series Representation LearningRui Cheng, Xiangfei Jia, Qing Li, Rong Xing et al.KDD 2025 · 2 citations
