Timer-XL: Long-Context Transformers for Unified Time Series Forecasting
Yong Liu, Guo Qin, Xiangdong Huang, Jianmin Wang, Mingsheng Long
摘要
We present Timer-XL, a causal Transformer for unified time series forecasting. To uniformly predict multidimensional time series, we generalize next token prediction, predominantly adopted for 1D token sequences, to multivariate next token prediction. The paradigm formulates various forecasting tasks as a long-context prediction problem. We opt for decoder-only Transformers that capture causal dependencies from varying-length contexts for unified forecasting, making predictions on non-stationary univariate time series, multivariate series with complicated dynamics and correlations, as well as covariate-informed contexts that include exogenous variables. Technically, we propose a universal TimeAttention to capture fine-grained intra-and inter-series dependencies of flattened time series tokens (patches), which is further enhanced by deft position embedding for temporal causality and variable equivalence. Timer-XL achieves state-of-the-art performance across task-specific forecasting benchmarks through a unified approach. Based on large-scale pre-training, Timer-XL achieves state-of-the-art zero-shot performance, making it a promising architecture for pre-trained time series models. Code is available at this repository: https://github.com/thuml/Timer-XL . * Equal Contribution Published as a conference paper at ICLR 2025 Tokens of Modality # Tokens of Context Unified Time Series Forecasting Answer the following mathematical questions: Q: If you have 12 apples and you give 5 to your friend, how many apples do you have now? A: The answer is 7.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper28
- TS-RAG: Retrieval-Augmented Generation based Time Series Foundation Models are Stronger Zero-Shot ForecasterKanghui Ning, Zijie Pan, Yu Liu, Yushan Jiang 等NeurIPS 2025 · 被引用 53 次
- xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar MemoriesMaurice Kraus, Felix Divo, Devendra Singh Dhami, Kristian KerstingNeurIPS 2025 · 被引用 27 次
- SciTS: Scientific Time Series Understanding and Generation with LLMsWen Wu, Ziyang Zhang, Liwei Liu, Xuenan Xu 等ICLR 2026 · 被引用 11 次
- BEDTime: A Unified Benchmark for Automatically Describing Time SeriesMedhasweta Sen, Zachary Gottesman, Jiaxing Qiu, C. Bayan Bruss 等ICML 2026 · 被引用 8 次
- TRACE: Grounding Time Series in Context for Multimodal Embedding and RetrievalJialin Chen, Ziyu Zhao, Gaukhar Nurbek, Aosong Feng 等NeurIPS 2025 · 被引用 8 次
它引用的顶会 Paper21
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 被引用 5,824 次
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra 等NeurIPS 2022 · 被引用 5,493 次
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 被引用 3,619 次
- iTransformer: Inverted Transformers Are Effective for Time Series ForecastingYong Liu, Tengge Hu, Haoran Zhang, Haixu Wu 等ICLR 2024 · 被引用 1,703 次
相关 Paper
- TimeXer: Empowering Transformers for Time Series Forecasting with Exogenous VariablesYuxuan Wang, Haixu Wu, Jiaxiang Dong, Guo Qin 等NeurIPS 2024 · 被引用 536 次
- UniTS: A Unified Multi-Task Time Series ModelShanghua Gao, Teddy Koker, Owen Queen, Tom Hartvigsen 等NeurIPS 2024 · 被引用 159 次
- TimePerceiver: An Encoder-Decoder Framework for Generalized Time-Series ForecastingJaebin Lee, Hankook LeeNeurIPS 2025 · 被引用 1 次
- Unified Training of Universal Time Series Forecasting TransformersGerald Woo, Chenghao Liu, Akshat Kumar, Caiming Xiong 等ICML 2024 · 被引用 513 次
- TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context LearningAndreas Auer, Patrick Podest, Daniel Klotz, Sebastian Böck 等NeurIPS 2025 · 被引用 126 次
