TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
Andreas Auer, Patrick Podest, Daniel Klotz, Sebastian Böck, Günter Klambauer, Sepp Hochreiter
Abstract
In-context learning, the ability of large language models to perform tasks using only examples provided in the prompt, has recently been adapted for time series forecasting. This paradigm enables zero-shot prediction, where past values serve as context for forecasting future values, making powerful forecasting tools accessible to non-experts and increasing the performance when training data are scarce. Most existing zero-shot forecasting approaches rely on transformer architectures, which, despite their success in language, often fall short of expectations in time series forecasting, where recurrent models like LSTMs frequently have the edge. Conversely, while LSTMs are well-suited for time series modeling due to their state-tracking capabilities, they lack strong in-context learning abilities. We introduce TiRex that closes this gap by leveraging xLSTM, an enhanced LSTM with competitive in-context learning skills. Unlike transformers, state-space models, or parallelizable RNNs such as RWKV, TiRex retains state-tracking, a critical property for long-horizon forecasting. To further facilitate its state-tracking ability, we propose a training-time masking strategy called CPM. TiRex sets a new state of the art in zero-shot time series forecasting on the HuggingFace benchmarks GiftEval and Chronos-ZS, outperforming significantly larger models including TabPFN-TS (Prior Labs), Chronos Bolt (Amazon), TimesFM (Google), and Moirai (Salesforce) across both short-and long-term forecasts. 0 500 1000 1500 2000 2500 (a) bitbrains_rnd/5T/medium Signal Prediction (5th quantile) Prediction range (2nd to 8th quantile)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar MemoriesMaurice Kraus, Felix Divo, Devendra Singh Dhami, Kristian KerstingNeurIPS 2025 · 27 citations
- CauKer: Classification Time Series Foundation Models Can Be Pretrained on Synthetic DataShifeng Xie, Vasilii Feofanov, Jianfeng Zhang, Themis Palpanas et al.ICLR 2026 · 15 citations
- Exploring Neural Granger Causality with xLSTMs: Unveiling Temporal Dependencies in Complex DataHarsh Poonia, Felix Divo, Kristian Kersting, Devendra Singh DhamiNeurIPS 2025 · 5 citations
- Adaptive Conformal Anomaly Detection with Time Series Foundation Models for Signal Monitoring.Natalia Martinez, Fearghal O'Donncha, Wesley M. Gifford, Nianjun Zhou et al.ICLR 2026 · 5 citations
- Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?Coen Adler, Yuxin Chang, Samar Abdi, Felix Draxler et al.ICLR 2026 · 3 citations
Builds on14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 3,619 citations
- Reversible Instance Normalization for Accurate Time-Series Forecasting against Distribution ShiftTaesung Kim, Jinhee Kim, Yunwon Tae, Cheonbok Park et al.ICLR 2022 · 1,020 citations
- A decoder-only foundation model for time-series forecastingAbhimanyu Das, Weihao Kong, Rajat Sen, Yichen ZhouICML 2024 · 601 citations
Related papers
- In-context Time Series PredictorJiecheng Lu, Yan Sun, Shihao YangICLR 2025
- Timer-XL: Long-Context Transformers for Unified Time Series ForecastingYong Liu, Guo Qin, Xiangdong Huang, Jianmin Wang et al.ICLR 2025
- In-Context Fine-Tuning for Time-Series Foundation ModelsMatthew Faw, Rajat Sen, Yichen Zhou, Abhimanyu DasICML 2025
- Enhancing Large Language Models for Time-Series Forecasting via Vector-Injected In-Context LearningJianqi Zhang, Jingyao Wang, Wenwen Qiang, Fanjiang Xu et al.WWW 2026
- Time-LLM: Time Series Forecasting by Reprogramming Large Language ModelsMing Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu et al.ICLR 2024 · 915 citations
