Understanding the Implicit Biases of Design Choices for Time Series Foundation Models
Annan Yu, Danielle C. Maddix, Boran Han, Xiyuan Zhang, Abdul Fatir Ansari, Oleksandr Shchur, Christos Faloutsos, Andrew Gordon Wilson, Michael W. Mahoney, Bernie Wang
Abstract
Time series foundation models (TSFMs) are a class of potentially powerful, general-purpose tools for time series forecasting and related temporal tasks, but their behavior is strongly shaped by subtle inductive biases in their design. Rather than developing a new model and claiming that it is better than existing TSFMs, e.g., by winning on existing well-established benchmarks, our objective is to understand how the various ``knobs''of the training process affect model quality. Using a mix of theory and controlled empirical evaluation, we identify several design choices (patch size, embedding choice, training objective, etc.) and show how they lead to implicit biases in fundamental model properties (temporal behavior, geometric structure, how aggressively or not the model regresses to the mean, etc.); and we show how these biases can be intuitive or very counterintuitive, depending on properties of the model and data. We also illustrate in a case study on outlier handling how multiple biases can interact in complex ways; and we discuss implications of our results for learning the bitter lesson and building TSFMs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Context parroting: A simple but tough-to-beat baseline for foundation models in scientific machine learningYuanzhao Zhang, William GilpinICLR 2026 · 16 citations
- Zero-shot Forecasting by Simulation AloneBoris N. Oreshkin, Mayank Jauhari, Ravi Kiran Selvam, Malcolm Wolff et al.ICLR 2026 · 4 citations
- Universal Redundancies in Time Series Foundation ModelsAnthony Bao, Venkata Hasith Vattikuti, Jeffrey Lai, William GilpinICML 2026 · 2 citations
- Baguan-TS: dual in-context learning model for time series forecasting with covariatesLinxiao Yang, Xue Jiang, Gezheng Xu, Tian Zhou et al.ICML 2026
Builds on23
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- iTransformer: Inverted Transformers Are Effective for Time Series ForecastingYong Liu, Tengge Hu, Haoran Zhang, Haixu Wu et al.ICLR 2024 · 1,703 citations
- N-BEATS: Neural basis expansion analysis for interpretable time series forecastingBoris N. Oreshkin, Dmitri Carpov, Nicolas Chapados, Yoshua BengioICLR 2020 · 1,550 citations
- One Fits All: Power General Time Series Analysis by Pretrained LMTian Zhou, Peisong Niu, Xue Wang, Liang Sun et al.NeurIPS 2023 · 1,178 citations
- Large Language Models Are Zero-Shot Time Series ForecastersNate Gruver, Marc Finzi, Shikai Qiu, Andrew Gordon WilsonNeurIPS 2023 · 898 citations
Related papers
- It's TIME: Towards the Next Generation of Time Series Forecasting BenchmarksZhongzheng Qiao, SHENG PAN, Anni Wang, Viktoriya Zhukova et al.ICML 2026 · 9 citations
- Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?Coen Adler, Yuxin Chang, Samar Abdi, Felix Draxler et al.ICLR 2026 · 3 citations
- Zeus: Towards Tuning-Free Foundation Model for Time Series AnalysisYisong Fu, Zezhi Shao, Chengqing Yu, Yujie Li et al.ICML 2026
- Investigating Hallucinations of Time Series Foundation Models through Signal Subspace AnalysisYufeng Zou, Zijian Wang, Diego Klabjan, Han LiuNeurIPS 2025 · 2 citations
- Towards Neural Scaling Laws for Time Series Foundation ModelsQingren Yao, Chao-Han Huck Yang, Renhe Jiang, Yuxuan Liang et al.ICLR 2025
