Are Language Models Actually Useful for Time Series Forecasting?
Mingtian Tan, Mike A. Merrill, Vinayak Gupta, Tim Althoff, Tom Hartvigsen
Abstract
Large language models (LLMs) are being applied to time series forecasting. But are language models actually useful for time series? In a series of ablation studies on three recent and popular LLM-based time series forecasting methods, we find that removing the LLM component or replacing it with a basic attention layer does not degrade forecasting performance -- in most cases, the results even improve! We also find that despite their significant computational cost, pretrained LLMs do no better than models trained from scratch, do not represent the sequential dependencies in time series, and do not assist in few-shot settings. Additionally, we explore time series encoders and find that patching and attention structures perform similarly to LLM-based forecasters.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dcdabcbb-ad20-44c0-9796-6848831b3122Cited by top-tier papers61
- UniTS: A Unified Multi-Task Time Series ModelShanghua Gao, Teddy Koker, Owen Queen, Tom Hartvigsen et al.NeurIPS 2024 · 159 citations
- TS-RAG: Retrieval-Augmented Generation based Time Series Foundation Models are Stronger Zero-Shot ForecasterKanghui Ning, Zijie Pan, Yu Liu, Yushan Jiang et al.NeurIPS 2025 · 53 citations
- Language in the Flow of Time: Time-Series-Paired Texts Weaved into a Unified Temporal NarrativeZihao Li, Xiao Lin, Zhining Liu, Jiaru Zou et al.ICLR 2026 · 41 citations
- True Zero-Shot Inference of Dynamical Systems Preserving Long-Term StatisticsChristoph Jürgen Hemmer, Daniel DurstewitzNeurIPS 2025 · 25 citations
- Multi-Modal View Enhanced Large Vision Models for Long-Term Time Series ForecastingChengAo Shen, Wenchao Yu, Ziming Zhao, Dongjin Song et al.NeurIPS 2025 · 14 citations
Builds on21
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 3,619 citations
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang et al.ICML 2022 · 2,912 citations
- iTransformer: Inverted Transformers Are Effective for Time Series ForecastingYong Liu, Tengge Hu, Haoran Zhang, Haixu Wu et al.ICLR 2024 · 1,703 citations
Related papers
- From Tokenizer Bias to Backbone Capability: A Controlled Study of LLMs for Time Series ForecastingXinyu Zhang, Shanshan Feng, Xutao Li, Kenghong Lin et al.KDD 2026 · 2 citations
- Understanding Why Large Language Models Can Be Ineffective in Time Series Analysis: The Impact of Modality AlignmentLiangwei Nathan Zheng, Chang George Dong, Wei Emma Zhang, Lin Yue et al.KDD 2025 · 1 citation
- AutoTimes: Autoregressive Time Series Forecasters via Large Language ModelsYong Liu, Guo Qin, Xiangdong Huang, Jianmin Wang et al.NeurIPS 2024 · 138 citations
- Time-LLM: Time Series Forecasting by Reprogramming Large Language ModelsMing Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu et al.ICLR 2024 · 915 citations
- A decoder-only foundation model for time-series forecastingAbhimanyu Das, Weihao Kong, Rajat Sen, Yichen ZhouICML 2024 · 601 citations
