Local Geometry Attention for Time Series Forecasting under Realistic Corruptions
Dongbin Kim, Youngjoo Park, Woojin Jeong, Jaewook Lee
Abstract
Transformers have demonstrated strong performance in time series forecasting, yet they often fail to capture the intrinsic structure of temporal data, making them susceptible to real-world noise and anomalies. Unlike in vision or language, the local geometry of temporal patterns is a critical feature in time series forecasting, but it is frequently disrupted by corruptions. In this work, we address this gap with two key contributions. First, we propose Local Geometry Attention (LGA), a novel attention mechanism theoretically grounded in local Gaussian process theory. LGA adapts to the intrinsic data geometry by learning query-specific distance metrics, enabling it to model complex temporal dependencies and enhance resilience to noise. Second, we introduce TSRBench, the first comprehensive benchmark for evaluating forecasting robustness under realistic, statistically-grounded corruptions. Experiments on TSRBench show that LGA significantly reduces performance degradation, consistently outperforming both Transformer and linear model. These results establish a foundation for developing robust time series models that can be deployed in real-world applications where data quality is not guaranteed. Our code is available at: https://github.com/dongbeank/LGA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on19
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
Related papers
- A Closer Look at Transformers for Time Series Forecasting: Understanding Why They Work and Where They StruggleYu Chen, Nathalia Céspedes, Payam M. BarnaghiICML 2025
- DeformableTST: Transformer for Time Series Forecasting without Over-reliance on PatchingDonghao Luo, Xue WangNeurIPS 2024 · 41 citations
- Learning to Rotate: Quaternion Transformer for Complicated Periodical Time Series ForecastingWeiqi Chen, Wenwei Wang, Bingqing Peng, Qingsong Wen et al.KDD 2022 · 61 citations
- Taming the Recent-Data Bias: Towards Robust Time Series Forecasting with Global ContextLonglong Xu, Zeyan Li, Xiao He, Zhaoyang Yu et al.ICML 2026
- WAVE: Weighted Autoregressive Varying Gate for Time Series ForecastingJiecheng Lu, Xu Han, Yan Sun, Shihao YangICML 2025
