Selective Learning for Deep Time Series Forecasting
Yisong Fu, Zezhi Shao, Chengqing Yu, Yujie Li, Zhulin An, Qi (Cheems) Wang, Yongjun Xu, Fei Wang
摘要
Benefiting from high capacity for capturing complex temporal patterns, deep learning (DL) has significantly advanced time series forecasting (TSF). However, deep models tend to suffer from severe overfitting due to the inherent vulnerability of time series to noise and anomalies. The prevailing DL paradigm uniformly optimizes all timesteps through the MSE loss and learns those uncertain and anomalous timesteps without difference, ultimately resulting in overfitting. To address this, we propose a novel selective learning strategy for deep TSF. Specifically, selective learning screens a subset of the whole timesteps to calculate the MSE loss in optimization, guiding the model to focus on generalizable timesteps while disregarding non-generalizable ones. Our framework introduces a dual-mask mechanism to target timesteps: (1) an uncertainty mask leveraging residual entropy to filter uncertain timesteps, and (2) an anomaly mask employing residual lower bound estimation to exclude anomalous timesteps. Extensive experiments across eight real-world datasets demonstrate that selective learning can significantly improve the predictive performance for typical state-of-the-art deep models, including 37.4% MSE reduction for Informer, 8.4% for TimesNet, and 6.5% for iTransformer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- DropoutTS: Sample-Adaptive Dropout for Robust Time Series ForecastingSiru Zhong, Yiqiu Liu, Zhiqing Cui, Zezhi Shao 等ICML 2026
- APT: Affine Prototype-Timestamp for Time Series Forecasting Under Distribution ShiftYujie Li, Zezhi Shao, Chengqing Yu, Yisong Fu 等AAAI 2026
- Zeus: Towards Tuning-Free Foundation Model for Time Series AnalysisYisong Fu, Zezhi Shao, Chengqing Yu, Yujie Li 等ICML 2026
它引用的顶会 Paper53
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 被引用 5,824 次
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 被引用 3,619 次
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang 等ICML 2022 · 被引用 2,912 次
- Connecting the Dots: Multivariate Time Series Forecasting with Graph Neural NetworksZonghan Wu, Shirui Pan, Guodong Long, Jing Jiang 等KDD 2020 · 被引用 1,738 次
相关 Paper
- Abstain Mask Retain Core: Time Series Prediction by Adaptive Masking Loss with Representation ConsistencyRenzhao Liang, Sizhe Xu, Chenggang Xie, Jingru Chen 等NeurIPS 2025 · 被引用 2 次
- From Observations to States: Latent Time Series ForecastingJie Yang, Yifan Hu, Yuante Li, Kexin Zhang 等ICML 2026 · 被引用 3 次
- Amortized Predictability-aware Training Framework for Time Series Forecasting and ClassificationXu Zhang, Peng Wang, Yichen Li, Wei WangWWW 2026
- Hierarchical Classification Auxiliary Network for Time Series ForecastingYanru Sun, Zongxia Xie, Dongyue Chen, Emadeldeen Eldele 等AAAI 2025 · 被引用 28 次
- Robust Inter-Series Dependency Modeling for Time Series Forecasting via Information-Theoretic AlignmentWuqing Yu, Weichen Guo, Jian Zhou, Shuyu Luo 等ICML 2026
