Amortized Predictability-aware Training Framework for Time Series Forecasting and Classification
Xu Zhang, Peng Wang, Yichen Li, Wei Wang
Abstract
Time series data are prone to noise in various domains, and training samples may contain low-predictability patterns that deviate from the normal data distribution, leading to training instability or convergence to poor local minima. Therefore, mitigating the adverse effects of low-predictability samples is crucial for time series analysis tasks such as time series forecasting (TSF) and time series classification (TSC). While many deep learning models have achieved promising performance, few consider how to identify and penalize low-predictability samples to improve model performance from the training perspective. To fill this gap, we propose a general Amortized Predictability-aware Training Framework (APTF) for both TSF and TSC. APTF introduces two key designs that enable the model to focus on high-predictability samples while still learning appropriately from low-predictability ones: (i) a Hierarchical Predictability-aware Loss (HPL) that dynamically identifies low-predictability samples and progressively expands their loss penalty as training evolves, and (ii) an amortization model that mitigates predictability estimation errors caused by model bias, further enhancing HPL's effectiveness. The code is available at https://github.com/Meteor-Stars/APTF . CCS Concepts • Information systems → Spatial-temporal systems; • Computing methodologies → Machine learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6b1a0a9b-134c-4d7c-bc48-0c00a56bdd1fCited by top-tier papers2
- Lost in the Non-convex Loss Landscape: How to Fine-tune the Large Time Series Model?Xu Zhang, Peng Wang, Wei WangICLR 2026 · 2 citations
- Diff-MN: Diffusion Parameterized MoE-NCDE for Continuous Time Series Generation with Irregular ObservationsXu Zhang, Junwei Deng, Chang Xu, Hao Li et al.ICML 2026 · 2 citations
Builds on18
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 3,619 citations
- Non-stationary Transformers: Exploring the Stationarity in Time Series ForecastingYong Liu, Haixu Wu, Jianmin Wang, Mingsheng LongNeurIPS 2022 · 1,080 citations
- How Do Vision Transformers Work?Namuk Park, Songkuk KimICLR 2022 · 653 citations
Related papers
- IdealTSF: Can Non-Ideal Data Contribute to Enhancing the Performance of Time Series Forecasting Models?Hua Wang, Jinghao Lu, Fan ZhangAAAI 2026 · 1 citation
- Selective Learning for Deep Time Series ForecastingYisong Fu, Zezhi Shao, Chengqing Yu, Yujie Li et al.NeurIPS 2025 · 10 citations
- From Observations to States: Latent Time Series ForecastingJie Yang, Yifan Hu, Yuante Li, Kexin Zhang et al.ICML 2026 · 3 citations
- RobustTSF: Towards Theory and Design of Robust Time Series Forecasting with AnomaliesHao Cheng, Qingsong Wen, Yang Liu, Liang SunICLR 2024 · 19 citations
- Hierarchical Classification Auxiliary Network for Time Series ForecastingYanru Sun, Zongxia Xie, Dongyue Chen, Emadeldeen Eldele et al.AAAI 2025 · 28 citations
