Addressing Prediction Delays in Time Series Forecasting: A Continuous GRU Approach with Derivative Regularization
Sheo Yon Jhin, Seojin Kim, Noseong Park
Abstract
Time series forecasting has been an essential field in many different application areas, including economic analysis, meteorology, and so forth. The majority of time series forecasting models are trained using the mean squared error (MSE). However, this training based on MSE causes a limitation known as prediction delay. The prediction delay, which implies the ground-truth precedes the prediction, can cause serious problems in a variety of fields, e.g., finance and weather forecasting --- as a matter of fact, predictions succeeding ground-truth observations are not practically meaningful although their MSEs can be low. This paper proposes a new perspective on traditional time series forecasting tasks and introduces a new solution to mitigate the prediction delay. We introduce a continuous-time gated recurrent unit (GRU) based on the neural ordinary differential equation (NODE) which can supervise explicit time-derivatives. We generalize the GRU architecture in a continuous-time manner and minimize the prediction delay through our time-derivative regularization. Our method outperforms in metrics such as MSE, Dynamic Time Warping (DTW) and Time Distortion Index (TDI). In addition, we demonstrate the low prediction delay of our method in a variety of datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50fefc44-f685-461f-a29d-24c9aec58e2fCited by top-tier papers3
- AliO: Output Alignment Matters in Long-Term Time Series ForecastingKwangryeol Park, Jaeho Kim, Seulki LeeNeurIPS 2025 · 2 citations
- SDE: A Simplified and Disentangled Dependency Encoding Framework for State Space Models in Time Series ForecastingZixuan Weng, Jindong Han, Wenzhao Jiang, Hao LiuKDD 2025 · 1 citation
- GFS: A Preemption-aware Scheduling Framework for GPU Clusters with Predictive Spot Instance ManagementJiaang Duan, Shenglin Xu, Shiyou Qian, Dingyu Yang et al.ASPLOS 2026 · 1 citation
Builds on10
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 5,824 citations
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 3,619 citations
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang et al.ICML 2022 · 2,912 citations
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 2,878 citations
Related papers
- Learnable Path in Neural Controlled Differential EquationsSheo Yon Jhin, Minju Jo, Seungji Kook, Noseong ParkAAAI 2023 · 13 citations
- Unveiling Delay Effects in Traffic Forecasting: A Perspective from Spatial-Temporal Delay Differential EquationsQingqing Long, Zheng Fang, Chen Fang, Chong Chen et al.WWW 2024 · 37 citations
- ODE-RSSM: Learning Stochastic Recurrent State Space Model from Irregularly Sampled DataZhaolin Yuan, Xiaojuan Ban, Zixuan Zhang, Xiaorui Li et al.AAAI 2023 · 8 citations
- EXIT: Extrapolation and Interpolation-based Neural Controlled Differential Equations for Time-series Classification and ForecastingSheo Yon Jhin, Jaehoon Lee, Minju Jo, Seungji Kook et al.WWW 2022 · 30 citations
- Neural Delay Differential EquationsQunxi Zhu, Yao Guo, Wei LinICLR 2021 · 3 citations
