Tuning the burn-in phase in training recurrent neural networks improves their performance
Julian D. Schiller, Malte Heinrich, Victor G. Lopez, Matthias A. Müller
摘要
Training recurrent neural networks (RNNs) with standard backpropagation through time (BPTT) can be challenging, especially in the presence of long input sequences. A practical alternative to reduce computational and memory overhead is to perform BPTT repeatedly over shorter segments of the training data set, corresponding to truncated BPTT. In this paper, we examine the training of RNNs when using such a truncated learning approach for time series tasks. Specifically, we establish theoretical bounds on the accuracy and performance loss when optimizing over subsequences instead of the full data sequence. This reveals that the burn-in phase of the RNN is an important tuning knob in its training, with significant impact on the performance guarantees. We validate our theoretical results through experiments on standard benchmarks from the fields of system identification and time series forecasting. In all experiments, we observe a strong influence of the burn-in phase on the training process, and proper tuning can lead to a reduction of the prediction error on the training and test data of more than 60% in some cases.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Resurrecting Recurrent Neural Networks for Long SequencesAntonio Orvieto, Samuel L. Smith, Albert Gu, Anushan Fernando 等ICML 2023 · 被引用 474 次
- Global Convergence and Stability of Stochastic Gradient DescentVivak Patel, Shushu Zhang, Bowen TianNeurIPS 2022 · 被引用 38 次
- RNNs of RNNs: Recursive Construction of Stable Assemblies of Recurrent Neural NetworksLeo Kozachkov, Michaela Ennis, Jean-Jacques E. SlotineNeurIPS 2022 · 被引用 30 次
- Improved Worst-Case Regret Bounds for Randomized Least-Squares Value IterationPriyank Agrawal, Jinglin Chen, Nan JiangAAAI 2021 · 被引用 24 次
相关 Paper
- RNNs Incrementally Evolving on an Equilibrium Manifold: A Panacea for Vanishing and Exploding Gradients?Anil Kag, Ziming Zhang, Venkatesh SaligramaICLR 2020 · 被引用 51 次
- Training Recurrent Neural Networks via Forward Propagation Through TimeAnil Kag, Venkatesh SaligramaICML 2021 · 被引用 48 次
- SkipW: Resource Adaptable RNN with Strict Upper Computational LimitTsiry Mayet, Anne Lambert, Pascal Leguyadec, Françoise Le Bolzer 等ICLR 2021
- Training Recurrent Neural Networks Online by Learning Explicit State VariablesSomjit Nath, Vincent Liu, Alan Chan, Xin Li 等ICLR 2020 · 被引用 9 次
- Balanced Resonate-and-Fire NeuronsSaya Higuchi, Sebastian Kairat, Sander M. Bohté, Sebastian OtteICML 2024 · 被引用 19 次
