PAC-Bayes Generalisation Bounds for Dynamical Systems including Stable RNNs
Deividas Eringis, John Leth, Zheng-Hua Tan, Rafael Wisniewski, Mihály Petreczky
Abstract
In this paper, we derive a PAC-Bayes bound on the generalisation gap, in a supervised time-series setting for a special class of discrete-time non-linear dynamical systems. This class includes stable recurrent neural networks (RNN), and the motivation for this work was its application to RNNs. In order to achieve the results, we impose some stability constraints, on the allowed models. Here, stability is understood in the sense of dynamical systems. For RNNs, these stability conditions can be expressed in terms of conditions on the weights. We assume the processes involved are essentially bounded and the loss functions are Lipschitz. The proposed bound on the generalisation gap depends on the mixing coefficient of the data distribution, and the essential supremum of the data. Furthermore, the bound converges to zero as the dataset size increases. In this paper, we 1) formalize the learning problem, 2) derive a PAC-Bayesian error bound for such systems, 3) discuss various consequences of this error bound, and 4) show an illustrative example, with discussions on computing the proposed bound. Unlike other available bounds the derived bound holds for non i.i.d. data (time-series) and it does not grow with the number of steps of the RNN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 65df2df4-2b5e-4a40-8d91-2ba3ebfe501cCited by top-tier papers2
- Generalization Bounds for Kolmogorov-Arnold Networks (KANs) and Enhanced KANs with Lower Lipschitz ComplexityPengqi Li, Lizhong Ding, Jiarun Fu, Chunhui Zhang et al.NeurIPS 2025 · 8 citations
- PAC-Bayesian Error Bound, via Rényi Divergence, for a Class of Linear Time-Invariant State-Space ModelsDeividas Eringis, John Leth, Zheng-Hua Tan, Rafal Wisniewski et al.ICML 2024 · 2 citations
Builds on7
- Transformers as Algorithms: Generalization and Stability in In-context LearningYingcong Li, Muhammed Emrullah Ildiz, Dimitris Papailiopoulos, Samet OymakICML 2023 · 242 citations
- Naive Exploration is Optimal for Online LQRMax Simchowitz, Dylan J. FosterICML 2020 · 209 citations
- Logarithmic Regret Bound in Partially Observable Linear Dynamical SystemsSahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima AnandkumarNeurIPS 2020 · 106 citations
- On Empirical Risk Minimization with Dependent and Heavy-Tailed DataAbhishek Roy, Krishnakumar Balasubramanian, Murat A. ErdogduNeurIPS 2021 · 22 citations
- Implicit Bias of Linear RNNsMelikasadat Emami, Mojtaba Sahraee-Ardakan, Parthe Pandit, Sundeep Rangan et al.ICML 2021 · 14 citations
Related papers
- Framing RNN as a kernel method: A neural ODE approachAdeline Fermanian, Pierre Marion, Jean-Philippe Vert, Gérard BiauNeurIPS 2021 · 34 citations
- Improved PAC-Bayesian Bounds for Linear RegressionVera Shalaeva, Alireza Fakhrizadeh Esfahani, Pascal Germain, Mihály PetreczkyAAAI 2020 · 20 citations
- HyRNN: Hybrid Recurrent Neural Networks for Approximating Hybrid Dynamical SystemsRicardo G. SanfeliceAAAI 2026
- Lipschitz Recurrent Neural NetworksN. Benjamin Erichson, Omri Azencot, Alejandro F. Queiruga, Liam Hodgkinson et al.ICLR 2021 · 32 citations
- On the Role of Noise in the Sample Complexity of Learning Recurrent Neural Networks: Exponential Gaps for Long SequencesAlireza Fathollah Pour, Hassan AshtianiNeurIPS 2023
