Generalization Error Bounds on Deep Learning with Markov Datasets
Lan V. Truong
摘要
In this paper, we derive upper bounds on generalization errors for deep neural networks with Markov datasets. These bounds are developed based on Koltchinskii and Panchenko's approach for bounding the generalization error of combined classifiers with i.i.d. datasets. The development of new symmetrization inequalities in high-dimensional probability for Markov chains is a key element in our extension, where the absolute spectral gap of the infinitesimal generator of the Markov chain plays a key parameter in these inequalities. We also propose a simple method to convert these bounds and other similar ones in traditional deep learning and machine learning to Bayesian counterparts for both i.i.d. and Markov datasets. Extensions to m-order homogeneous Markov chains such as AR and ARMA models and mixtures of several Markov data services are given. * Use footnote for providing further information about author (webpage, alternative address)-not for acknowledging funding agencies. 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Theoretical analysis of deep neural networks for temporally dependent observationsMingliang Ma, Abolfazl SafikhaniNeurIPS 2022 · 被引用 22 次
- Streaming PCA for Markovian DataSyamantak Kumar, Purnamrita SarkarNeurIPS 2023 · 被引用 16 次
相关 Paper
- On the Stochastic Stability of Deep Markov ModelsJán Drgona, Sayak Mukherjee, Jiaxin Zhang, Frank Liu 等NeurIPS 2021 · 被引用 10 次
- PAC-Bayes Generalisation Bounds for Dynamical Systems including Stable RNNsDeividas Eringis, John Leth, Zheng-Hua Tan, Rafael Wisniewski 等AAAI 2024 · 被引用 5 次
- Why High-rank Neural Networks Generalize?: An Algebraic Framework with RKHSsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Masahiro IkedaICLR 2026 · 被引用 1 次
- On the Importance of Gradient Norm in PAC-Bayesian BoundsItai Gat, Yossi Adi, Alexander G. Schwing, Tamir HazanNeurIPS 2022 · 被引用 7 次
- Dynamics of neural scaling laws in random feature regression with powerlaw-distributed kernel eigenvaluesJakob Kramp, Javed Lindner, Moritz HeliasICML 2026 · 被引用 2 次
