A New Family of Generalization Bounds Using Samplewise Evaluated CMI
Fredrik Hellström, Giuseppe Durisi
摘要
We present a new family of information-theoretic generalization bounds, in which the training loss and the population loss are compared through a jointly convex function. This function is upper-bounded in terms of the disintegrated, samplewise, evaluated conditional mutual information (CMI), an information measure that depends on the losses incurred by the selected hypothesis, rather than on the hypothesis itself, as is common in probably approximately correct (PAC)-Bayesian results. We demonstrate the generality of this framework by recovering and extending previously known information-theoretic bounds. Furthermore, using the evaluated CMI, we derive a samplewise, average version of Seeger's PAC-Bayesian bound, where the convex function is the binary KL divergence. In some scenarios, this novel bound results in a tighter characterization of the population loss of deep neural networks than previous bounds. Finally, we derive high-probability versions of some of these average bounds. We demonstrate the unifying nature of the evaluated CMI bounds by using them to recover average and high-probability generalization bounds for multiclass classification with finite Natarajan dimension.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Fantastic Generalization Measures are Nowhere to be FoundMichael Gastpar, Ido Nachum, Jonathan Shafer, Thomas WeinbergerICLR 2024 · 被引用 29 次
- Tighter Information-Theoretic Generalization Bounds from SupersamplesZiqiao Wang, Yongyi MaoICML 2023 · 被引用 23 次
- Information-theoretic Generalization Analysis for Expected Calibration ErrorFutoshi Futami, Masahiro FujisawaNeurIPS 2024 · 被引用 22 次
- Minimum Description Length and Generalization Guarantees for Representation LearningMilad Sefidgaran, Abdellatif Zaidi, Piotr KrasnowskiNeurIPS 2023 · 被引用 17 次
- On -Divergence Principled Domain Adaptation: An Improved FrameworkZiqiao Wang, Yongyi MaoNeurIPS 2024 · 被引用 13 次
它引用的顶会 Paper6
- Sharpened Generalization Bounds based on Conditional Mutual Information and an Application to Noisy, Iterative AlgorithmsMahdi Haghifam, Jeffrey Negrea, Ashish Khisti, Daniel M. Roy 等NeurIPS 2020 · 被引用 124 次
- PAC-Bayes Analysis Beyond the Usual BoundsOmar Rivasplata, Ilja Kuzborskij, Csaba Szepesvári, John Shawe-TaylorNeurIPS 2020 · 被引用 101 次
- Conditioning and Processing: Techniques to Improve Information-Theoretic Generalization BoundsHassan Hafez-Kolahi, Zeinab Golgooni, Shohreh Kasaei, Mahdieh SoleymaniNeurIPS 2020 · 被引用 63 次
- Information-theoretic generalization bounds for black-box learning algorithmsHrayr Harutyunyan, Maxim Raginsky, Greg Ver Steeg, Aram GalstyanNeurIPS 2021 · 被引用 61 次
- Towards a Unified Information-Theoretic Framework for GeneralizationMahdi Haghifam, Gintare Karolina Dziugaite, Shay Moran, Daniel M. RoyNeurIPS 2021 · 被引用 38 次
相关 Paper
- On Leave-One-Out Conditional Mutual Information For GeneralizationMohamad Rida Rammal, Alessandro Achille, Aditya Golatkar, Suhas N. Diggavi 等NeurIPS 2022 · 被引用 11 次
- Controlling Multiple Errors Simultaneously with a PAC-Bayes BoundReuben Adams, John Shawe-Taylor, Benjamin GuedjNeurIPS 2024
- Slicing Mutual Information Generalization Bounds for Neural NetworksKimia Nadjahi, Kristjan H. Greenewald, Rickard Brüel Gabrielsson, Justin SolomonICML 2024 · 被引用 5 次
- Generalization in Federated Learning: A Conditional Mutual Information FrameworkZiqiao Wang, Cheng Long, Yongyi MaoICML 2025
- Evaluated CMI Bounds for Meta Learning: Tightness and ExpressivenessFredrik Hellström, Giuseppe DurisiNeurIPS 2022 · 被引用 15 次
