Generalization Bounds via Conditional f-Information
Ziqiao Wang, Yongyi Mao
Abstract
In this work, we introduce novel information-theoretic generalization bounds using the conditional -information framework, an extension of the traditional conditional mutual information (MI) framework. We provide a generic approach to derive generalization bounds via -information in the supersample setting, applicable to both bounded and unbounded loss functions. Unlike previous MI-based bounds, our proof strategy does not rely on upper bounding the cumulant-generating function (CGF) in the variational formula of MI. Instead, we set the CGF or its upper bound to zero by carefully selecting the measurable function invoked in the variational formula. Although some of our techniques are partially inspired by recent advances in the coin-betting framework (e.g., Jang et al. (2023)), our results are independent of any previous findings from regret guarantees of online gambling algorithms. Additionally, our newly derived MI-based bound recovers many previous results and improves our understanding of their potential limitations. Finally, we empirically compare various -information measures for generalization, demonstrating the improvement of our new bounds over the previous bounds.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7a2aaf59-9260-4adb-8b76-9d0856786c99Cited by top-tier papers2
- Exactly Tight Information-theoretic Generalization Bounds via Binary Jensen-Shannon DivergenceYuxin Dong, Haoran Guo, Tieliang Gong, Wen Wen et al.ICML 2025
- Generalization in Federated Learning: A Conditional Mutual Information FrameworkZiqiao Wang, Cheng Long, Yongyi MaoICML 2025
Builds on19
- Sharpened Generalization Bounds based on Conditional Mutual Information and an Application to Noisy, Iterative AlgorithmsMahdi Haghifam, Jeffrey Negrea, Ashish Khisti, Daniel M. Roy et al.NeurIPS 2020 · 124 citations
- How Does Information Bottleneck Help Deep Learning?Kenji Kawaguchi, Zhun Deng, Xu Ji, Jiaoyang HuangICML 2023 · 117 citations
- Conditioning and Processing: Techniques to Improve Information-Theoretic Generalization BoundsHassan Hafez-Kolahi, Zeinab Golgooni, Shohreh Kasaei, Mahdieh SoleymaniNeurIPS 2020 · 63 citations
- Information-theoretic generalization bounds for black-box learning algorithmsHrayr Harutyunyan, Maxim Raginsky, Greg Ver Steeg, Aram GalstyanNeurIPS 2021 · 61 citations
- Tighter Expected Generalization Error Bounds via Wasserstein DistanceBorja Rodríguez Gálvez, Germán Bassi, Ragnar Thobaben, Mikael SkoglundNeurIPS 2021 · 52 citations
Related papers
- Tighter Information-Theoretic Generalization Bounds from SupersamplesZiqiao Wang, Yongyi MaoICML 2023 · 23 citations
- Evaluated CMI Bounds for Meta Learning: Tightness and ExpressivenessFredrik Hellström, Giuseppe DurisiNeurIPS 2022 · 15 citations
- On Leave-One-Out Conditional Mutual Information For GeneralizationMohamad Rida Rammal, Alessandro Achille, Aditya Golatkar, Suhas N. Diggavi et al.NeurIPS 2022 · 11 citations
- A unified framework for information-theoretic generalization boundsYifeng Chu, Maxim RaginskyNeurIPS 2023 · 29 citations
- A New Family of Generalization Bounds Using Samplewise Evaluated CMIFredrik Hellström, Giuseppe DurisiNeurIPS 2022 · 32 citations
