Tighter Expected Generalization Error Bounds via Wasserstein Distance
Borja Rodríguez Gálvez, Germán Bassi, Ragnar Thobaben, Mikael Skoglund
Abstract
This work presents several expected generalization error bounds based on the Wasserstein distance. More specifically, it introduces full-dataset, single-letter, and random-subset bounds, and their analogues in the randomized subsample setting from Steinke and Zakynthinou [1]. Moreover, when the loss function is bounded and the geometry of the space is ignored by the choice of the metric in the Wasserstein distance, these bounds recover from below (and thus, are tighter than) current bounds based on the relative entropy. In particular, they generate new, non-vacuous bounds based on the relative entropy. Therefore, these results can be seen as a bridge between works that account for the geometry of the hypothesis space and those based on the relative entropy, which is agnostic to such geometry. Furthermore, it is shown how to produce various new bounds based on different information measures (e.g., the lautum information or several -divergences) based on these bounds and how to derive similar bounds with respect to the backward channel using the presented proof techniques.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- An Exact Characterization of the Generalization Error for the Gibbs AlgorithmGholamali Aminian, Yuheng Bu, Laura Toni, Miguel R. D. Rodrigues et al.NeurIPS 2021 · 75 citations
- A unified framework for information-theoretic generalization boundsYifeng Chu, Maxim RaginskyNeurIPS 2023 · 29 citations
- Integral Probability Metrics PAC-Bayes BoundsRon Amit, Baruch Epstein, Shay Moran, Ron MeirNeurIPS 2022 · 25 citations
- Tighter Information-Theoretic Generalization Bounds from SupersamplesZiqiao Wang, Yongyi MaoICML 2023 · 23 citations
- Learning via Wasserstein-Based High Probability Generalisation BoundsPaul Viallard, Maxime Haddouche, Umut Simsekli, Benjamin GuedjNeurIPS 2023 · 16 citations
Builds on3
- Sharpened Generalization Bounds based on Conditional Mutual Information and an Application to Noisy, Iterative AlgorithmsMahdi Haghifam, Jeffrey Negrea, Ashish Khisti, Daniel M. Roy et al.NeurIPS 2020 · 124 citations
- Conditioning and Processing: Techniques to Improve Information-Theoretic Generalization BoundsHassan Hafez-Kolahi, Zeinab Golgooni, Shohreh Kasaei, Mahdieh SoleymaniNeurIPS 2020 · 63 citations
- Information-Theoretic Understanding of Population Risk Improvement with Model CompressionYuheng Bu, Weihao Gao, Shaofeng Zou, Venugopal V. VeeravalliAAAI 2020 · 18 citations
Related papers
- PAC-Bayesian Generalization Bounds for Adversarial Generative ModelsSokhna Diarra Mbacke, Florence Clerc, Pascal GermainICML 2023 · 12 citations
- Generalization Analysis of Machine Learning Algorithms via the Worst-Case Data-Generating Probability MeasureXinying Zou, Samir M. Perlaza, Iñaki Esnaola, Eitan AltmanAAAI 2024 · 28 citations
- Quantifying the Empirical Wasserstein Distance to a Set of Measures: Beating the Curse of DimensionalityNian Si, Jose H. Blanchet, Soumyadip Ghosh, Mark S. SquillanteNeurIPS 2020 · 16 citations
- Exactly Tight Information-theoretic Generalization Bounds via Binary Jensen-Shannon DivergenceYuxin Dong, Haoran Guo, Tieliang Gong, Wen Wen et al.ICML 2025
- On Leave-One-Out Conditional Mutual Information For GeneralizationMohamad Rida Rammal, Alessandro Achille, Aditya Golatkar, Suhas N. Diggavi et al.NeurIPS 2022 · 11 citations
