How Tight Can PAC-Bayes be in the Small Data Regime?
Andrew Y. K. Foong, Wessel P. Bruinsma, David R. Burt, Richard E. Turner
摘要
In this paper, we investigate the question: Given a small number of datapoints, for example N = 30 , how tight can PAC-Bayes and test set bounds be made? For such small datasets, test set bounds adversely affect generalisation performance by withholding data from the training procedure. In this setting, PAC-Bayes bounds are especially attractive, due to their ability to use all the data to simultaneously learn a posterior and bound its generalisation risk. We focus on the case of i.i.d. data with a bounded loss and consider the generic PAC-Bayes theorem of Germain et al. While their theorem is known to recover many existing PAC-Bayes bounds, it is unclear what the tightest bound derivable from their framework is. For a fixed learning algorithm and dataset, we show that the tightest possible bound coincides with a bound considered by Catoni; and, in the more natural case of distributions over datasets, we establish a lower bound on the best bound achievable in expectation. Interestingly, this lower bound recovers the Chernoff test set bound if the posterior is equal to the prior. Moreover, to illustrate how tight these bounds can be, we study synthetic one-dimensional classification tasks in which it is feasible to meta-learn both the prior and the form of the bound to numerically optimise for the tightest bounds possible. We find that in this simple, controlled scenario, PAC-Bayes bounds are competitive with comparable, commonly used Chernoff test set bounds. However, the sharpest test set bounds still lead to better guarantees on the generalisation error than the PAC-Bayes bounds we consider. data-generating distributions, and aim to find the best expected bounds for this distribution achievable by an optimised algorithm. 10 We choose especially simple learning tasks — synthetic 1-dimensional binary classification problems, generated by thresholding Gaussian process (GP) samples — which allows us to fully control the task distribution and easily inspect predictive distributions visually to diagnose learning. Appendix I.1 contains full details.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- The Pick-to-Learn Algorithm: Empowering Compression for Tight Generalization Bounds and Improved Post-training PerformanceDario Paccagnan, Marco C. Campi, Simone GarattiNeurIPS 2023 · 被引用 15 次
- PAC-Bayes-Chernoff bounds for unbounded lossesIoar Casado, Luis A. Ortega Andrés, Aritz Pérez, Andrés R. MasegosaNeurIPS 2024 · 被引用 14 次
- On Margins and Generalisation for Voting ClassifiersFelix Biggs, Valentina Zantedeschi, Benjamin GuedjNeurIPS 2022 · 被引用 10 次
- Rethinking Information-theoretic Generalization: Loss Entropy Induced PAC BoundsYuxin Dong, Tieliang Gong, Hong Chen, Shujian Yu 等ICLR 2024 · 被引用 8 次
- On the Importance of Gradient Norm in PAC-Bayesian BoundsItai Gat, Yossi Adi, Alexander G. Schwing, Tamir HazanNeurIPS 2022 · 被引用 7 次
它引用的顶会 Paper4
- PACOH: Bayes-Optimal Meta-Learning with PAC-GuaranteesJonas Rothfuss, Vincent Fortuin, Martin Josifoski, Andreas KrauseICML 2021 · 被引用 136 次
- PAC-Bayes Analysis Beyond the Usual BoundsOmar Rivasplata, Ilja Kuzborskij, Csaba Szepesvári, John Shawe-TaylorNeurIPS 2020 · 被引用 101 次
- Meta-Learning Stationary Stochastic Process Prediction with Convolutional Neural ProcessesAndrew Y. K. Foong, Wessel P. Bruinsma, Jonathan Gordon, Yann Dubois 等NeurIPS 2020 · 被引用 96 次
- Second Order PAC-Bayesian Bounds for the Weighted Majority VoteAndrés R. Masegosa, Stephan Sloth Lorenzen, Christian Igel, Yevgeny SeldinNeurIPS 2020 · 被引用 48 次
相关 Paper
- A Limitation of the PAC-Bayes FrameworkRoi Livni, Shay MoranNeurIPS 2020 · 被引用 26 次
- Fast-Rate PAC-Bayesian Generalization Bounds for Meta-LearningJiechao Guan, Zhiwu LuICML 2022 · 被引用 18 次
- Improved PAC-Bayesian Bounds for Linear RegressionVera Shalaeva, Alireza Fakhrizadeh Esfahani, Pascal Germain, Mihály PetreczkyAAAI 2020 · 被引用 20 次
- A Unified View on PAC-Bayes Bounds for Meta-LearningArezou RezazadehICML 2022 · 被引用 13 次
- Generalization Bounds for Meta-Learning via PAC-Bayes and Uniform StabilityAlec Farid, Anirudha MajumdarNeurIPS 2021 · 被引用 46 次
