Learning Optimal Representations with the Decodable Information Bottleneck
Yann Dubois, Douwe Kiela, David J. Schwab, Ramakrishna Vedantam
Abstract
We address the question of characterizing and finding optimal representations for supervised learning. Traditionally, this question has been tackled using the Information Bottleneck, which compresses the inputs while retaining information about the targets, in a decoder-agnostic fashion. In machine learning, however, our goal is not compression but rather generalization, which is intimately linked to the predictive family or decoder of interest (e.g. linear classifier). We propose the Decodable Information Bottleneck (DIB) that considers information retention and compression from the perspective of the desired predictive family. As a result, DIB gives rise to representations that are optimal in terms of expected test performance and can be estimated with guarantees. Empirically, we show that the framework can be used to enforce a small generalization gap on downstream classifiers and to predict the generalization ability of neural networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0153009c-b1ff-4730-b337-4b80c4d4074eCited by top-tier papers20
- Graph Structure Learning with Variational Information BottleneckQingyun Sun, Jianxin Li, Hao Peng, Jia Wu et al.AAAI 2022 · 224 citations
- Reducing Information Bottleneck for Weakly Supervised Semantic SegmentationJungbeom Lee, Jooyoung Choi, Jisoo Mok, Sungroh YoonNeurIPS 2021 · 174 citations
- Lossy Compression for Lossless PredictionYann Dubois, Benjamin Bloem-Reddy, Karen Ullrich, Chris J. MaddisonNeurIPS 2021 · 82 citations
- Optimal Representations for Covariate ShiftYangjun Ruan, Yann Dubois, Chris J. MaddisonICLR 2022 · 77 citations
- Compressive Visual RepresentationsKuang-Huei Lee, Anurag Arnab, Sergio Guadarrama, John F. Canny et al.NeurIPS 2021 · 55 citations
Builds on3
- Fantastic Generalization Measures and Where to Find ThemYiding Jiang, Behnam Neyshabur, Hossein Mobahi, Dilip Krishnan et al.ICLR 2020 · 705 citations
- On Mutual Information Maximization for Representation LearningMichael Tschannen, Josip Djolonga, Paul K. Rubenstein, Sylvain Gelly et al.ICLR 2020 · 559 citations
- A Theory of Usable Information under Computational ConstraintsYilun Xu, Shengjia Zhao, Jiaming Song, Russell Stewart et al.ICLR 2020 · 211 citations
Related papers
- Disentangled Information BottleneckZiqi Pan, Li Niu, Jianfu Zhang, Liqing ZhangAAAI 2021 · 55 citations
- Information Retention via Learning Supplemental FeaturesZhipeng Xie, Yahe LiICLR 2024 · 1 citation
- Minimum Description Length and Generalization Guarantees for Representation LearningMilad Sefidgaran, Abdellatif Zaidi, Piotr KrasnowskiNeurIPS 2023 · 17 citations
- Cauchy-Schwarz Divergence Information Bottleneck for RegressionShujian Yu, Xi Yu, Sigurd Løkse, Robert Jenssen et al.ICLR 2024 · 16 citations
- Structured IB: Improving Information Bottleneck with Structured Feature LearningHanzhe Yang, Youlong Wu, Dingzhu Wen, Yong Zhou et al.AAAI 2025 · 6 citations
