Information Bottleneck Analysis of Deep Neural Networks via Lossy Compression
Ivan Butakov, Aleksander Tolmachev, Sofia Malanchuk, Anna Neopryatnaya, Alexey A. Frolov, Kirill Andreev
Abstract
The Information Bottleneck (IB) principle offers an information-theoretic framework for analyzing the training process of deep neural networks (DNNs). Its essence lies in tracking the dynamics of two mutual information (MI) values: between the hidden layer output and the DNN input/target. According to the hypothesis put forth by Shwartz-Ziv & Tishby (2017), the training process consists of two distinct phases: fitting and compression. The latter phase is believed to account for the good generalization performance exhibited by DNNs. Due to the challenging nature of estimating MI between high-dimensional random vectors, this hypothesis was only partially verified for NNs of tiny sizes or specific types, such as quantized NNs. In this paper, we introduce a framework for conducting IB analysis of general NNs. Our approach leverages the stochastic NN method proposed by Goldfeld et al. (2019) and incorporates a compression step to overcome the obstacles associated with high dimensionality. In other words, we estimate the MI between the compressed representations of high-dimensional random vectors. The proposed method is supported by both theoretical and practical justifications. Notably, we demonstrate the accuracy of our estimator through synthetic experiments featuring predefined MI values and comparison with MINE (Belghazi et al., 2018). Finally, we perform IB analysis on a close-to-real-scale convolutional DNN, which reveals new features of the MI dynamics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dd4c7852-4b62-414d-927a-a09e5ca093beCited by top-tier papers12
- Mutual Information Estimation via Normalizing FlowsIvan Butakov, Aleksander Tolmachev, Sofia Malanchuk, Anna Neopryatnaya et al.NeurIPS 2024 · 30 citations
- InfoBridge: Mutual Information estimation via Bridge MatchingSergei Kholkin, Ivan Butakov, Evgeny Burnaev, Nikita Gushchin et al.ICLR 2026 · 7 citations
- Gated Relational Alignment via Confidence-based Distillation for Efficient VLMsYanlong Chen, Amir Habibian, Luca Benini, Yawei LiICML 2026 · 5 citations
- Explaining Grokking and Information Bottleneck through Neural Collapse EmergenceKeitaro Sakamoto, Issei SatoICLR 2026 · 5 citations
- Information-Bottleneck Driven Binary Neural Network for Change DetectionKaijie Yin, Zhiyuan Zhang, Shu Kong, Tian Gao et al.ICCV 2025 · 4 citations
Builds on5
- Algorithmic Transparency via Quantitative Input Influence: Theory and Experiments with Learning SystemsAnupam Datta, Shayak Sen, Yair ZickS&P 2016 · 774 citations
- Information Bottleneck: Exact Analysis of (Quantized) Neural NetworksStephan Sloth Lorenzen, Christian Igel, Mads NielsenICLR 2022 · 24 citations
- Improved Mutual Information EstimationYoussef Mroueh, Igor Melnyk, Pierre L. Dognin, Jarret Ross et al.AAAI 2021 · 15 citations
- Invariant Representations with Stochastically Quantized Neural NetworksMattia Cerrato, Marius Köppel, Roberto Esposito, Stefan KramerAAAI 2023 · 5 citations
- Information Plane Analysis for Dropout Neural NetworksLinara Adilova, Bernhard C. Geiger, Asja FischerICLR 2023 · 1 citation
Related papers
- Cauchy-Schwarz Divergence Information Bottleneck for RegressionShujian Yu, Xi Yu, Sigurd Løkse, Robert Jenssen et al.ICLR 2024 · 16 citations
- FlowNIB: An Information Bottleneck Analysis of Bidirectional vs. Unidirectional Language ModelsMd Kowsher, Nusrat Jahan Prottasha, Shiyun Xu, Shetu Mohanto et al.ICLR 2026
- PAC-Bayes Information BottleneckZifeng Wang, Shao-Lun Huang, Ercan Engin Kuruoglu, Jimeng Sun et al.ICLR 2022 · 42 citations
- Understanding and Leveraging the Learning Phases of Neural NetworksJohannes Schneider, Mohit PrabhushankarAAAI 2024 · 5 citations
- Differentiable Information Bottleneck for Deterministic Multi-View ClusteringXiaoqiang Yan, Zhixiang Jin, Fengshou Han, Yangdong YeCVPR 2024 · 19 citations
