A Theory of Usable Information under Computational Constraints
Yilun Xu, Shengjia Zhao, Jiaming Song, Russell Stewart, Stefano Ermon
Abstract
We propose a new framework for reasoning about information in complex systems. Our foundation is based on a variational extension of Shannon's information theory that takes into account the modeling power and computational constraints of the observer. The resulting predictive -information encompasses mutual information and other notions of informativeness such as the coefficient of determination. Unlike Shannon's mutual information and in violation of the data processing inequality, -information can be created through computation. This is consistent with deep neural networks extracting hierarchies of progressively more informative features in representation learning. Additionally, we show that by incorporating computational constraints, -information can be reliably estimated from data even in high dimensions with PAC-style guarantees. Empirically, we demonstrate predictive -information is more effective than mutual information for structure learning and fair representation learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d5bebf59-c095-4996-84ea-942f8baddcc9Cited by top-tier papers68
- On Mutual Information Maximization for Representation LearningMichael Tschannen, Josip Djolonga, Paul K. Rubenstein, Sylvain Gelly et al.ICLR 2020 · 559 citations
- Understanding Dataset Difficulty with V-Usable InformationKawin Ethayarajh, Yejin Choi, Swabha SwayamdiptaICML 2022 · 337 citations
- LEACE: Perfect linear concept erasure in closed formNora Belrose, David Schneider-Joseph, Shauli Ravfogel, Ryan Cotterell et al.NeurIPS 2023 · 305 citations
- RePrompt: Automatic Prompt Editing to Refine AI-Generative Art Towards Precise ExpressionsYunlong Wang, Shuyuan Shen, Brian Y. LimCHI 2023 · 118 citations
- Lossy Compression for Lossless PredictionYann Dubois, Benjamin Bloem-Reddy, Karen Ullrich, Chris J. MaddisonNeurIPS 2021 · 82 citations
Builds on1
Related papers
- Information Plane Analysis for Dropout Neural NetworksLinara Adilova, Bernhard C. Geiger, Asja FischerICLR 2023 · 1 citation
- Readout Representation: Redefining Neural Codes by Input RecoveryShunsuke Onoo, Yoshihiro Nagano, Yukiyasu KamitaniICLR 2026 · 2 citations
- Invariant Representations with Stochastically Quantized Neural NetworksMattia Cerrato, Marius Köppel, Roberto Esposito, Stefan KramerAAAI 2023 · 5 citations
- Max-Sliced Mutual InformationDor Tsur, Ziv Goldfeld, Kristjan H. GreenewaldNeurIPS 2023 · 20 citations
- Connecting Jensen-Shannon and Kullback-Leibler Divergences: A New Bound for Representation LearningReuben Dorent, Polina Golland, William (Sandy) WellsNeurIPS 2025 · 7 citations
