REMEDI: Corrective Transformations for Improved Neural Entropy Estimation
Viktor Nilsson, Anirban Samaddar, Sandeep Madireddy, Pierre Nyquist
摘要
Information theoretic quantities play a central role in machine learning. The recent surge in the complexity of data and models has increased the demand for accurate estimation of these quantities. However, as the dimension grows the estimation presents significant challenges, with existing methods struggling already in relatively low dimensions. To address this issue, in this work, we introduce for efficient and accurate estimation of differential entropy, a fundamental information theoretic quantity. The approach combines the minimization of the cross-entropy for simple, adaptive base models and the estimation of their deviation, in terms of the relative entropy, from the data density. Our approach demonstrates improvement across a broad spectrum of estimation tasks, encompassing entropy estimation on both synthetic and natural data. Further, we extend important theoretical consistency results to a more generalized setting required by our approach. We illustrate how the framework can be naturally extended to information theoretic supervised learning models, with a specific focus on the Information Bottleneck approach. It is demonstrated that the method delivers better accuracy compared to the existing methods in Information Bottleneck. In addition, we explore a natural connection between and generative modeling using rejection sampling and Langevin dynamics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
- A Differential Entropy Estimator for Training Neural NetworksGeorg Pichler, Pierre Jean A. Colombo, Malik Boudiaf, Günther Koliander 等ICML 2022 · 被引用 28 次
相关 Paper
- Connecting Jensen-Shannon and Kullback-Leibler Divergences: A New Bound for Representation LearningReuben Dorent, Polina Golland, William (Sandy) WellsNeurIPS 2025 · 被引用 7 次
- InfoBridge: Mutual Information estimation via Bridge MatchingSergei Kholkin, Ivan Butakov, Evgeny Burnaev, Nikita Gushchin 等ICLR 2026 · 被引用 7 次
- Local Intrinsic Dimensional EntropyRohan Ghosh, Mehul MotaniAAAI 2023 · 被引用 2 次
- Information Estimation with Discrete DiffusionAlberto Foresti, Giulio Franzese, Pietro MichiardiICLR 2026
- FlowNIB: An Information Bottleneck Analysis of Bidirectional vs. Unidirectional Language ModelsMd Kowsher, Nusrat Jahan Prottasha, Shiyun Xu, Shetu Mohanto 等ICLR 2026
