Local Intrinsic Dimensional Entropy
Rohan Ghosh, Mehul Motani
摘要
Most entropy measures depend on the spread of the probability distribution over the sample space |X|, and the maximum entropy achievable scales proportionately with the sample space cardinality |X|. For a finite |X|, this yields robust entropy measures which satisfy many important properties, such as invariance to bijections, while the same is not true for continuous spaces (where |X|=infinity). Furthermore, since R and R^d (d in Z+) have the same cardinality (from Cantor's correspondence argument), cardinality-dependent entropy measures cannot encode the data dimensionality. In this work, we question the role of cardinality and distribution spread in defining entropy measures for continuous spaces, which can undergo multiple rounds of transformations and distortions, e.g., in neural networks. We find that the average value of the local intrinsic dimension of a distribution, denoted as ID-Entropy, can serve as a robust entropy measure for continuous spaces, while capturing the data dimensionality. We find that ID-Entropy satisfies many desirable properties and can be extended to conditional entropy, joint entropy and mutual-information variants. ID-Entropy also yields new information bottleneck principles and also links to causality. In the context of deep learning, for feedforward architectures, we show, theoretically and empirically, that the ID-Entropy of a hidden layer directly controls the generalization gap for both classifiers and auto-encoders, when the target function is Lipschitz continuous. Our work primarily shows that, for continuous spaces, taking a structural rather than a statistical approach yields entropy measures which preserve intrinsic data dimensionality, while being relevant for studying various architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Deep Regression Representation Learning with TopologyShihao Zhang, Kenji Kawaguchi, Angela YaoICML 2024 · 被引用 4 次
- Intrinsic Dimension Correlation: uncovering nonlinear connections in multimodal representationsLorenzo Basile, Santiago Acevedo, Luca Bortolussi, Fabio Anselmi 等ICLR 2025 · 被引用 1 次
- CuMPerLay: Learning Cubical Multiparameter Persistence VectorizationsCaner Korkmaz, Brighton Nuwagira, Baris Coskunuzer, Tolga BirdalICCV 2025 · 被引用 1 次
- Tab-PET: Graph-Based Positional Encodings for Tabular TransformersYunze Leng, Rohan Ghosh, Mehul MotaniAAAI 2026
它引用的顶会 Paper3
- On the geometry of generalization and memorization in deep neural networksCory Stephenson, Suchismita Padhy, Abhinav Ganesh, Yue Hui 等ICLR 2021 · 被引用 95 次
- Intrinsic Dimension, Persistent Homology and Generalization in Neural NetworksTolga Birdal, Aaron Lou, Leonidas J. Guibas, Umut SimsekliNeurIPS 2021 · 被引用 94 次
- The Effect of the Intrinsic Dimension on the Generalization of Quadratic ClassifiersFabian Latorre, Leello Tadesse Dadi, Paul Rolland, Volkan CevherNeurIPS 2021 · 被引用 13 次
相关 Paper
- How Does Information Bottleneck Help Deep Learning?Kenji Kawaguchi, Zhun Deng, Xu Ji, Jiaoyang HuangICML 2023 · 被引用 117 次
- Cauchy-Schwarz Divergence Information Bottleneck for RegressionShujian Yu, Xi Yu, Sigurd Løkse, Robert Jenssen 等ICLR 2024 · 被引用 16 次
- Information-Theoretic Generalization Bounds for VAEs: A Role of Encoder and Latent VariableFutoshi Futami, Masahiro FujisawaICML 2026
- Information Bottleneck: Exact Analysis of (Quantized) Neural NetworksStephan Sloth Lorenzen, Christian Igel, Mads NielsenICLR 2022 · 被引用 24 次
- Learning Diverse and Discriminative Representations via the Principle of Maximal Coding Rate ReductionYaodong Yu, Kwan Ho Ryan Chan, Chong You, Chaobing Song 等NeurIPS 2020 · 被引用 265 次
