Deep Regression Representation Learning with Topology
Shihao Zhang, Kenji Kawaguchi, Angela Yao
摘要
Most works studying representation learning focus only on classification and neglect regression. Yet, the learning objectives and, therefore, the representation topologies of the two tasks are fundamentally different: classification targets class separation, leading to disconnected representations, whereas regression requires ordinality with respect to the target, leading to continuous representations. We thus wonder how the effectiveness of a regression representation is influenced by its topology, with evaluation based on the Information Bottleneck (IB) principle. The IB principle is an important framework that provides principles for learning effective representations. We establish two connections between it and the topology of regression representations. The first connection reveals that a lower intrinsic dimension of the feature space implies a reduced complexity of the representation Z. This complexity can be quantified as the conditional entropy of Z on the target Y, and serves as an upper bound on the generalization error. The second connection suggests a feature space that is topologically similar to the target space will better align with the IB principle. Based on these two connections, we introduce PH-Reg, a regularizer specific to regression that matches the intrinsic dimension and topology of the feature space with the target space. Experiments on synthetic and real-world regression tasks demonstrate the benefits of PH-Reg. Code: https: //github.com/needylove/PH-Reg .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Discretized Density-Guided Source-Free Adaptation for Continuous TargetsGezheng Xu, Qi CHEN, QIUHAO Zeng, Charles X. Ling 等ICML 2026
- Improving Deep Regression with TightnessShihao Zhang, Yuguang Yan, Angela YaoICLR 2025
- Matching without Group Barrier for Heterogeneous Treatment Effect EstimationYuguang Yan, Haolin Yang, Shihao Zhang, Weilin Chen 等ICLR 2026
它引用的顶会 Paper11
- Delving into Deep Imbalanced RegressionYuzhe Yang, Kaiwen Zha, Ying-Cong Chen, Hao Wang 等ICML 2021 · 被引用 385 次
- Topological AutoencodersMichael Moor, Max Horn, Bastian Rieck, Karsten M. BorgwardtICML 2020 · 被引用 192 次
- Intrinsic Dimension Estimation for Robust Detection of AI-Generated TextsEduard Tulchinskii, Kristian Kuznetsov, Laida Kushnareva, Daniil Cherniavskii 等NeurIPS 2023 · 被引用 163 次
- How Does Information Bottleneck Help Deep Learning?Kenji Kawaguchi, Zhun Deng, Xu Ji, Jiaoyang HuangICML 2023 · 被引用 117 次
- Intrinsic Dimension, Persistent Homology and Generalization in Neural NetworksTolga Birdal, Aaron Lou, Leonidas J. Guibas, Umut SimsekliNeurIPS 2021 · 被引用 94 次
相关 Paper
- Improving Deep Regression with Ordinal EntropyShihao Zhang, Linlin Yang, Michael Bi Mi, Xiaoxu Zheng 等ICLR 2023 · 被引用 9 次
- Cauchy-Schwarz Divergence Information Bottleneck for RegressionShujian Yu, Xi Yu, Sigurd Løkse, Robert Jenssen 等ICLR 2024 · 被引用 16 次
- Representation Learning with Conditional Information Flow MaximizationDou Hu, Lingwei Wei, Wei Zhou, Songlin HuACL 2024
- Information Retention via Learning Supplemental FeaturesZhipeng Xie, Yahe LiICLR 2024 · 被引用 1 次
- FlowNIB: An Information Bottleneck Analysis of Bidirectional vs. Unidirectional Language ModelsMd Kowsher, Nusrat Jahan Prottasha, Shiyun Xu, Shetu Mohanto 等ICLR 2026
