Structural Entropy Guided Probabilistic Coding
Xiang Huang, Hao Peng, Li Sun, Hui Lin, Chunyang Liu, Jiang Cao, Philip S. Yu
摘要
Probabilistic embeddings have several advantages over deterministic embeddings as they map each data point to a distribution, which better describes the uncertainty and complexity of data. Many works focus on adjusting the distribution constraint under the Information Bottleneck (IB) principle to enhance representation learning. However, these proposed regularization terms only consider the constraint of each latent variable, omitting the structural information between latent variables. In this paper, we propose a novel structural entropy-guided probabilistic coding model, named SEPC. Specifically, we incorporate the relationship between latent variables into the optimization by proposing a structural entropy regularization loss. Besides, as traditional structural information theory is not well-suited for regression tasks, we propose a probabilistic encoding tree, transferring regression tasks to classification tasks while diminishing the influence of the transformation. Experimental results across 12 natural language understanding tasks, including both classification and regression tasks, demonstrate the superior performance of SEPC compared to other state-of-the-art models in terms of effectiveness, generalization capability, and robustness to label noise. The codes and datasets are available at https://github.com/SELGroup/SEPC .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper17
- Supervised Contrastive Learning for Pre-trained Language Model Fine-tuningBeliz Gunel, Jingfei Du, Alexis Conneau, Veselin StoyanovICLR 2021 · 被引用 595 次
- Probabilistic Face EmbeddingsYichun Shi, Anil K. JainICCV 2019 · 被引用 362 次
- Graph Structure Learning with Variational Information BottleneckQingyun Sun, Jianxin Li, Hao Peng, Jia Wu 等AAAI 2022 · 被引用 224 次
- SCTNet: Single-Branch CNN with Transformer Semantic Information for Real-Time SegmentationZhengze Xu, Dongyue Wu, Changqian Yu, Xiangxiang Chu 等AAAI 2024 · 被引用 166 次
- Variational Information Bottleneck for Effective Low-Resource Fine-TuningRabeeh Karimi Mahabadi, Yonatan Belinkov, James HendersonICLR 2021 · 被引用 88 次
相关 Paper
- SECodec: Structural Entropy-based Compressive Speech Representation Codec for Speech Language ModelsLinqin Wang, Yaping Liu, Zhengtao Yu, Shengxiang Gao 等AAAI 2025 · 被引用 3 次
- An Information-Theoretic Regularizer for Lossy Neural Image CompressionYingwen Zhang, Meng Wang, Xihua Sheng, Peilin Chen 等ICCV 2025
- Representation Learning with Conditional Information Flow MaximizationDou Hu, Lingwei Wei, Wei Zhou, Songlin HuACL 2024
- Self-Supervised Learning via Maximum Entropy CodingXin Liu, Zhongdao Wang, Yali Li, Shengjin WangNeurIPS 2022 · 被引用 65 次
- Structured Probabilistic CodingDou Hu, Lingwei Wei, Yaxin Liu, Wei Zhou 等AAAI 2024
