Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Features Model
Hien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat Ho
摘要
The current paradigm of training deep neural networks for classification tasks includes minimizing the empirical risk, pushing the training loss value towards zero even after the training classification error has vanished. In this terminal phase of training, it has been observed that the last-layer features collapse to their class-means and these class-means converge to the vertices of a simplex Equiangular Tight Frame (ETF). This phenomenon is termed as Neural Collapse (N C). However, this characterization only holds in classbalanced datasets where every class has the same number of training samples. When the training dataset is class-imbalanced, some N C properties will no longer hold true, for example, the geometry of class-means will skew away from the simplex ETF. In this paper, we generalize N C to imbalanced regime for cross-entropy loss under the unconstrained ReLU features model. We demonstrate that while the within-class features collapse property still holds in this setting, the class-means will converge to a structure consisting of orthogonal vectors with lengths dependent on the number of training samples. Furthermore, we find that the classifier weights (i.e., the last-layer linear classifier) are aligned to the scaled and centered class-means, with scaling factors dependent on the number of training samples of each class. This generalizes N C in the class-balanced setting. We empirically validate our results through experiments on practical architectures and dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Unifying Low Dimensional Spectra in Deep LearningConnall Garrod, Jonathan KeatingICML 2026 · 被引用 12 次
- Neural Collapse is Globally Optimal in Deep Regularized ResNets and TransformersPeter Súkeník, Christoph H. Lampert, Marco MondelliNeurIPS 2025 · 被引用 12 次
- Self-Supervised Contrastive Learning is Approximately Supervised Contrastive LearningAchleshwar Luthra, Tianbao Yang, Tomer GalantiNeurIPS 2025 · 被引用 7 次
- Neural Collapse in Cumulative Link Models for Ordinal Regression: An Analysis with Unconstrained Feature ModelChuang Ma, Tomoyuki Obuchi, Toshiyuki TanakaNeurIPS 2025 · 被引用 6 次
- Hephaestus: Mixture Generative Modeling with Energy Guidance for Large-scale QoS DegradationNguyen Do, Bach Ngo, Youval Kashuv, Canh V. Pham 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper14
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- A Geometric Analysis of Neural Collapse with Unconstrained FeaturesZhihui Zhu, Tianyu Ding, Jinxin Zhou, Xiao Li 等NeurIPS 2021 · 被引用 303 次
- Exploring Balanced Feature Spaces for Representation LearningBingyi Kang, Yu Li, Sa Xie, Zehuan Yuan 等ICLR 2021 · 被引用 296 次
- Neural Collapse Under MSE Loss: Proximity to and Dynamics on the Central PathX. Y. Han, Vardan Papyan, David L. DonohoICLR 2022 · 被引用 182 次
- Inducing Neural Collapse in Imbalanced Learning: Do We Really Need a Learnable Classifier at the End of Deep Neural Network?Yibo Yang, Shixiang Chen, Xiangtai Li, Liang Xie 等NeurIPS 2022 · 被引用 144 次
相关 Paper
- Neural Collapse in Deep Linear Networks: From Balanced to Imbalanced DataHien Dang, Tho Tran Huu, Stanley J. Osher, Hung Tran-The 等ICML 2023 · 被引用 44 次
- Extended Unconstrained Features Model for Exploring Deep Neural CollapseTom Tirer, Joan BrunaICML 2022 · 被引用 118 次
- On the Optimization Landscape of Neural Collapse under MSE Loss: Global Optimality with Unconstrained FeaturesJinxin Zhou, Xiao Li, Tianyu Ding, Chong You 等ICML 2022 · 被引用 122 次
- Inducing Neural Collapse to a Fixed Hierarchy-Aware Frame for Reducing Mistake SeverityTong Liang, Jim DavisICCV 2023 · 被引用 14 次
- MLC-NC: Long-Tailed Multi-Label Image Classification Through the Lens of Neural CollapseZijian Tao, Shao-Yuan Li, Wenhai Wan, Jinpeng Zheng 等AAAI 2025 · 被引用 7 次
