Inducing Neural Collapse in Imbalanced Learning: Do We Really Need a Learnable Classifier at the End of Deep Neural Network?
Yibo Yang, Shixiang Chen, Xiangtai Li, Liang Xie, Zhouchen Lin, Dacheng Tao
Abstract
Modern deep neural networks for classification usually jointly learn a backbone for representation and a linear classifier to output the logit of each class. A recent study has shown a phenomenon called neural collapse that the within-class means of features and the classifier vectors converge to the vertices of a simplex equiangular tight frame (ETF) at the terminal phase of training on a balanced dataset. Since the ETF geometric structure maximally separates the pair-wise angles of all classes in the classifier, it is natural to raise the question, why do we spend an effort to learn a classifier when we know its optimal geometric structure? In this paper, we study the potential of learning a neural network for classification with the classifier randomly initialized as an ETF and fixed during training. Our analytical work based on the layer-peeled model indicates that the feature learning with a fixed ETF classifier naturally leads to the neural collapse state even when the dataset is imbalanced among classes. We further show that in this case the cross entropy (CE) loss is not necessary and can be replaced by a simple squared loss that shares the same global optimality but enjoys a better convergence property. Our experimental results show that our method is able to bring significant improvements with faster convergence on multiple imbalanced datasets. Code address: link.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers44
- Long-Tail Learning with Foundation Model: Heavy Fine-Tuning HurtsJiang-Xin Shi, Tong Wei, Zhi Zhou, Jie-Jing Shao et al.ICML 2024 · 78 citations
- Linguistic Collapse: Neural Collapse in (Large) Language ModelsRobert Wu, Vardan PapyanNeurIPS 2024 · 45 citations
- Neural Collapse in Deep Linear Networks: From Balanced to Imbalanced DataHien Dang, Tho Tran Huu, Stanley J. Osher, Hung Tran-The et al.ICML 2023 · 44 citations
- NECO: NEural Collapse Based Out-of-distribution detectionMouïn Ben Ammar, Nacim Belkhir, Sebastian Popescu, Antoine Manzanera et al.ICLR 2024 · 40 citations
- Federated Learning with Bilateral Curation for Partially Class-Disjoint DataZiqing Fan, Ruipeng Zhang, Jiangchao Yao, Bo Han et al.NeurIPS 2023 · 28 citations
Builds on12
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- A Geometric Analysis of Neural Collapse with Unconstrained FeaturesZhihui Zhu, Tianyu Ding, Jinxin Zhou, Xiao Li et al.NeurIPS 2021 · 303 citations
- Neural Collapse Under MSE Loss: Proximity to and Dynamics on the Central PathX. Y. Han, Vardan Papyan, David L. DonohoICLR 2022 · 182 citations
- On the Optimization Landscape of Neural Collapse under MSE Loss: Global Optimality with Unconstrained FeaturesJinxin Zhou, Xiao Li, Tianyu Ding, Chong You et al.ICML 2022 · 122 citations
- Extended Unconstrained Features Model for Exploring Deep Neural CollapseTom Tirer, Joan BrunaICML 2022 · 118 citations
Related papers
- Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Features ModelHien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat HoICML 2024 · 19 citations
- Inducing Neural Collapse to a Fixed Hierarchy-Aware Frame for Reducing Mistake SeverityTong Liang, Jim DavisICCV 2023 · 14 citations
- Guiding Neural Collapse: Optimising Towards the Nearest Simplex Equiangular Tight FrameEvan Markou, Thalaiyasingam Ajanthan, Stephen GouldNeurIPS 2024 · 16 citations
- Neural Collapse Inspired Feature-Classifier Alignment for Few-Shot Class-Incremental LearningYibo Yang, Haobo Yuan, Xiangtai Li, Zhouchen Lin et al.ICLR 2023 · 22 citations
- MLC-NC: Long-Tailed Multi-Label Image Classification Through the Lens of Neural CollapseZijian Tao, Shao-Yuan Li, Wenhai Wan, Jinpeng Zheng et al.AAAI 2025 · 7 citations
