Fixed Non-negative Orthogonal Classifier: Inducing Zero-mean Neural Collapse with Feature Dimension Separation
Hoyong Kim, Kangil Kim
摘要
Fixed classifiers in neural networks for classification problems have demonstrated cost efficiency and even outperformed learnable classifiers in some popular benchmarks when incorporating orthogonality (Pernici et al., 2021a). Despite these advantages, prior research has yet to investigate the training dynamics of fixed classifiers on neural collapse. Ensuring this phenomenon is critical for obtaining global optimality in a layer-peeled model, potentially leading to enhanced performance in practice. However, the neural collapse cannot explain the collapse phenomenon in the fixed classifier when its shape is not a simplex ETF. To overcome the limits, we exploit additional constraints to the layer-peeled model: non-negativity and orthogonality. Then, we propose a fixed non-negative orthogonal classifier, which makes a layer-peeled model with the fixed classifier have the global optimality and the max-margin in decision by inducing zero-mean neural collapse. Building on this foundation, we exploit a feature dimension separation inherent in our classifier for further purposes: (1) enhances softmax masking by mitigating feature interference in continual learning and (2) tackles the limitations of mixup on the hypersphere in imbalanced learning. We conducted comprehensive experiments on various datasets and demonstrated significant performance improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- The Prevalence of Neural Collapse in Neural Multivariate RegressionGeorge Andriopoulos, Zixuan Dong, Li Guo, Zifan Zhao 等NeurIPS 2024 · 被引用 24 次
- Neural Collapse Inspired Knowledge DistillationShuoxi Zhang, Zijian Song, Kun HeAAAI 2025 · 被引用 2 次
- Neural Collapse by Design: Learning Class Prototypes on the HyperspherePanagiotis Koromilas, Theodoros Giannakopoulos, Mihalis Nicolaou, Yannis PanagakisICML 2026
它引用的顶会 Paper46
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Neural Sheaf Diffusion: A Topological Perspective on Heterophily and Oversmoothing in GNNsCristian Bodnar, Francesco Di Giovanni, Benjamin Paul Chamberlain, Pietro Lió 等NeurIPS 2022 · 被引用 313 次
- A Geometric Analysis of Neural Collapse with Unconstrained FeaturesZhihui Zhu, Tianyu Ding, Jinxin Zhou, Xiao Li 等NeurIPS 2021 · 被引用 303 次
- Using Hindsight to Anchor Past Knowledge in Continual LearningArslan Chaudhry, Albert Gordo, Puneet K. Dokania, Philip H. S. Torr 等AAAI 2021 · 被引用 279 次
相关 Paper
- Inducing Neural Collapse in Imbalanced Learning: Do We Really Need a Learnable Classifier at the End of Deep Neural Network?Yibo Yang, Shixiang Chen, Xiangtai Li, Liang Xie 等NeurIPS 2022 · 被引用 144 次
- An Unconstrained Layer-Peeled Perspective on Neural CollapseWenlong Ji, Yiping Lu, Yiliang Zhang, Zhun Deng 等ICLR 2022 · 被引用 101 次
- Neural Collapse Inspired Debiased Representation Learning for Min-max FairnessShenyu Lu, Junyi Chai, Xiaoqian WangKDD 2024 · 被引用 1 次
- Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Features ModelHien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat HoICML 2024 · 被引用 19 次
- Imbalance Trouble: Revisiting Neural-Collapse GeometryChristos Thrampoulidis, Ganesh Ramachandra Kini, Vala Vakilian, Tina BehniaNeurIPS 2022 · 被引用 101 次
