Maximum Class Separation as Inductive Bias in One Matrix
Tejaswi Kasarla, Gertjan J. Burghouts, Max van Spengler, Elise van der Pol, Rita Cucchiara, Pascal Mettes
摘要
Maximizing the separation between classes constitutes a well-known inductive bias in machine learning and a pillar of many traditional algorithms. By default, deep networks are not equipped with this inductive bias and therefore many alternative solutions have been proposed through differential optimization. Current approaches tend to optimize classification and separation jointly: aligning inputs with class vectors and separating class vectors angularly. This paper proposes a simple alternative: encoding maximum separation as an inductive bias in the network by adding one fixed matrix multiplication before computing the softmax activations. The main observation behind our approach is that separation does not require optimization but can be solved in closed-form prior to training and plugged into a network. We outline a recursive approach to obtain the matrix consisting of maximally separable vectors for any number of classes, which can be added with negligible engineering effort and computational overhead. Despite its simple nature, this one matrix multiplication provides real impact. We show that our proposal directly boosts classification, long-tailed recognition, out-of-distribution detection, and open-set recognition, from CIFAR to ImageNet. We find empirically that maximum separation works best as a fixed bias; making the matrix learnable adds nothing to the performance. The closed-form implementation and code to reproduce the experiments are available on github.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Combating Representation Learning Disparity with Geometric HarmonizationZhihan Zhou, Jiangchao Yao, Feng Hong, Ya Zhang 等NeurIPS 2023 · 被引用 20 次
- Improved Balanced Classification with Theoretically Grounded Loss FunctionsCorinna Cortes, Mehryar Mohri, Yutao ZhongNeurIPS 2025 · 被引用 19 次
- Forgetting, Ignorance or Myopia: Revisiting Key Challenges in Online Continual LearningXinrui Wang, Chuanxing Geng, Wenhai Wan, Shao-Yuan Li 等NeurIPS 2024 · 被引用 16 次
- Hyperspherical Classification with Dynamic Label-to-Prototype AssignmentMohammad Saeed Ebrahimi Saadabadi, Ali Dabouei, Sahar Rahimi Malakshan, Nasser M. NasrabadiCVPR 2024 · 被引用 9 次
- Pairwise Similarity Learning is SimPLEYandong Wen, Weiyang Liu, Yao Feng, Bhiksha Raj 等ICCV 2023 · 被引用 9 次
它引用的顶会 Paper21
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- Directional Message Passing for Molecular GraphsJohannes Klicpera, Janek Groß, Stephan GünnemannICLR 2020 · 被引用 1,079 次
- Open-Set Recognition: A Good Closed-Set Classifier is All You NeedSagar Vaze, Kai Han, Andrea Vedaldi, Andrew ZissermanICLR 2022 · 被引用 594 次
- ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive BiasYufei Xu, Qiming Zhang, Jing Zhang, Dacheng TaoNeurIPS 2021 · 被引用 429 次
- Graph Convolutional Reinforcement LearningJiechuan Jiang, Chen Dun, Tiejun Huang, Zongqing LuICLR 2020 · 被引用 415 次
相关 Paper
- Open-Sampling: Exploring Out-of-Distribution data for Re-balancing Long-tailed datasetsHongxin Wei, Lue Tao, Renchunzi Xie, Lei Feng 等ICML 2022 · 被引用 46 次
- Unlocking Better Closed-Set Alignment Based on Neural Collapse for Open-Set RecognitionChaohua Li, Enhao Zhang, Chuanxing Geng, Songcan ChenAAAI 2025 · 被引用 1 次
- Pursuing Feature Separation based on Neural Collapse for Out-of-Distribution DetectionYingwen Wu, Ruiji Yu, Xinwen Cheng, Zhengbao He 等ICLR 2025
- Few-Shot Open-Set Recognition Using Meta-LearningBo Liu, Hao Kang, Haoxiang Li, Gang Hua 等CVPR 2020
- Multi-Class Data Description for Out-of-distribution DetectionDongha Lee, Sehun Yu, Hwanjo YuKDD 2020 · 被引用 22 次
