Learning Towards The Largest Margins
Xiong Zhou, Xianming Liu, Deming Zhai, Junjun Jiang, Xin Gao, Xiangyang Ji
Abstract
One of the main challenges for feature representation in deep learning-based classification is the design of appropriate loss functions that exhibit strong discriminative power. The classical softmax loss does not explicitly encourage discriminative learning of features. A popular direction of research is to incorporate margins in well-established losses in order to enforce extra intra-class compactness and inter-class separability, which, however, were developed through heuristic means, as opposed to rigorous mathematical principles. In this work, we attempt to address this limitation by formulating the principled optimization objective as learning towards the largest margins. Specifically, we firstly define the class margin as the measure of inter-class separability, and the sample margin as the measure of intra-class compactness. Accordingly, to encourage discriminative representation of features, the loss function should promote the largest possible margins for both classes and samples. Furthermore, we derive a generalized margin softmax loss to draw general conclusions for the existing margin-based losses. Not only does this principled framework offer new perspectives to understand and interpret existing margin-based losses, but it also provides new insights that can guide the design of new tools, including sample margin regularization and largest margin softmax loss for the class-balanced case, and zero-centroid regularization for the class-imbalanced case. Experimental results demonstrate the effectiveness of our strategy on a variety of tasks, including visual classification, imbalanced classification, person re-identification, and face verification.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c14be5f3-1b41-44de-9866-5c7ca78c9fedCited by top-tier papers7
- Generalized Neural Collapse for a Large Number of ClassesJiachen Jiang, Jinxin Zhou, Peng Wang, Qing Qu et al.ICML 2024 · 44 citations
- Maximum Class Separation as Inductive Bias in One MatrixTejaswi Kasarla, Gertjan J. Burghouts, Max van Spengler, Elise van der Pol et al.NeurIPS 2022 · 29 citations
- Quantifying the Variability Collapse of Neural NetworksJing Xu, Haoxiong LiuICML 2023 · 10 citations
- Prototype-Anchored Learning for Learning with Imperfect AnnotationsXiong Zhou, Xianming Liu, Deming Zhai, Junjun Jiang et al.ICML 2022 · 8 citations
- Zero-Mean Regularized Spectral Contrastive Learning: Implicitly Mitigating Wrong Connections in Positive-Pair GraphsXiong Zhou, Xianming Liu, Feilong Zhang, Gang Wu et al.ICLR 2024 · 4 citations
Builds on6
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 2,360 citations
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- High-Performance Large-Scale Image Recognition Without NormalizationAndy Brock, Soham De, Samuel L. Smith, Karen SimonyanICML 2021 · 613 citations
Related papers
- Mis-Classified Vector Guided Softmax Loss for Face RecognitionXiaobo Wang, Shifeng Zhang, Shuo Wang, Tianyu Fu et al.AAAI 2020 · 188 citations
- Gaussian Affinity for Max-Margin Class Imbalanced LearningMunawar Hayat, Salman H. Khan, Syed Waqas Zamir, Jianbing Shen et al.ICCV 2019 · 71 citations
- Reducing Class-Wise Performance Disparity via Margin RegularizationBeier Zhu, Kesen Zhao, Jiequan Cui, Qianru Sun et al.ICLR 2026 · 1 citation
- Multi-Class Support Vector Machine with Maximizing Minimum MarginFeiping Nie, Zhezheng Hao, Rong WangAAAI 2024 · 30 citations
- MaxSup: Overcoming Representation Collapse in Label SmoothingYuxuan Zhou, Heng Li, Zhi-Qi Cheng, Xudong Yan et al.NeurIPS 2025 · 5 citations
