Orthogonal Projection Loss
Kanchana Ranasinghe, Muzammal Naseer, Munawar Hayat, Salman H. Khan, Fahad Shahbaz Khan
摘要
Deep neural networks have achieved remarkable performance on a range of classification tasks, with softmax crossentropy (CE) loss emerging as the de-facto objective function. The CE loss encourages features of a class to have a higher projection score on the true class-vector compared to the negative classes. However, this is a relative constraint and does not explicitly force different class features to be well-separated. Motivated by the observation that ground-truth class representations in CE loss are orthogonal (one-hot encoded vectors), we develop a novel loss function termed ‘Orthogonal Projection Loss' (OPL) which imposes orthogonality in the feature space. OPL augments the properties of CE loss and directly enforces inter-class separation alongside intra-class clustering in the feature space through orthogonality constraints on the mini-batch level. As compared to other alternatives of CE, OPL offers unique advantages e.g., no additional learnable parameters, does not require careful negative mining and is not sensitive to the batch size. Given the plug-and-play nature of OPL, we evaluate it on a diverse range of tasks including image recognition (CIFAR-100), large-scale classification (ImageNet), domain generalization (PACS) and few-shot learning (mini-ImageNet, CIFAR-FS, tiered-ImageNet and Meta-dataset) and demonstrate its effectiveness across the board. Furthermore, OPL offers better robustness against practical nuisances such as adversarial attacks and label noise. Code is available at: https://github.com/kahnchana/opl.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Object-Aware Domain Generalization for Object DetectionWooju Lee, Dasol Hong, Hyungtae Lim, Hyun MyungAAAI 2024 · 被引用 58 次
- Auxiliary Losses for Learning Generalizable Concept-based ModelsIvaxi Sheth, Samira Ebrahimi KahouNeurIPS 2023 · 被引用 52 次
- Concept Bottleneck Generative ModelsAya Abdelsalam Ismail, Julius Adebayo, Héctor Corrada Bravo, Stephen Ra 等ICLR 2024 · 被引用 43 次
- The Principle of Diversity: Training Stronger Vision Transformers Calls for Reducing All Levels of RedundancyTianlong Chen, Zhenyu Zhang, Yu Cheng, Ahmed Awadallah 等CVPR 2022 · 被引用 36 次
- Maximum Class Separation as Inductive Bias in One MatrixTejaswi Kasarla, Gertjan J. Burghouts, Max van Spengler, Elise van der Pol 等NeurIPS 2022 · 被引用 29 次
它引用的顶会 Paper11
- Reliable evaluation of adversarial robustness with an ensemble of diverse parameter-free attacksFrancesco Croce, Matthias HeinICML 2020 · 被引用 2,337 次
- Improving Adversarial Robustness Requires Revisiting Misclassified ExamplesYisen Wang, Difan Zou, Jinfeng Yi, James Bailey 等ICLR 2020 · 被引用 829 次
- Meta-Dataset: A Dataset of Datasets for Learning to Learn from Few ExamplesEleni Triantafillou, Tyler Zhu, Vincent Dumoulin, Pascal Lamblin 等ICLR 2020 · 被引用 692 次
- Few-Shot Learning With Embedded Class Models and Shot-Free Meta TrainingAvinash Ravichandran, Rahul Bhotika, Stefano SoattoICCV 2019 · 被引用 191 次
- Unraveling Meta-Learning: Understanding Feature Representations for Few-Shot TasksMicah Goldblum, Steven Reich, Liam Fowl, Renkun Ni 等ICML 2020 · 被引用 82 次
相关 Paper
- Gaussian Affinity for Max-Margin Class Imbalanced LearningMunawar Hayat, Salman H. Khan, Syed Waqas Zamir, Jianbing Shen 等ICCV 2019 · 被引用 71 次
- BCE vs. CE in Deep Feature LearningQiufu Li, Huibin Xiao, Linlin ShenICML 2025
- The Devil is in the Margin: Margin-based Label Smoothing for Network CalibrationBingyuan Liu, Ismail Ben Ayed, Adrian Galdran, Jose DolzCVPR 2022 · 被引用 61 次
- Beyond Max-Margin: Class Margin Equilibrium for Few-Shot Object DetectionBohao Li, Boyu Yang, Chang Liu, Feng Liu 等CVPR 2021
- UniFace: Unified Cross-Entropy Loss for Deep Face RecognitionJiancan Zhou, Xi Jia, Qiufu Li, Linlin Shen 等ICCV 2023 · 被引用 38 次
