Stochastic smoothing of the top-K calibrated hinge loss for deep imbalanced classification
Camille Garcin, Maximilien Servajean, Alexis Joly, Joseph Salmon
Abstract
In modern classification tasks, the number of labels is getting larger and larger, as is the size of the datasets encountered in practice. As the number of classes increases, class ambiguity and class imbalance become more and more problematic to achieve high top-1 accuracy. Meanwhile, Top-K metrics (metrics allowing K guesses) have become popular, especially for performance reporting. Yet, proposing top-K losses tailored for deep learning remains a challenge, both theoretically and practically. In this paper we introduce a stochastic top-K hinge loss inspired by recent developments on top-K calibrated losses. Our proposal is based on the smoothing of the top-K operator building on the flexible "perturbed optimizer" framework. We show that our loss function performs very well in the case of balanced datasets, while benefiting from a significantly lower computational time than the state-of-the-art top-K loss function. In addition, we propose a simple variant of our loss for the imbalanced case. Experiments on a heavy-tailed dataset show that our loss function significantly outperforms other baseline loss functions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a371a8a1-7471-48e0-bb81-93508aba3dbdCited by top-tier papers6
- Bilevel Coreset Selection in Continual Learning: A New Formulation and AlgorithmJie Hao, Kaiyi Ji, Mingrui LiuNeurIPS 2023 · 43 citations
- Efficient Augmentation for Imbalanced Deep LearningDamien A. Dablain, Colin Bellinger, Bartosz Krawczyk, Nitesh V. ChawlaICDE 2023 · 19 citations
- Conformal Prediction for Long-Tailed ClassificationTiffany Ding, Jean-Baptiste Fermanian, Joseph SalmonICLR 2026 · 9 citations
- First-Order Minimax Bilevel OptimizationYifan Yang, Zhaofeng Si, Siwei Lyu, Kaiyi JiNeurIPS 2024 · 3 citations
- SoftMoE: Soft Differentiable Routing for Mixture-of-Experts in LLMsMikołaj Zasada, Łukasz Struski, Jacek Tabor, Marcin KurdzielICML 2026 · 2 citations
Builds on4
- Long-tailed Recognition by Routing Diverse Distribution-Aware ExpertsXudong Wang, Long Lian, Zhongqi Miao, Ziwei Liu et al.ICLR 2021 · 481 citations
- Differentiable Top-k with Optimal TransportYujia Xie, Hanjun Dai, Minshuo Chen, Bo Dai et al.NeurIPS 2020 · 124 citations
- On the consistency of top-k surrogate lossesForest Yang, Sanmi KoyejoICML 2020 · 54 citations
- BBN: Bilateral-Branch Network With Cumulative Learning for Long-Tailed Visual RecognitionBoyan Zhou, Quan Cui, Xiu-Shen Wei, Zhao-Min ChenCVPR 2020
Related papers
- ST: A Scalable Module for Solving Top-k ProblemsHanchen Xia, Weidong Liu, Xiaojun MaoNeurIPS 2024 · 1 citation
- Differentiable Top-k Classification LearningFelix Petersen, Hilde Kuehne, Christian Borgelt, Oliver DeussenICML 2022 · 48 citations
- Hierarchical classification at multiple operating pointsJack ValmadreNeurIPS 2022 · 27 citations
- Cardinality-Aware Set Prediction and Top- ClassificationCorinna Cortes, Anqi Mao, Christopher Mohri, Mehryar Mohri et al.NeurIPS 2024 · 29 citations
- Gaussian Affinity for Max-Margin Class Imbalanced LearningMunawar Hayat, Salman H. Khan, Syed Waqas Zamir, Jianbing Shen et al.ICCV 2019 · 71 citations
