T-distributed Spherical Feature Representation for Imbalanced Classification
Xiaoyu Yang, Yufei Chen, Xiaodong Yue, Shaoxun Xu, Chao Ma
Abstract
Real-world classification tasks often show an extremely imbalanced problem. The extreme imbalance will cause a strong bias that the decision boundary of the classifier is completely dominated by the categories with abundant samples, which are also called the head categories. Current methods have alleviated the imbalanced impact from mainly three aspects: class re-balance, decoupling and domain adaptation. However, the existing criterion with the winner-take-all strategy still leads to the crowding problem in the eigenspace. The head categories with many samples can extract features more accurately, but occupy most of the eigenspace. The tail categories sharing the rest of the narrow eigenspace are too crowded together to accurately extract features. Above these issues, we propose a novel T-distributed spherical metric for equalized eigenspace in the imbalanced classification, which has the following innovations: 1) We design the T-distributed spherical metric, which has the characteristics of high kurtosis. Instead of the winner-take-all strategy, the T-distributed spherical metric produces a high logit only when the extracted feature is close enough to the category center, without a strong bias against other categories. 2) The T-distributed spherical metric is integrated into the classifier, which is able to equalize the eigenspace for alleviating the crowding issue in the imbalanced problem. The equalized eigenspace by the T-distributed spherical classifier is capable of improving the accuracy of the tail categories while maintaining the accuracy of the head, which significantly promotes the intraclass compactness and interclass separability of features. Extensive experiments on large-scale imbalanced datasets verify our method, which shows superior results in the long-tailed CIFAR-100/-10 with the imbalanced ratio IR = 100/50.
Our method also achieves excellent results on the large-scale ImageNet-LT dataset and the iNaturalist dataset with various backbones. In addition, we provide a case study of the real clinical classification of pancreatic tumor subtypes with 6 categories. Among them, the largest number of PDAC accounts for 315 cases, and the least CP has only 8 cases. After 4-fold cross-validation, we achieved a top-1 accuracy of 69.04%.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7888cda2-8e02-4243-b733-d2d9eb7cfc83Cited by top-tier papers6
- Walking the Tightrope: Autonomous Disentangling Beneficial and Detrimental Drifts in Non-Stationary Custom-TuningXiaoyu Yang, Jie Lu, En YuNeurIPS 2025 · 22 citations
- From Newborn to Impact: Bias-Aware Citation PredictionMingfei Lu, Mengjia Wu, Jiawei Xu, Weikai Li et al.WWW 2026 · 6 citations
- Stop Guessing: Choosing the Optimization-Consistent Uncertainty Measurement for Evidential Deep LearningLinye Li, Yufei Chen, Xiaodong Yue, Xujing Zhou et al.ICLR 2026
- One Leaf Reveals the Season: Occlusion-Based Contrastive Learning with Semantic-Aware Views for Efficient Visual RepresentationXiaoyu Yang, Lijian Xu, Hongsheng Li, Shaoting ZhangICML 2025
- Adapting Multi-modal Large Language Model to Concept Drift From Pre-training OnwardsXiaoyu Yang, Jie Lu, En YuICLR 2025
Builds on12
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Long-Tailed Classification by Keeping the Good and Removing the Bad Momentum Causal EffectKaihua Tang, Jianqiang Huang, Hanwang ZhangNeurIPS 2020 · 533 citations
- Long-tailed Recognition by Routing Diverse Distribution-Aware ExpertsXudong Wang, Long Lian, Zhongqi Miao, Ziwei Liu et al.ICLR 2021 · 481 citations
- Self-Adaptive Training: beyond Empirical Risk MinimizationLang Huang, Chao Zhang, Hongyang ZhangNeurIPS 2020 · 256 citations
- ACE: Ally Complementary Experts for Solving Long-Tailed Recognition in One-ShotJiarui Cai, Yizhou Wang, Jenq-Neng HwangICCV 2021 · 176 citations
Related papers
- Geometry of Long-Tailed Representation Learning: Rebalancing Features for Skewed DistributionsLingjie Yi, Jiachen Yao, Weimin Lyu, Haibin Ling et al.ICLR 2025
- Inverse Weight-Balancing for Deep Long-Tailed LearningWenqi Dang, Zhou Yang, Weisheng Dong, Xin Li et al.AAAI 2024 · 8 citations
- Equalization Loss v2: A New Gradient Balance Approach for Long-Tailed Object DetectionJingru Tan, Xin Lu, Gang Zhang, Changqing Yin et al.CVPR 2021
- Distributional Robustness Loss for Long-tail LearningDvir Samuel, Gal ChechikICCV 2021 · 128 citations
- Feature Fusion from Head to Tail for Long-Tailed Visual RecognitionMengke Li, Zhikai Hu, Yang Lu, Weichao Lan et al.AAAI 2024 · 59 citations
