Self-Supervised Aggregation of Diverse Experts for Test-Agnostic Long-Tailed Recognition
Yifan Zhang, Bryan Hooi, Lanqing Hong, Jiashi Feng
Abstract
Existing long-tailed recognition methods, aiming to train class-balanced models from long-tailed data, generally assume the models would be evaluated on the uniform test class distribution. However, practical test class distributions often violate this assumption (e.g., being either long-tailed or even inversely long-tailed), which may lead existing methods to fail in real applications. In this paper, we study a more practical yet challenging task, called test-agnostic long-tailed recognition, where the training class distribution is long-tailed while the test class distribution is agnostic and not necessarily uniform. In addition to the issue of class imbalance, this task poses another challenge: the class distribution shift between the training and test data is unknown. To tackle this task, we propose a novel approach, called Self-supervised Aggregation of Diverse Experts, which consists of two strategies: (i) a new skill-diverse expert learning strategy that trains multiple experts from a single and stationary long-tailed dataset to separately handle different class distributions; (ii) a novel test-time expert aggregation strategy that leverages self-supervision to aggregate the learned multiple experts for handling unknown test class distributions. We theoretically show that our self-supervised strategy has a provable ability to simulate test-agnostic class distributions. Promising empirical results demonstrate the effectiveness of our method on both vanilla and test-agnostic long-tailed recognition. Code is available at https://github.com/Vanint/SADE-AgnosticLT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers58
- Expanding Small-Scale Datasets with Guided ImaginationYifan Zhang, Daquan Zhou, Bryan Hooi, Kai Wang et al.NeurIPS 2023 · 84 citations
- Efficient Test-Time Adaptation for Super-Resolution with Second-Order Degradation and ReconstructionZeshuai Deng, Zhuokun Chen, Shuaicheng Niu, Thomas H. Li et al.NeurIPS 2023 · 37 citations
- MDCS: More Diverse Experts with Consistency Self-distillation for Long-tailed RecognitionQihao Zhao, Chen Jiang, Wei Hu, Fan Zhang et al.ICCV 2023 · 35 citations
- Local and Global Logit Adjustments for Long-Tailed LearningYingfan Tao, Jingna Sun, Hao Yang, Li Chen et al.ICCV 2023 · 32 citations
- Decoupled Contrastive Learning for Long-Tailed RecognitionShiyu Xuan, Shiliang ZhangAAAI 2024 · 29 citations
Builds on39
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen et al.ICLR 2021 · 1,731 citations
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain et al.ICLR 2021 · 937 citations
Related papers
- A Square Peg in a Square Hole: Meta-Expert for Long-Tailed Semi-Supervised LearningYaxin Hou, Yuheng JiaICML 2025
- Exploring Balanced Feature Spaces for Representation LearningBingyi Kang, Yu Li, Sa Xie, Zehuan Yuan et al.ICLR 2021 · 296 citations
- Three Heads Are Better than One: Complementary Experts for Long-Tailed Semi-supervised LearningChengcheng Ma, Ismail Elezi, Jiankang Deng, Weiming Dong et al.AAAI 2024 · 20 citations
- Distilling Long-tailed DatasetsZhenghao Zhao, Haoxuan Wang, Yuzhang Shang, Kai Wang et al.CVPR 2025
- Breaking Long-Tailed Learning Bottlenecks: A Controllable Paradigm with Hypernetwork-Generated Diverse ExpertsZhe Zhao, Haibin Wen, Zikang Wang, Pengkun Wang et al.NeurIPS 2024 · 12 citations
