Harnessing Hierarchical Label Distribution Variations in Test Agnostic Long-tail Recognition
Zhiyong Yang, Qianqian Xu, Zitai Wang, Sicong Li, Boyu Han, Shilong Bao, Xiaochun Cao, Qingming Huang
摘要
This paper explores test-agnostic long-tail recognition, a challenging long-tail task where the test label distributions are unknown and arbitrarily imbalanced. We argue that the variation in these distributions can be broken down hierarchically into global and local levels. The global ones reflect a broad range of diversity, while the local ones typically arise from milder changes, often focused on a particular neighbor. Traditional methods predominantly use a Mixture-of-Expert (MoE) approach, targeting a few fixed test label distributions that exhibit substantial global variations. However, the local variations are left unconsidered. To address this issue, we propose a new MoE strategy, <inline-formula><tex-math notation="LaTeX"></tex-math><alternatives>mml:math<mml:mi mathvariant="sans-serif">DirMixE</mml:mi></mml:math><inline-graphic xlink:href="yang-ieq1-3647124.gif"/></alternatives></inline-formula>, which assigns experts to different Dirichlet meta-distributions of the label distribution, each targeting a specific aspect of local variations. Additionally, the diversity among these Dirichlet meta-distributions inherently captures global variations. This dual-level approach also leads to a more stable objective function, allowing us to sample different test distributions better to quantify the mean and variance of performance outcomes. Building on this idea, we develop a general Latent Skill Finetuning (LSF) framework for parameter-efficient finetuning of foundation models. We provide implementations based on LoRA and Adapter. Theoretically, we derive upper bounds on the generalization error for both standard learning and PEFT. Under mild assumptions, we show that the variance-based regularization helps tighten these bounds. Furthermore, we prove that the covering number of the PEFT hypothesis class scales with the number of trainable parameters. Finally, extensive experiments on CIFAR-10-LT, CIFAR-100-LT, ImageNet-LT, and iNaturalist validate the effectiveness of <inline-formula><tex-math notation="LaTeX"></tex-math><alternatives>mml:math<mml:mi mathvariant="sans-serif">DirMixE</mml:mi></mml:math><inline-graphic xlink:href="yang-ieq2-3647124.gif"/></alternatives></inline-formula>.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Improved Balanced Classification with Theoretically Grounded Loss FunctionsCorinna Cortes, Mehryar Mohri, Yutao ZhongNeurIPS 2025 · 被引用 19 次
- DiffuLT: Diffusion for Long-tail Recognition Without External KnowledgeJie Shao, Ke Zhu, Hanxiao Zhang, Jianxin WuNeurIPS 2024 · 被引用 17 次
- A Unified Generalization Analysis of Re-Weighting and Logit-Adjustment for Imbalanced LearningZitai Wang, Qianqian Xu, Zhiyong Yang, Yuan He 等NeurIPS 2023 · 被引用 15 次
- Neural Collapse To Multiple Centers For Imbalanced DataHongren Yan, Yuhua Qian, Furong Peng, Jiachen Luo 等NeurIPS 2024 · 被引用 13 次
- Optimized Deferral for Imbalanced SettingsCorinna Cortes, Anqi Mao, Mehryar Mohri, Yutao ZhongICML 2026 · 被引用 7 次
它引用的顶会 Paper29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- AdaptFormer: Adapting Vision Transformers for Scalable Visual RecognitionShoufa Chen, Chongjian Ge, Zhan Tong, Jiangliu Wang 等NeurIPS 2022 · 被引用 1,291 次
相关 Paper
- MoA: Heterogeneous Mixture of Adapters for Parameter-Efficient Fine-Tuning of Large Language ModelsJie Cao, Tianwei Lin, Bo Yuan, Rolan Yan 等ACL 2026 · 被引用 2 次
- Parameter-Efficient Complementary Expert Learning for Long-Tailed Visual RecognitionLixiang Ru, Xin Guo, Lei Yu, Yingying Zhang 等ACM MM 2024 · 被引用 2 次
- Self-Supervised Aggregation of Diverse Experts for Test-Agnostic Long-Tailed RecognitionYifan Zhang, Bryan Hooi, Lanqing Hong, Jiashi FengNeurIPS 2022 · 被引用 214 次
- LT-Soups: Bridging Head and Tail Classes via Subsampled Model SoupsMasih Aminbeidokhti, Subhankar Roy, Eric Granger, Elisa Ricci 等NeurIPS 2025 · 被引用 1 次
- HiMoLE: Towards OOD-Robust LoRA via Hierarchical Mixture of ExpertsYinuo Jiang, Xiaodong Yan, Keyan Ding, Deng Zhao 等NeurIPS 2025
