Focal-SAM: Focal Sharpness-Aware Minimization for Long-Tailed Classification
Sicong Li, Qianqian Xu, Zhiyong Yang, Zitai Wang, Linchao Zhang, Xiaochun Cao, Qingming Huang
Abstract
Real-world datasets often follow a long-tailed distribution, making generalization to tail classes difficult. Recent methods resorted to long-tail variants of Sharpness-Aware Minimization (SAM), such as ImbSAM and CC-SAM, to improve generalization by flattening the loss landscape. However, these attempts face a trade-off between computational efficiency and control over the loss landscape. On the one hand, ImbSAM is efficient but offers only coarse control as it excludes head classes from the SAM process. On the other hand, CC-SAM provides fine-grained control through class-dependent perturbations but at the cost of efficiency due to multiple backpropagations. Seeing this dilemma, we introduce Focal-SAM, which assigns different penalties to class-wise sharpness, achieving fine-grained control without extra backpropagations, thus maintaining efficiency. Furthermore, we theoretically analyze Focal-SAM's generalization ability and derive a sharper generalization bound. Extensive experiments on both traditional and foundation models validate the effectiveness of Focal-SAM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 01dec756-a09d-4245-85ff-43a36b303315Cited by top-tier papers15
- Harnessing Hierarchical Label Distribution Variations in Test Agnostic Long-tail RecognitionZhiyong Yang, Qianqian Xu, Zitai Wang, Sicong Li et al.ICML 2024 · 21 citations
- Improved Balanced Classification with Theoretically Grounded Loss FunctionsCorinna Cortes, Mehryar Mohri, Yutao ZhongNeurIPS 2025 · 19 citations
- Confusion-Aware Spectral Regularizer for Long-Tailed RecognitionZiquan Zhu, Gaojie Jin, Hanruo Zhu, Si-Yuan Lu et al.CVPR 2026 · 4 citations
- Reframing Long-Tailed Learning via Loss Landscape GeometryShenghan Chen, Yiming Liu, Yanzhen Wang, Yujia Wang et al.CVPR 2026 · 2 citations
- Making Training-Free Diffusion Segmentors Scale with the Generative PowerBenyuan Meng, Qianqian Xu, Zitai Wang, Xiaochun Cao et al.CVPR 2026 · 2 citations
Builds on33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 1,861 citations
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Conditional Prompt Learning for Vision-Language ModelsKaiyang Zhou, Jingkang Yang, Chen Change Loy, Ziwei LiuCVPR 2022 · 1,438 citations
Related papers
- SSE-SAM: Balancing Head and Tail Classes Gradually Through Stage-Wise SAMXingyu Lyu, Qianqian Xu, Zhiyong Yang, Shaojie Lyu et al.AAAI 2025 · 2 citations
- ImbSAM: A Closer Look at Sharpness-Aware Minimization in Class-Imbalanced RecognitionYixuan Zhou, Yi Qu, Xing Xu, Hengtao ShenICCV 2023 · 35 citations
- Improving Visual Prompt Tuning by Gaussian Neighborhood Minimization for Long-Tailed Visual RecognitionMengke Li, Ye Liu, Yang Lu, Yiqun Zhang et al.NeurIPS 2024 · 27 citations
- Class-Conditional Sharpness-Aware Minimization for Deep Long-Tailed RecognitionZhipeng Zhou, Lanqing Li, Peilin Zhao, Pheng-Ann Heng et al.CVPR 2023
- Balanced Gradient Penalty Improves Deep Long-Tailed LearningDong Wang, Yicheng Liu, Liangji Fang, Fanhua Shang et al.ACM MM 2022 · 7 citations
