Semantic-Guided Global-Local Collaborative Prompt Learning for Few-Shot Class Incremental Learning
yongxin yan, Weisen Chen, Xingye Chen, Yuanjie Shao, Zhengrong Zuo, Wenming Tan, Wenqi Ren, Changxin Gao, Nong Sang
Abstract
Few-Shot Class-Incremental Learning (FSCIL) poses a critical challenge in machine learning, requiring models to continuously integrate novel classes with limited samples while preserving knowledge of previously seen classes. While existing FSCIL approaches have demonstrated promising results, they still suffer from catastrophic forgetting and few-shot overfitting due to the challenge of balancing old knowledge retention with new knowledge acquisition. To address these challenges, we propose an innovative Semantic-Guided Global-Local Collaborative Prompt Learning (SGLC) framework. Built upon powerful pre-trained Vision-Language Models (VLMs), the framework first introduces a dual-alignment mechanism: globally aligning visual features with visual-textual prototypes and locally aligning multi-view visual features with local textual attribute features, which facilitates effective knowledge learning while preserving existing knowledge via frozen prototypes of previous classes. Furthermore, to alleviate overfitting, we incorporate Large Language Models (LLMs) to generate semantically rich textual descriptions, which simultaneously guide both global and local prompt learning through knowledge distillation. Extensive experiments on the miniImageNet, CIFAR-100, and CUB200 datasets demonstrate that SGLC performs favorably against the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 26567ab0-9b05-4e76-9ebd-8ff9cd7e1c06Builds on25
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Conditional Prompt Learning for Vision-Language ModelsKaiyang Zhou, Jingkang Yang, Chen Change Loy, Ziwei LiuCVPR 2022 · 1,438 citations
- Self-regulating Prompts: Foundational Model Adaptation without ForgettingMuhammad Uzair Khattak, Syed Talal Wasim, Muzammal Naseer, Salman Khan et al.ICCV 2023 · 365 citations
- MetaFSCIL: A Meta-Learning Approach for Few-Shot Class Incremental LearningZhixiang Chi, Li Gu, Huan Liu, Yang Wang et al.CVPR 2022 · 149 citations
Related papers
- Feature Decomposition-Recomposition in Large Vision-Language Model for Few-Shot Class-Incremental LearningZongyao Xue, Meina Kan, Shiguang Shan, Xilin ChenICCV 2025 · 1 citation
- Pre-trained Vision and Language Transformers are Few-Shot Incremental LearnersKeon-Hee Park, Kyungwoo Song, Gyeong-Moon ParkCVPR 2024 · 27 citations
- Quantized Residuals to Continuous Prompts for Few-Shot Class Incremental Learning in Vision-Language ModelsAbhishek Kumar Sinha, Nitant Dube, Soma BiswasCVPR 2026
- SEC-Prompt: SEmantic Complementary Prompting for Few-Shot Class-Incremental LearningYe Liu, Meng YangCVPR 2025
- Imagining Vision From Language for Few-Shot Class-Incremental LearningShuo Li, Xingchen Liu, Fang Liu, Licheng Jiao et al.ACM MM 2025 · 2 citations
