Multi-Label Few-Shot Image Classification via Pairwise Feature Augmentation and Flexible Prompt Learning
Han Liu, Yuanyuan Wang, Xiaotong Zhang, Feng Zhang, Wei Wang, Fenglong Ma, Hong Yu
摘要
Multi-label few-shot image classification is a crucial and challenging task due to limited annotated data and elusive category specificity. However, research on this topic is still in the rudimentary stage and few methods are available. Existing methods either leverage data augmentation to alleviate data scarcity or utilize label features as auxiliary knowledge to eliminate the negative effect caused by irrelevant categories, but they ignore the utilization of image region features for data augmentation, and overlook to learn appropriate text feature to better match the image features of specific categories. Moreover, these methods only focus on one side and do not effectively tackle the above two issues simultaneously. In this paper, we introduce a novel prototype-based multi-label fewshot learning framework that seamlessly integrates pairwise feature augmentation and flexible prompt learning. Specifically, by pairwise feature augmentation, we leverage the region features of images in the support set to generate more image features and construct image prototypes, thus alleviating the issue of data scarcity. By flexible prompt learning, we adaptively acquire class-specific prompts to build text prototypes that highly match the image features of specific classes, thereby mitigating the impact of irrelevant classes. Finally, with adaptive learnable parameters, we merge image and text prototypes to obtain the final prototypes, achieving a more powerful classifier for multi-label few-shot image classification. Extensive experimental results demonstrate that our proposed method can push the performance to a higher level.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Unified Named Entity Recognition as Word-Word Relation ClassificationJingye Li, Hao Fei, Jiang Liu, Shengqiong Wu 等AAAI 2022 · 被引用 340 次
- Relational Embedding for Few-Shot ClassificationDahyun Kang, Heeseung Kwon, Juhong Min, Minsu ChoICCV 2021 · 被引用 254 次
- DualCoOp: Fast Adaptation to Multi-Label Recognition with Limited AnnotationsXimeng Sun, Ping Hu, Kate SaenkoNeurIPS 2022 · 被引用 199 次
相关 Paper
- Semantic Prompt for Few-Shot Image RecognitionCVPR 2023
- MPA: Multimodal Prototype Augmentation for Few-Shot LearningLiwen Wu, Wei Wang, Lei Zhao, Zhan Gao 等AAAI 2026
- Provably Improving Generalization of Few-shot models with Synthetic DataLan-Cuong Nguyen, Quan Nguyen-Tri, Bang Tran Khanh, Dung D. Le 等ICML 2025
- Prompt-Based Meta-Learning For Few-shot Text ClassificationHaoxing Zhang, Xiaofeng Zhang, Haibo Huang, Lei YuEMNLP 2022 · 被引用 34 次
- SEC-Prompt: SEmantic Complementary Prompting for Few-Shot Class-Incremental LearningYe Liu, Meng YangCVPR 2025
