Class-Aware Patch Embedding Adaptation for Few-Shot Image Classification
Fusheng Hao, Fengxiang He, Liu Liu, Fuxiang Wu, Dacheng Tao, Jun Cheng
摘要
"A picture is worth a thousand words", significantly beyond mere a categorization. Accompanied by that, many patches of the image could have completely irrelevant meanings with the categorization if they were independently observed. This could significantly reduce the efficiency of a large family of few-shot learning algorithms, which have limited data and highly rely on the comparison of image patches. To address this issue, we propose a Class-aware Patch Embedding Adaptation (CPEA) method to learn "class-aware embeddings" of the image patches. The key idea of CPEA is to integrate patch embeddings with class-aware embeddings to make them class-relevant. Furthermore, we define a dense score matrix between class-relevant patch embeddings across images, based on which the degree of similarity between paired images is quantified. Visualization results show that CPEA concentrates patch embeddings by class, thus making them class-relevant. Extensive experiments on four benchmark datasets, miniImageNet, tieredImageNet, CIFAR-FS, and FC-100, indicate that our CPEA significantly outperforms the existing state-of-the-art methods. The source code is available at https://github.com/FushengHao/CPEA.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Simple Semantic-Aided Few-Shot LearningHai Zhang, Junzhe Xu, Shanlin Jiang, Zhenan HeCVPR 2024 · 被引用 33 次
- VT-FSL: Bridging Vision and Text with LLMs for Few-Shot LearningWenhao Li, Qiangchang Wang, Xianjing Meng, Zhibin Wu 等NeurIPS 2025 · 被引用 10 次
- OVMR: Open-Vocabulary Recognition with Multi-Modal ReferencesZehong Ma, Shiliang Zhang, Longhui Wei, Qi TianCVPR 2024 · 被引用 6 次
- DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot LearningWenhao Li, Xianjing Meng, Qiangchang Wang, Zhongyi Han 等ICLR 2026 · 被引用 4 次
- MPA: Multimodal Prototype Augmentation for Few-Shot LearningLiwen Wu, Wei Wang, Lei Zhao, Zhan Gao 等AAAI 2026
它引用的顶会 Paper43
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
相关 Paper
- Learning a Few-shot Embedding Model with Contrastive LearningChen Liu, Yanwei Fu, Chengming Xu, Siqian Yang 等AAAI 2021 · 被引用 202 次
- Generating Representative Samples for Few-Shot ClassificationJingyi Xu, Hieu LeCVPR 2022 · 被引用 96 次
- Semantic Prompt for Few-Shot Image RecognitionCVPR 2023
- Relational Embedding for Few-Shot ClassificationDahyun Kang, Heeseung Kwon, Juhong Min, Minsu ChoICCV 2021 · 被引用 254 次
- Matching Feature Sets for Few-Shot Image ClassificationArman Afrasiyabi, Hugo Larochelle, Jean-François Lalonde, Christian GagnéCVPR 2022 · 被引用 124 次
