Attributes-Guided and Pure-Visual Attention Alignment for Few-Shot Recognition
Siteng Huang, Min Zhang, Yachen Kang, Donglin Wang
Abstract
The purpose of few-shot recognition is to recognize novel categories with a limited number of labeled examples in each class. To encourage learning from a supplementary view, recent approaches have introduced auxiliary semantic modalities into effective metric-learning frameworks that aim to learn a feature similarity between training samples (support set) and test samples (query set). However, these approaches only augment the representations of samples with available semantics while ignoring the query set, which loses the potential for the improvement and may lead to a shift between the modalities combination and the pure-visual representation. In this paper, we devise an attributes-guided attention module (AGAM) to utilize human-annotated attributes and learn more discriminative features. This plug-and-play module enables visual contents and corresponding attributes to collectively focus on important channels and regions for the support set. And the feature selection is also achieved for query set with only visual information while the attributes are not available. Therefore, representations from both sets are improved in a fine-grained manner. Moreover, an attention alignment mechanism is proposed to distill knowledge from the guidance of attributes to the pure-visual branch for samples without attributes. Extensive experiments and analysis show that our proposed module can significantly improve simple metric-based approaches to achieve state-of-the-art performance on different datasets and settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4caf3a93-b359-4d29-bfa5-d0e59c31a848Cited by top-tier papers5
- What does a platypus look like? Generating customized prompts for zero-shot image classificationSarah M. Pratt, Ian Covert, Rosanne Liu, Ali FarhadiICCV 2023 · 343 citations
- Compositional Prototypical Networks for Few-Shot ClassificationQiang Lyu, Weiqiang WangAAAI 2023 · 16 citations
- Open-Set Image Tagging with Multi-Grained Text SupervisionXinyu Huang, Yi-Jie Huang, Youcai Zhang, Weiwei Tian et al.ACM MM 2025 · 13 citations
- Do Large Language Models Pay Similar Attention Like Human Programmers When Generating Code?Bonan Kou, Shengmai Chen, Zhijie Wang, Lei Ma et al.FSE 2024 · 8 citations
- ATTEQ-NN: Attention-based QoE-aware Evasive Backdoor AttacksXueluan Gong, Yanjiao Chen, Jianshuo Dong, Qian WangNDSS 2022
Builds on3
- Diversity With Cooperation: Ensemble Methods for Few-Shot ClassificationNikita Dvornik, Julien Mairal, Cordelia SchmidICCV 2019 · 210 citations
- Learning Compositional Representations for Few-Shot RecognitionPavel Tokmakov, Yu-Xiong Wang, Martial HebertICCV 2019 · 133 citations
- Few-Shot Learning via Embedding Adaptation With Set-to-Set FunctionsHan-Jia Ye, Hexiang Hu, De-Chuan Zhan, Fei ShaCVPR 2020
Related papers
- Frequency Guidance Matters in Few-Shot LearningHao Cheng, Siyuan Yang, Joey Tianyi Zhou, Lanqing Guo et al.ICCV 2023 · 48 citations
- Dynamic Extension Nets for Few-shot Semantic SegmentationLizhao Liu, Junyi Cao, Minqian Liu, Yong Guo et al.ACM MM 2020 · 55 citations
- Dual Attention Networks for Few-Shot Fine-Grained RecognitionShu-Lin Xu, Faen Zhang, Xiu-Shen Wei, Jianhua WangAAAI 2022 · 43 citations
- Attentive Weights Generation for Few Shot Learning via Information MaximizationYiluan Guo, Ngai-Man CheungCVPR 2020
- Goal-Oriented Gaze Estimation for Zero-Shot LearningYang Liu, Lei Zhou, Xiao Bai, Yifei Huang et al.CVPR 2021
