Attributes-Guided and Pure-Visual Attention Alignment for Few-Shot Recognition
Siteng Huang, Min Zhang, Yachen Kang, Donglin Wang
摘要
The purpose of few-shot recognition is to recognize novel categories with a limited number of labeled examples in each class. To encourage learning from a supplementary view, recent approaches have introduced auxiliary semantic modalities into effective metric-learning frameworks that aim to learn a feature similarity between training samples (support set) and test samples (query set). However, these approaches only augment the representations of samples with available semantics while ignoring the query set, which loses the potential for the improvement and may lead to a shift between the modalities combination and the pure-visual representation. In this paper, we devise an attributes-guided attention module (AGAM) to utilize human-annotated attributes and learn more discriminative features. This plug-and-play module enables visual contents and corresponding attributes to collectively focus on important channels and regions for the support set. And the feature selection is also achieved for query set with only visual information while the attributes are not available. Therefore, representations from both sets are improved in a fine-grained manner. Moreover, an attention alignment mechanism is proposed to distill knowledge from the guidance of attributes to the pure-visual branch for samples without attributes. Extensive experiments and analysis show that our proposed module can significantly improve simple metric-based approaches to achieve state-of-the-art performance on different datasets and settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- What does a platypus look like? Generating customized prompts for zero-shot image classificationSarah M. Pratt, Ian Covert, Rosanne Liu, Ali FarhadiICCV 2023 · 被引用 343 次
- Compositional Prototypical Networks for Few-Shot ClassificationQiang Lyu, Weiqiang WangAAAI 2023 · 被引用 16 次
- Open-Set Image Tagging with Multi-Grained Text SupervisionXinyu Huang, Yi-Jie Huang, Youcai Zhang, Weiwei Tian 等ACM MM 2025 · 被引用 13 次
- Do Large Language Models Pay Similar Attention Like Human Programmers When Generating Code?Bonan Kou, Shengmai Chen, Zhijie Wang, Lei Ma 等FSE 2024 · 被引用 8 次
- ATTEQ-NN: Attention-based QoE-aware Evasive Backdoor AttacksXueluan Gong, Yanjiao Chen, Jianshuo Dong, Qian WangNDSS 2022
它引用的顶会 Paper3
- Diversity With Cooperation: Ensemble Methods for Few-Shot ClassificationNikita Dvornik, Julien Mairal, Cordelia SchmidICCV 2019 · 被引用 210 次
- Learning Compositional Representations for Few-Shot RecognitionPavel Tokmakov, Yu-Xiong Wang, Martial HebertICCV 2019 · 被引用 133 次
- Few-Shot Learning via Embedding Adaptation With Set-to-Set FunctionsHan-Jia Ye, Hexiang Hu, De-Chuan Zhan, Fei ShaCVPR 2020
相关 Paper
- Frequency Guidance Matters in Few-Shot LearningHao Cheng, Siyuan Yang, Joey Tianyi Zhou, Lanqing Guo 等ICCV 2023 · 被引用 48 次
- Dynamic Extension Nets for Few-shot Semantic SegmentationLizhao Liu, Junyi Cao, Minqian Liu, Yong Guo 等ACM MM 2020 · 被引用 55 次
- Dual Attention Networks for Few-Shot Fine-Grained RecognitionShu-Lin Xu, Faen Zhang, Xiu-Shen Wei, Jianhua WangAAAI 2022 · 被引用 43 次
- Attentive Weights Generation for Few Shot Learning via Information MaximizationYiluan Guo, Ngai-Man CheungCVPR 2020
- Goal-Oriented Gaze Estimation for Zero-Shot LearningYang Liu, Lei Zhou, Xiao Bai, Yifei Huang 等CVPR 2021
