Dual Attention Networks for Few-Shot Fine-Grained Recognition
Shu-Lin Xu, Faen Zhang, Xiu-Shen Wei, Jianhua Wang
摘要
The task of few-shot fine-grained recognition is to classify images belonging to subordinate categories merely depending on few examples. Due to the fine-grained nature, it is desirable to capture subtle but discriminative part-level patterns from limited training data, which makes it a challenging problem. In this paper, to generate fine-grained tailored representations for few-shot recognition, we propose a Dual Attention Network (Dual Att-Net) consisting of two dual branches of both hard- and soft-attentions. Specifically, by producing attention guidance from deep activations of input images, our hard-attention is realized by keeping a few useful deep descriptors and forming them as a bag of multi-instance learning. Since these deep descriptors could correspond to objects' parts, the advantage of modeling as a multi-instance bag is able to exploit inherent correlation of these fine-grained parts. On the other side, a soft attended activation representation can be obtained by applying attention guidance upon original activations, which brings comprehensive attention information as the counterpart of hard-attention. After that, both outputs of dual branches are aggregated as a holistic image embedding w.r.t. input images. By performing meta-learning, we can learn a powerful image embedding in such a metric space to generalize to novel classes. Experiments on three popular fine-grained benchmark datasets show that our Dual Att-Net obviously outperforms other existing state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Cross-Layer and Cross-Sample Feature Optimization Network for Few-Shot Fine-Grained Image ClassificationZhen-Xiang Ma, Zhen-Duo Chen, Li-Jun Zhao, Zi-Chao Zhang 等AAAI 2024 · 被引用 57 次
- Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive EvaluationHong-Tao Yu, Yuxin Peng, Serge J. Belongie, Xiu-Shen WeiICLR 2026 · 被引用 21 次
- Hyperbolic Space with Hierarchical Margin Boosts Fine-Grained Learning from Coarse LabelsShu-Lin Xu, Yifan Sun, Faen Zhang, Anqi Xu 等NeurIPS 2023 · 被引用 16 次
- Channel-Spatial Support-Query Cross-Attention for Fine-Grained Few-Shot Image ClassificationShicheng Yang, Xiaoxu Li, Dongliang Chang, Zhanyu Ma 等ACM MM 2024 · 被引用 12 次
- VT-FSL: Bridging Vision and Text with LLMs for Few-Shot LearningWenhao Li, Qiangchang Wang, Xianjing Meng, Zhibin Wu 等NeurIPS 2025 · 被引用 10 次
它引用的顶会 Paper6
- Filtration and Distillation: Enhancing Region Attention for Fine-Grained Visual CategorizationChuanbin Liu, Hongtao Xie, Zheng-Jun Zha, Lingfeng Ma 等AAAI 2020 · 被引用 179 次
- Graph-Propagation Based Correlation Learning for Weakly Supervised Fine-Grained Image ClassificationZhuhui Wang, Shijie Wang, Haojie Li, Zhi Dou 等AAAI 2020 · 被引用 107 次
- Multiple Instance Active Learning for Object DetectionTianning Yuan, Fang Wan, Mengying Fu, Jianzhuang Liu 等CVPR 2021
- Attentive Weights Generation for Few Shot Learning via Information MaximizationYiluan Guo, Ngai-Man CheungCVPR 2020
- MIST: Multiple Instance Spatial TransformerBaptiste Angles, Yuhe Jin, Simon Kornblith, Andrea Tagliasacchi 等CVPR 2021
相关 Paper
- Relational Embedding for Few-Shot ClassificationDahyun Kang, Heeseung Kwon, Juhong Min, Minsu ChoICCV 2021 · 被引用 254 次
- Dense Relation Distillation With Context-Aware Aggregation for Few-Shot Object DetectionHanzhe Hu, Shuai Bai, Aoxue Li, Jinshi Cui 等CVPR 2021
- Twofold Debiasing Enhances Fine-Grained Learning with Coarse LabelsXin-yang Zhao, Jian Jin, Yangyang Li, Yazhou YaoAAAI 2025 · 被引用 2 次
- Towards Cross-Granularity Few-Shot Learning: Coarse-to-Fine Pseudo-Labeling with Visual-Semantic Meta-EmbeddingJinhai Yang, Hua Yang, Lin ChenACM MM 2021 · 被引用 16 次
- Few-shot Fine-Grained Action Recognition via Bidirectional Attention and Contrastive Meta-LearningJiahao Wang, Yunhong Wang, Sheng Liu, Annan LiACM MM 2021 · 被引用 15 次
