Category-specific Semantic Coherency Learning for Fine-grained Image Recognition
Shijie Wang, Zhihui Wang, Haojie Li, Wanli Ouyang
Abstract
Existing deep learning based weakly supervised fine-grained image recognition (WFGIR) methods usually pick out the discriminative regions from the high-level feature (HLF) maps directly. However, as HLF maps are derived based on spatial aggregation of convolution which is basically a pattern matching process that applies fixed filters, it is ineffective to model visual contents of same semantic but varying posture or perspective. We argue that this will cause the selected discriminative regions of same sub-category are not semantically corresponding and thus degrade the WFGIR performance. To address this issue, we propose an end-to-end Category-specific Semantic Coherency Network (CSC-Net) to semantically align the discriminative regions of the same subcategory. Specifically, CSC-Net consists of: 1) Local-to-Attribute Projecting Module (LPM), which automatically learns a set of latent attributes via collecting the category-specific semantic details while eliminating the varying spatial distributions from the local regions. 2) Latent Attribute Aligning (LAA), which aligns the latent attributes to specific semantic via graph convolution based on their discriminability, to achieve category-specific semantic coherency; 3) Attribute-to-Local Resuming Module (ARM), which resumes the original Euclidean space of latent attributes and construct latent attribute aligned feature maps by a location-embedding graph unpooling operation. Finally, the new feature maps are used which applies the category-specific semantic coherency implicitly for more accurate discriminative regions localization. Extensive experiments verify that CSC-Net yields the best performance under the same settings with most competitive approaches, on CUB Bird, Stanford-Cars, and FGVC Aircraft datasets.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 3879a1c4-e59b-4467-a8da-56c72be42b02Cited by top-tier papers6
- Dynamic Position-aware Network for Fine-grained Image RecognitionShijie Wang, Haojie Li, Zhihui Wang, Wanli OuyangAAAI 2021 · 36 citations
- Fine-Grained Retrieval Prompt TuningShijie Wang, Jianlong Chang, Zhihui Wang, Haojie Li et al.AAAI 2023 · 27 citations
- Learning to Parameterize Visual Attributes for Open-set Fine-grained RetrievalShijie Wang, Jianlong Chang, Haojie Li, Zhihui Wang et al.NeurIPS 2023 · 13 citations
- Category-Specific Nuance Exploration Network for Fine-Grained Object RetrievalShijie Wang, Zhihui Wang, Haojie Li, Wanli OuyangAAAI 2022 · 12 citations
- Trusted Fine-Grained Image Classification through Hierarchical Evidence FusionZhikang Xu, Xiaodong Yue, Ying Lv, Wei Liu et al.AAAI 2023 · 12 citations
Related papers
- Weakly Supervised Fine-Grained Image Classification via Guassian Mixture Model Oriented Discriminative LearningZhihui Wang, Shijie Wang, Shuhui Yang, Haojie Li et al.CVPR 2020
- Graph-Propagation Based Correlation Learning for Weakly Supervised Fine-Grained Image ClassificationZhuhui Wang, Shijie Wang, Haojie Li, Zhi Dou et al.AAAI 2020 · 107 citations
- Learning Hierarchal Channel Attention for Fine-grained Visual ClassificationXiang Guan, Guoqing Wang, Xing Xu, Yi BinACM MM 2021 · 12 citations
- A Weakly Supervised Fine Label Classifier Enhanced by Coarse SupervisionFariborz Taherkhani, Hadi Kazemi, Ali Dabouei, Jeremy M. Dawson et al.ICCV 2019 · 30 citations
- DANet: Divergent Activation for Weakly Supervised Object LocalizationHaolan Xue, Chang Liu, Fang Wan, Jianbin Jiao et al.ICCV 2019 · 192 citations
