Dual Part Discovery Network for Zero-Shot Learning
Jiannan Ge, Hongtao Xie, Shaobo Min, Pandeng Li, Yongdong Zhang
Abstract
Zero-Shot Learning (ZSL) aims to recognize unseen classes by transferring knowledge from seen classes. Recent methods focus on learning a common semantic space to align visual and attribute information. However, they always over-relied on provided attributes and ignored the category discriminative information that contributes to accurate unseen class recognition, resulting in weak transferability. To this end, we propose a novel Dual Part Discovery Network (DPDN) that considers both attribute and category discriminative information by discovering attribute-guided parts and category-guided parts simultaneously to improve knowledge transfer. Specifically, for attribute-guided parts discovery, DPDN can localize the regions with specific attribute information and significantly bridge the gap between visual and semantic information guided by the given attributes. For category-guided parts discovery, the local parts are explored to discover other important regions that bring latent crucial details ignored by attributes, with the guidance of adaptive category prototypes. To better mine the transferable knowledge, we impose class correlations constraints to regularize the category prototypes. Finally, attribute- and category-guided parts complement each other and provide adequate discriminative subtle information for more accurate unseen class recognition. Extensive experimental results demonstrate that DPDN can discover discriminative parts and outperform state-of-the-art methods on three standard benchmarks.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers5
- Towards Balanced Alignment: Modal-Enhanced Semantic Modeling for Video Moment RetrievalZhihang Liu, Jun Li, Hongtao Xie, Pandeng Li et al.AAAI 2024 · 49 citations
- Visual-Semantic Decomposition and Partial Alignment for Document-based Zero-Shot LearningXiangyan Qu, Jing Yu, Keke Gai, Jiamin Zhuang et al.ACM MM 2024 · 5 citations
- CLIP-Adapted Region-to-Text Learning for Generative Open-Vocabulary Semantic SegmentationJiannan Ge, Lingxi Xie, Hongtao Xie, Pandeng Li et al.ICCV 2025 · 3 citations
- Causality-Guided Prompt Learning for Vision-Language Models via Visual GranulationMengyu Gao, Qiulei DongICCV 2025 · 2 citations
- ProS: Prompting-to-Simulate Generalized Knowledge for Universal Cross-Domain RetrievalKaipeng Fang, Jingkuan Song, Lianli Gao, Pengpeng Zeng et al.CVPR 2024
Related papers
- Dual Progressive Prototype Network for Generalized Zero-Shot LearningChaoqun Wang, Shaobo Min, Xuejin Chen, Xiaoyan Sun et al.NeurIPS 2021 · 72 citations
- MSDN: Mutually Semantic Distillation Network for Zero-Shot LearningShiming Chen, Ziming Hong, Guo-Sen Xie, Wenhan Yang et al.CVPR 2022 · 141 citations
- Task-Independent Knowledge Makes for Transferable Representations for Generalized Zero-Shot LearningChaoqun Wang, Xuejin Chen, Shaobo Min, Xiaoyan Sun et al.AAAI 2021 · 22 citations
- TransZero: Attribute-Guided Transformer for Zero-Shot LearningShiming Chen, Ziming Hong, Yang Liu, Guo-Sen Xie et al.AAAI 2022 · 185 citations
- Class Semantic Attribute Perception Guided Zero-Shot LearningQin Yue, Junbiao Cui, Jianqing Liang, Liang BaiAAAI 2025 · 1 citation
