SAGE: Structured Attribute-Guided Enhancement for GZSL
Zao Zhang, Liguo Sun, Pin Lyu
Abstract
Embedding-based generalized zero-shot learning (GZSL) models often first forge robust latent semantic correlations between visual and attribute features so that knowledge can generalize to unseen categories. Despite leveraging attributes as priors and learning a shared embedding space, current methods exhibit two critical flaws. First, attributes with heterogeneous granularity are treated uniformly, leading to semantic ambiguity. Second, the source of class-level misclassification seldom aligns with attribute-level errors, preventing models from targeting the specific attributes responsible. To overcome these limitations, we introduce Structured Attribute-Guided Enhancement (SAGE), a unified framework for GZSL. Consensus-aware bidirectional attention first synchronizes visual–semantic focus regions via a mutual-distillation scheme. Next, we partition all attributes into pairwise-disjoint subsets—Global, Context, and Local—and couple them with visual features extracted at matching spatial scales. Finally, we design a cross-sample, subset-aware distillation mechanism—when a sample is misclassified, SAGE identifies the culpable attribute subset, retrieves high-confidence prototypes from a memory bank, and applies a Kullback–Leibler (KL) divergence constraint to the corresponding feature branch. Comprehensive experiments and ablations on the challenging AwA2, CUB, and SUN benchmarks demonstrate the contribution of each component, with SAGE achieving a new state-of-the-art throughout. These findings underscore SAGE’s robustness and versatility, marking a substantial advance in generalized zero-shot learning and paving the way for broader zero-resource recognition.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- HSVA: Hierarchical Semantic-Visual Adaptation for Zero-Shot LearningShiming Chen, Guo-Sen Xie, Yang Liu, Qinmu Peng et al.NeurIPS 2021 · 190 citations
- MSDN: Mutually Semantic Distillation Network for Zero-Shot LearningShiming Chen, Ziming Hong, Guo-Sen Xie, Wenhan Yang et al.CVPR 2022 · 141 citations
- En-Compactness: Self-Distillation Embedding & Contrastive Generation for Generalized Zero-Shot LearningXia Kong, Zuodong Gao, Xiaofan Li, Ming Hong et al.CVPR 2022 · 70 citations
- Closed-form Sample Probing for Learning Generative Models in Zero-shot LearningSamet Çetin, Orhun Bugra Baran, Ramazan Gokberk CinbisICLR 2022 · 10 citations
- Progressive Semantic-Guided Vision Transformer for Zero-Shot LearningShiming Chen, Wenjin Hou, Salman H. Khan, Fahad Shahbaz KhanCVPR 2024
Related papers
- Goal-Oriented Gaze Estimation for Zero-Shot LearningYang Liu, Lei Zhou, Xiao Bai, Yifei Huang et al.CVPR 2021
- Attribute Attention for Semantic Disambiguation in Zero-Shot LearningYang Liu, Jishun Guo, Deng Cai, Xiaofei HeICCV 2019 · 163 citations
- Enhancing Domain-Invariant Parts for Generalized Zero-Shot LearningYang Zhang, Songhe FengACM MM 2023 · 6 citations
- Fine-Grained Generalized Zero-Shot Learning via Dense Attribute-Based AttentionDat Huynh, Ehsan ElhamifarCVPR 2020
- Class Semantic Attribute Perception Guided Zero-Shot LearningQin Yue, Junbiao Cui, Jianqing Liang, Liang BaiAAAI 2025 · 1 citation
