SAGE: Structured Attribute-Guided Enhancement for GZSL
Zao Zhang, Liguo Sun, Pin Lyu
摘要
Embedding-based generalized zero-shot learning (GZSL) models often first forge robust latent semantic correlations between visual and attribute features so that knowledge can generalize to unseen categories. Despite leveraging attributes as priors and learning a shared embedding space, current methods exhibit two critical flaws. First, attributes with heterogeneous granularity are treated uniformly, leading to semantic ambiguity. Second, the source of class-level misclassification seldom aligns with attribute-level errors, preventing models from targeting the specific attributes responsible. To overcome these limitations, we introduce Structured Attribute-Guided Enhancement (SAGE), a unified framework for GZSL. Consensus-aware bidirectional attention first synchronizes visual–semantic focus regions via a mutual-distillation scheme. Next, we partition all attributes into pairwise-disjoint subsets—Global, Context, and Local—and couple them with visual features extracted at matching spatial scales. Finally, we design a cross-sample, subset-aware distillation mechanism—when a sample is misclassified, SAGE identifies the culpable attribute subset, retrieves high-confidence prototypes from a memory bank, and applies a Kullback–Leibler (KL) divergence constraint to the corresponding feature branch. Comprehensive experiments and ablations on the challenging AwA2, CUB, and SUN benchmarks demonstrate the contribution of each component, with SAGE achieving a new state-of-the-art throughout. These findings underscore SAGE’s robustness and versatility, marking a substantial advance in generalized zero-shot learning and paving the way for broader zero-resource recognition.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- HSVA: Hierarchical Semantic-Visual Adaptation for Zero-Shot LearningShiming Chen, Guo-Sen Xie, Yang Liu, Qinmu Peng 等NeurIPS 2021 · 被引用 190 次
- MSDN: Mutually Semantic Distillation Network for Zero-Shot LearningShiming Chen, Ziming Hong, Guo-Sen Xie, Wenhan Yang 等CVPR 2022 · 被引用 141 次
- En-Compactness: Self-Distillation Embedding & Contrastive Generation for Generalized Zero-Shot LearningXia Kong, Zuodong Gao, Xiaofan Li, Ming Hong 等CVPR 2022 · 被引用 70 次
- Closed-form Sample Probing for Learning Generative Models in Zero-shot LearningSamet Çetin, Orhun Bugra Baran, Ramazan Gokberk CinbisICLR 2022 · 被引用 10 次
- Progressive Semantic-Guided Vision Transformer for Zero-Shot LearningShiming Chen, Wenjin Hou, Salman H. Khan, Fahad Shahbaz KhanCVPR 2024
相关 Paper
- Goal-Oriented Gaze Estimation for Zero-Shot LearningYang Liu, Lei Zhou, Xiao Bai, Yifei Huang 等CVPR 2021
- Attribute Attention for Semantic Disambiguation in Zero-Shot LearningYang Liu, Jishun Guo, Deng Cai, Xiaofei HeICCV 2019 · 被引用 163 次
- Enhancing Domain-Invariant Parts for Generalized Zero-Shot LearningYang Zhang, Songhe FengACM MM 2023 · 被引用 6 次
- Fine-Grained Generalized Zero-Shot Learning via Dense Attribute-Based AttentionDat Huynh, Ehsan ElhamifarCVPR 2020
- Class Semantic Attribute Perception Guided Zero-Shot LearningQin Yue, Junbiao Cui, Jianqing Liang, Liang BaiAAAI 2025 · 被引用 1 次
