Visual-Augmented Dynamic Semantic Prototype for Generative Zero-Shot Learning
Wenjin Hou, Shiming Chen, Shuhuang Chen, Ziming Hong, Yan Wang, Xuetao Feng, Salman H. Khan, Fahad Shahbaz Khan, Xinge You
摘要
Generative Zero-shot learning (ZSL) learns a generator to synthesize visual samples for unseen classes, which is an effective way to advance ZSL. However, existing generative methods rely on the conditions of Gaussian noise and the predefined semantic prototype, which limit the generator only optimized on specific seen classes rather than characterizing each visual instance, resulting in poor generalizations (e.g., overfitting to seen classes). To address this issue, we propose a novel Visual-Augmented Dynamic Semantic prototype method (termed VADS) to boost the generator to learn accurate semantic-visual mapping by fully exploiting the visual-augmented knowledge into semantic conditions. In detail, VADS consists of two modules: (1) Visual-aware Domain Knowledge Learning module (VDKL) learns the local bias and global prior of the visual features (referred to as domain visual knowledge), which replace pure Gaussian noise to provide richer prior noise information; (2) Vision-Oriented Semantic Updation module (VOSU) updates the semantic prototype according to the visual representations of the samples. Ultimately, we concatenate their output as a dynamic semantic prototype, which serves as the condition of the generator. Extensive experiments demonstrate that our VADS achieves superior CZSL and GZSL performances on three prominent datasets and outperforms other state-of-the-art methods with averaging increases by 6.4%, 5.9% and 4.2% on SUN, CUB and AWA2, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Machine Vision Therapy: Multimodal Large Language Models Can Enhance Visual Robustness via Denoising In-Context LearningZhuo Huang, Chang Liu, Yinpeng Dong, Hang Su 等ICML 2024 · 被引用 31 次
- ZeroMamba: Exploring Visual State Space Model for Zero-Shot LearningWenjin Hou, Dingjie Fu, Kun Li, Shiming Chen 等AAAI 2025 · 被引用 4 次
- SAGE: Structured Attribute-Guided Enhancement for GZSLZao Zhang, Liguo Sun, Pin LyuAAAI 2026
- Hierarchical Divide-And-Conquer Grouping for Classification Adaptation of Pre-Trained ModelsZiqian Lu, Yunlong Yu, Qinyue Tong, Jun LiuICCV 2025
- Jailbreaking the Non-Transferable Barrier via Test-Time Data DisguisingYongli Xiang, Ziming Hong, Lina Yao, Dadong Wang 等CVPR 2025
它引用的顶会 Paper29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Conditional Prompt Learning for Vision-Language ModelsKaiyang Zhou, Jingkang Yang, Chen Change Loy, Ziwei LiuCVPR 2022 · 被引用 1,438 次
- Attribute Prototype Network for Zero-Shot LearningWenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele 等NeurIPS 2020 · 被引用 392 次
相关 Paper
- Evolving Semantic Prototype Improves Generative Zero-Shot LearningShiming Chen, Wenjin Hou, Ziming Hong, Xiaohan Ding 等ICML 2023 · 被引用 33 次
- Adaptive and Generative Zero-Shot LearningYu-Ying Chou, Hsuan-Tien Lin, Tyng-Luh LiuICLR 2021 · 被引用 25 次
- Episode-Based Prototype Generating Network for Zero-Shot LearningYunlong Yu, Zhong Ji, Jungong Han, Zhongfei ZhangCVPR 2020
- Meta-Learning for Generalized Zero-Shot LearningVinay Kumar Verma, Dhanajit Brahma, Piyush RaiAAAI 2020 · 被引用 112 次
- Distinguishing Unseen from Seen for Generalized Zero-shot LearningHongzu Su, Jingjing Li, Zhi Chen, Lei Zhu 等CVPR 2022 · 被引用 40 次
