Rethinking Zero-Shot Learning: A Conditional Visual Classification Perspective
Kai Li, Martin Renqiang Min, Yun Fu
Abstract
Zero-shot learning (ZSL) aims to recognize instances of unseen classes solely based on the semantic descriptions of the classes. Existing algorithms usually formulate it as a semantic-visual correspondence problem, by learning mappings from one feature space to the other. Despite being reasonable, previous approaches essentially discard the highly precious discriminative power of visual features in an implicit way, and thus produce undesirable results. We instead reformulate ZSL as a conditioned visual classification problem, i.e., classifying visual features based on the classifiers learned from the semantic descriptions. With this reformulation, we develop algorithms targeting various ZSL settings: For the conventional setting, we propose to train a deep neural network that directly generates visual feature classifiers from the semantic attributes with an episode-based training scheme; For the generalized setting, we concatenate the learned highly discriminative classifiers for seen classes and the generated classifiers for unseen classes to classify visual features of all classes; For the transductive setting, we exploit unlabeled data to effectively calibrate the classifier generator using a novel learning-without-forgetting self-training mechanism and guide the process by a robust generalized cross-entropy loss. Extensive experiments show that our proposed algorithms significantly outperform state-of-the-art methods by large margins on most benchmark datasets in all the ZSL settings. Our code is available at https://github. com/kailigo/cvcZSL
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4ac64ca5-9845-4a17-9c08-d40535c8f55dCited by top-tier papers32
- HSVA: Hierarchical Semantic-Visual Adaptation for Zero-Shot LearningShiming Chen, Guo-Sen Xie, Yang Liu, Qinmu Peng et al.NeurIPS 2021 · 190 citations
- FREE: Feature Refinement for Generalized Zero-Shot LearningShiming Chen, Wenjie Wang, Beihao Xia, Qinmu Peng et al.ICCV 2021 · 171 citations
- MSDN: Mutually Semantic Distillation Network for Zero-Shot LearningShiming Chen, Ziming Hong, Guo-Sen Xie, Wenhan Yang et al.CVPR 2022 · 141 citations
- Learning Feature-to-Feature Translator by Alternating Back-Propagation for Generative Zero-Shot LearningYizhe Zhu, Jianwen Xie, Bingchen Liu, Ahmed ElgammalICCV 2019 · 98 citations
- Multimodal Style Transfer via Graph CutsYulun Zhang, Chen Fang, Yilin Wang, Zhaowen Wang et al.ICCV 2019 · 92 citations
Builds on3
- Visual Semantic Reasoning for Image-Text MatchingKunpeng Li, Yulun Zhang, Kai Li, Yuanyuan Li et al.ICCV 2019 · 598 citations
- Multimodal Style Transfer via Graph CutsYulun Zhang, Chen Fang, Yilin Wang, Zhaowen Wang et al.ICCV 2019 · 92 citations
- Attention Bridging Network for Knowledge TransferKunpeng Li, Yulun Zhang, Kai Li, Yuanyuan Li et al.ICCV 2019 · 28 citations
Related papers
- Episode-Based Prototype Generating Network for Zero-Shot LearningYunlong Yu, Zhong Ji, Jungong Han, Zhongfei ZhangCVPR 2020
- Semantics Disentangling for Generalized Zero-Shot LearningZhi Chen, Yadan Luo, Ruihong Qiu, Sen Wang et al.ICCV 2021 · 143 citations
- Progressive Semantic-Guided Vision Transformer for Zero-Shot LearningShiming Chen, Wenjin Hou, Salman H. Khan, Fahad Shahbaz KhanCVPR 2024
- Causal Visual-semantic Correlation for Zero-shot LearningShuhuang Chen, Dingjie Fu, Shiming Chen, Shuo Ye et al.ACM MM 2024 · 11 citations
- Adaptive and Generative Zero-Shot LearningYu-Ying Chou, Hsuan-Tien Lin, Tyng-Luh LiuICLR 2021 · 25 citations
