HSVA: Hierarchical Semantic-Visual Adaptation for Zero-Shot Learning
Shiming Chen, Guo-Sen Xie, Yang Liu, Qinmu Peng, Baigui Sun, Hao Li, Xinge You, Ling Shao
Abstract
Zero-shot learning (ZSL) tackles the unseen class recognition problem, transferring semantic knowledge from seen classes to unseen ones. Typically, to guarantee desirable knowledge transfer, a common (latent) space is adopted for associating the visual and semantic domains in ZSL. However, existing common space learning methods align the semantic and visual domains by merely mitigating distribution disagreement through one-step adaptation. This strategy is usually ineffective due to the heterogeneous nature of the feature representations in the two domains, which intrinsically contain both distribution and structure variations. To address this and advance ZSL, we propose a novel hierarchical semantic-visual adaptation (HSVA) framework. Specifically, HSVA aligns the semantic and visual domains by adopting a hierarchical two-step adaptation, i.e., structure adaptation and distribution adaptation. In the structure adaptation step, we take two task-specific encoders to encode the source data (visual domain) and the target data (semantic domain) into a structure-aligned common space. To this end, a supervised adversarial discrepancy (SAD) module is proposed to adversarially minimize the discrepancy between the predictions of two task-specific classifiers, thus making the visual and semantic feature manifolds more closely aligned. In the distribution adaptation step, we directly minimize the Wasserstein distance between the latent multivariate Gaussian distributions to align the visual and semantic distributions using a common encoder. Finally, the structure and distribution adaptation are derived in a unified framework under two partially-aligned variational autoencoders. Extensive experiments on four benchmark datasets demonstrate that HSVA achieves superior performance on both conventional and generalized ZSL. The code is available at https://github.com/shiming-chen/HSVA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8318386d-962c-4f55-97a4-a8b7425690beCited by top-tier papers31
- Model Adaptation: Historical Contrastive Learning for Unsupervised Domain Adaptation without Source DataJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuNeurIPS 2021 · 301 citations
- TransZero: Attribute-Guided Transformer for Zero-Shot LearningShiming Chen, Ziming Hong, Yang Liu, Guo-Sen Xie et al.AAAI 2022 · 185 citations
- MSDN: Mutually Semantic Distillation Network for Zero-Shot LearningShiming Chen, Ziming Hong, Guo-Sen Xie, Wenhan Yang et al.CVPR 2022 · 141 citations
- CHiLS: Zero-Shot Image Classification with Hierarchical Label SetsZachary Novack, Julian J. McAuley, Zachary Chase Lipton, Saurabh GargICML 2023 · 127 citations
- DUET: Cross-Modal Semantic Grounding for Contrastive Zero-Shot LearningZhuo Chen, Yufeng Huang, Jiaoyan Chen, Yuxia Geng et al.AAAI 2023 · 97 citations
Builds on10
- Attribute Prototype Network for Zero-Shot LearningWenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele et al.NeurIPS 2020 · 392 citations
- Transferable Contrastive Network for Generalized Zero-Shot LearningHuajie Jiang, Ruiping Wang, Shiguang Shan, Xilin ChenICCV 2019 · 200 citations
- FREE: Feature Refinement for Generalized Zero-Shot LearningShiming Chen, Wenjie Wang, Beihao Xia, Qinmu Peng et al.ICCV 2021 · 171 citations
- Rethinking Zero-Shot Learning: A Conditional Visual Classification PerspectiveKai Li, Martin Renqiang Min, Yun FuICCV 2019 · 151 citations
- Compositional Zero-Shot Learning via Fine-Grained Dense Feature CompositionDat Huynh, Ehsan ElhamifarNeurIPS 2020 · 89 citations
Related papers
- Distinguishing Unseen from Seen for Generalized Zero-shot LearningHongzu Su, Jingjing Li, Zhi Chen, Lei Zhu et al.CVPR 2022 · 40 citations
- Learning Modality-Invariant Latent Representations for Generalized Zero-shot LearningJingjing Li, Mengmeng Jing, Lei Zhu, Zhengming Ding et al.ACM MM 2020 · 35 citations
- A Variational Autoencoder with Deep Embedding Model for Generalized Zero-Shot LearningPeirong Ma, Xiao HuAAAI 2020 · 43 citations
- Causal Visual-semantic Correlation for Zero-shot LearningShuhuang Chen, Dingjie Fu, Shiming Chen, Shuo Ye et al.ACM MM 2024 · 11 citations
- Domain-Aware Visual Bias Eliminating for Generalized Zero-Shot LearningShaobo Min, Hantao Yao, Hongtao Xie, Chaoqun Wang et al.CVPR 2020
