Learning Modality-Invariant Latent Representations for Generalized Zero-shot Learning
Jingjing Li, Mengmeng Jing, Lei Zhu, Zhengming Ding, Ke Lu, Yang Yang
摘要
Recently, feature generating methods have been successfully applied to zero-shot learning (ZSL). However, most previous approaches only generate visual representations for zero-shot recognition. In fact, typical ZSL is a classic multi-modal learning protocol which consists of a visual space and a semantic space. In this paper, therefore, we present a new method which can simultaneously generate both visual representations and semantic representations so that the essential multi-modal information associated with unseen classes can be captured. Specifically, we address the most challenging issue in such a paradigm, i.e., how to handle the domain shift and thus guarantee that the learned representations are modality-invariant. To this end, we propose two strategies: 1) leveraging the mutual information between the latent visual representations and the semantic representations; 2) maximizing the entropy of the joint distribution of the two latent representations. By leveraging the two strategies, we argue that the two modalities can be well aligned. At last, extensive experiments on five widely used datasets verify that the proposed method is able to significantly outperform previous the state-of-the-arts.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- Distinguishing Unseen from Seen for Generalized Zero-shot LearningHongzu Su, Jingjing Li, Zhi Chen, Lei Zhu 等CVPR 2022 · 被引用 40 次
- Mitigating Generation Shifts for Generalized Zero-Shot LearningZhi Chen, Yadan Luo, Sen Wang, Ruihong Qiu 等ACM MM 2021 · 被引用 30 次
- Learning Aligned Cross-Modal Representation for Generalized Zero-Shot ClassificationZhiyu Fang, Xiaobin Zhu, Chun Yang, Zheng Han 等AAAI 2022 · 被引用 26 次
- RMIB: Representation Matching Information Bottleneck for Matching Text RepresentationsHaihui Pan, Zhifang Liao, Wenrui Xie, Kun HanICML 2024 · 被引用 1 次
- (ML)2P-Encoder: On Exploration of Channel-Class Correlation for Multi-Label Zero-Shot LearningZiming Liu, Song Guo, Xiaocheng Lu, Jingcai Guo 等CVPR 2023
相关 Paper
- Episode-Based Prototype Generating Network for Zero-Shot LearningYunlong Yu, Zhong Ji, Jungong Han, Zhongfei ZhangCVPR 2020
- HSVA: Hierarchical Semantic-Visual Adaptation for Zero-Shot LearningShiming Chen, Guo-Sen Xie, Yang Liu, Qinmu Peng 等NeurIPS 2021 · 被引用 190 次
- A Variational Autoencoder with Deep Embedding Model for Generalized Zero-Shot LearningPeirong Ma, Xiao HuAAAI 2020 · 被引用 43 次
- Generalized Zero-shot Learning with Multi-source Semantic Embeddings for Scene RecognitionXinhang Song, Haitao Zeng, Sixian Zhang, Luis Herranz 等ACM MM 2020 · 被引用 9 次
- Evolving Semantic Prototype Improves Generative Zero-Shot LearningShiming Chen, Wenjin Hou, Ziming Hong, Xiaohan Ding 等ICML 2023 · 被引用 33 次
