VGSE: Visually-Grounded Semantic Embeddings for Zero-Shot Learning
Wenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele, Zeynep Akata
Abstract
Human-annotated attributes serve as powerful semantic embeddings in zero-shot learning. However, their annotation process is labor-intensive and needs expert supervision. Current unsupervised semantic embeddings, i.e., word embeddings, enable knowledge transfer between classes. However, word embeddings do not always reflect visual similarities and result in inferior zero-shot performance. We propose to discover semantic embeddings containing discriminative visual properties for zero-shot learning, without requiring any human annotation. Our model visually divides a set of images from seen classes into clusters of local image regions according to their visual similarity, and further imposes their class discrimination and semantic relatedness. To associate these clusters with previously unseen classes, we use external knowledge, e.g., word embeddings and propose a novel class relation discovery module. Through quantitative and qualitative evaluation, we demonstrate that our model discovers semantic embeddings that model the visual properties of both seen and unseen classes. Furthermore, we demonstrate on three benchmarks that our visually-grounded semantic embeddings further improve performance over word embeddings across various ZSL models by a large margin. Code is available at https://github.com/wenjiaXu/VGSE
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6da558be-7937-45d6-99ac-b3d2d74d533aCited by top-tier papers18
- I2DFormer: Learning Image to Document Attention for Zero-Shot Image ClassificationMuhammad Ferjad Naeem, Yongqin Xian, Luc Van Gool, Federico TombariNeurIPS 2022 · 63 citations
- Audiovisual Generalised Zero-shot Learning with Cross-modal Attention and LanguageOtniel-Bogdan Mercea, Lukas Riesch, A. Sophia Koepke, Zeynep AkataCVPR 2022 · 54 citations
- Graph Knows Unknowns: Reformulate Zero-Shot Learning as Sample-Level Graph RecognitionJingcai Guo, Song Guo, Qihua Zhou, Ziming Liu et al.AAAI 2023 · 42 citations
- Data Distribution Distilled Generative Model for Generalized Zero-Shot RecognitionYijie Wang, Mingjian Hong, Luwen Huangfu, Sheng HuangAAAI 2024 · 21 citations
- Image-free Classifier Injection for Zero-Shot ClassificationAnders Christensen, Massimiliano Mancini, A. Sophia Koepke, Ole Winther et al.ICCV 2023 · 21 citations
Builds on7
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Attribute Prototype Network for Zero-Shot LearningWenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele et al.NeurIPS 2020 · 392 citations
- Towards Latent Attribute Discovery From Triplet SimilaritiesIshan Nigam, Pavel Tokmakov, Deva RamananICCV 2019 · 11 citations
- Field-Guide-Inspired Zero-Shot LearningUtkarsh Mall, Bharath Hariharan, Kavita BalaICCV 2021 · 10 citations
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
Related papers
- TransZero: Attribute-Guided Transformer for Zero-Shot LearningShiming Chen, Ziming Hong, Yang Liu, Guo-Sen Xie et al.AAAI 2022 · 185 citations
- Goal-Oriented Gaze Estimation for Zero-Shot LearningYang Liu, Lei Zhou, Xiao Bai, Yifei Huang et al.CVPR 2021
- Generalized Zero-shot Learning with Multi-source Semantic Embeddings for Scene RecognitionXinhang Song, Haitao Zeng, Sixian Zhang, Luis Herranz et al.ACM MM 2020 · 9 citations
- Class Semantic Attribute Perception Guided Zero-Shot LearningQin Yue, Junbiao Cui, Jianqing Liang, Liang BaiAAAI 2025 · 1 citation
- Progressive Semantic-Guided Vision Transformer for Zero-Shot LearningShiming Chen, Wenjin Hou, Salman H. Khan, Fahad Shahbaz KhanCVPR 2024
