Learning Graph Embeddings for Compositional Zero-Shot Learning
Muhammad Ferjad Naeem, Yongqin Xian, Federico Tombari, Zeynep Akata
Abstract
In compositional zero-shot learning, the goal is to recognize unseen compositions (e.g. old dog) of observed visual primitives states (e.g. old, cute) and objects (e.g. car, dog) in the training set. This is challenging because the same state can for example alter the visual appearance of a dog drastically differently from a car. As a solution, we propose a novel graph formulation called Compositional Graph Embedding (CGE) that learns image features, compositional classifiers and latent representations of visual primitives in an end-to-end manner. The key to our approach is exploiting the dependency between states, objects and their compositions within a graph structure to enforce the relevant knowledge transfer from seen to unseen compositions. By learning a joint compatibility that encodes semantics between concepts, our model allows for generalization to unseen compositions without relying on an external knowledge base like WordNet. We show that in the challenging generalized compositional zero-shot setting our CGE significantly outperforms the state of the art on MIT-States and UT-Zappos. We also propose a new benchmark for this task based on the recent GQA dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e7f6d09a-a636-4475-ad68-5e62cc0b2564Cited by top-tier papers46
- BatchFormer: Learning to Explore Sample Relationships for Robust Representation LearningZhi Hou, Baosheng Yu, Dacheng TaoCVPR 2022 · 92 citations
- Siamese Contrastive Embedding Network for Compositional Zero-Shot LearningXiangyu Li, Xu Yang, Kun Wei, Cheng Deng et al.CVPR 2022 · 87 citations
- Independent Prototype Propagation for Zero-Shot CompositionalityFrank Ruis, Gertjan J. Burghouts, Doina BucurNeurIPS 2021 · 77 citations
- Disentangling Visual Embeddings for Attributes and ObjectsNirat Saini, Khoi Pham, Abhinav ShrivastavaCVPR 2022 · 74 citations
- Introducing Language Guidance in Prompt-based Continual LearningMuhammad Gul Zain Ali Khan, Muhammad Ferjad Naeem, Luc Van Gool, Didier Stricker et al.ICCV 2023 · 71 citations
Builds on7
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding et al.ICML 2020 · 1,910 citations
- DropEdge: Towards Deep Graph Convolutional Networks on Node ClassificationYu Rong, Wenbing Huang, Tingyang Xu, Junzhou HuangICLR 2020 · 1,599 citations
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 1,586 citations
- Task-Driven Modular Networks for Zero-Shot Compositional LearningSenthil Purushwalkam, Maximilian Nickel, Abhinav Gupta, Marc'Aurelio RanzatoICCV 2019 · 222 citations
- A causal view of compositional zero-shot recognitionYuval Atzmon, Felix Kreuk, Uri Shalit, Gal ChechikNeurIPS 2020 · 163 citations
Related papers
- On Leveraging Variational Graph Embeddings for Open World Compositional Zero-Shot LearningMuhammad Umer Anwaar, Zhihui Pan, Martin KleinsteuberACM MM 2022 · 19 citations
- Open World Compositional Zero-Shot LearningMassimiliano Mancini, Muhammad Ferjad Naeem, Yongqin Xian, Zeynep AkataCVPR 2021
- Learning Conditional Attributes for Compositional Zero-Shot LearningQingsheng Wang, Lingqiao Liu, Chenchen Jing, Hao Chen et al.CVPR 2023
- ProCC: Progressive Cross-Primitive Compatibility for Open-World Compositional Zero-Shot LearningFushuo Huo, Wenchao Xu, Song Guo, Jingcai Guo et al.AAAI 2024 · 17 citations
- KG-SP: Knowledge Guided Simple Primitives for Open World Compositional Zero-Shot LearningShyamgopal Karthik, Massimiliano Mancini, Zeynep AkataCVPR 2022 · 60 citations
