Learning Graph Embeddings for Compositional Zero-Shot Learning
Muhammad Ferjad Naeem, Yongqin Xian, Federico Tombari, Zeynep Akata
摘要
In compositional zero-shot learning, the goal is to recognize unseen compositions (e.g. old dog) of observed visual primitives states (e.g. old, cute) and objects (e.g. car, dog) in the training set. This is challenging because the same state can for example alter the visual appearance of a dog drastically differently from a car. As a solution, we propose a novel graph formulation called Compositional Graph Embedding (CGE) that learns image features, compositional classifiers and latent representations of visual primitives in an end-to-end manner. The key to our approach is exploiting the dependency between states, objects and their compositions within a graph structure to enforce the relevant knowledge transfer from seen to unseen compositions. By learning a joint compatibility that encodes semantics between concepts, our model allows for generalization to unseen compositions without relying on an external knowledge base like WordNet. We show that in the challenging generalized compositional zero-shot setting our CGE significantly outperforms the state of the art on MIT-States and UT-Zappos. We also propose a new benchmark for this task based on the recent GQA dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper46
- BatchFormer: Learning to Explore Sample Relationships for Robust Representation LearningZhi Hou, Baosheng Yu, Dacheng TaoCVPR 2022 · 被引用 92 次
- Siamese Contrastive Embedding Network for Compositional Zero-Shot LearningXiangyu Li, Xu Yang, Kun Wei, Cheng Deng 等CVPR 2022 · 被引用 87 次
- Independent Prototype Propagation for Zero-Shot CompositionalityFrank Ruis, Gertjan J. Burghouts, Doina BucurNeurIPS 2021 · 被引用 77 次
- Disentangling Visual Embeddings for Attributes and ObjectsNirat Saini, Khoi Pham, Abhinav ShrivastavaCVPR 2022 · 被引用 74 次
- Introducing Language Guidance in Prompt-based Continual LearningMuhammad Gul Zain Ali Khan, Muhammad Ferjad Naeem, Luc Van Gool, Didier Stricker 等ICCV 2023 · 被引用 71 次
它引用的顶会 Paper7
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- DropEdge: Towards Deep Graph Convolutional Networks on Node ClassificationYu Rong, Wenbing Huang, Tingyang Xu, Junzhou HuangICLR 2020 · 被引用 1,599 次
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 被引用 1,586 次
- Task-Driven Modular Networks for Zero-Shot Compositional LearningSenthil Purushwalkam, Maximilian Nickel, Abhinav Gupta, Marc'Aurelio RanzatoICCV 2019 · 被引用 222 次
- A causal view of compositional zero-shot recognitionYuval Atzmon, Felix Kreuk, Uri Shalit, Gal ChechikNeurIPS 2020 · 被引用 163 次
相关 Paper
- On Leveraging Variational Graph Embeddings for Open World Compositional Zero-Shot LearningMuhammad Umer Anwaar, Zhihui Pan, Martin KleinsteuberACM MM 2022 · 被引用 19 次
- Open World Compositional Zero-Shot LearningMassimiliano Mancini, Muhammad Ferjad Naeem, Yongqin Xian, Zeynep AkataCVPR 2021
- Learning Conditional Attributes for Compositional Zero-Shot LearningQingsheng Wang, Lingqiao Liu, Chenchen Jing, Hao Chen 等CVPR 2023
- ProCC: Progressive Cross-Primitive Compatibility for Open-World Compositional Zero-Shot LearningFushuo Huo, Wenchao Xu, Song Guo, Jingcai Guo 等AAAI 2024 · 被引用 17 次
- KG-SP: Knowledge Guided Simple Primitives for Open World Compositional Zero-Shot LearningShyamgopal Karthik, Massimiliano Mancini, Zeynep AkataCVPR 2022 · 被引用 60 次
