Learning Similarity Conditions Without Explicit Supervision
Reuben Tan, Mariya I. Vasileva, Kate Saenko, Bryan A. Plummer
摘要
Many real-world tasks require models to compare images along multiple similarity conditions (e.g. similarity in color, category or shape). Existing methods often reason about these complex similarity relationships by learning condition-aware embeddings. While such embeddings aid models in learning different notions of similarity, they also limit their capability to generalize to unseen categories since they require explicit labels at test time. To address this deficiency, we propose an approach that jointly learns representations for the different similarity conditions and their contributions as a latent variable without explicit supervision. Comprehensive experiments 1 across three datasets, Polyvore-Outfits, Maryland-Polyvore and UT-Zappos50k, demonstrate the effectiveness of our approach: our model outperforms the state-of-the-art methods, even those that are strongly supervised with pre-defined similarity conditions, on fill-in-the-blank, outfit compatibility prediction and triplet prediction tasks. Finally, we show that our model learns different visually-relevant semantic sub-spaces that allow it to generalize well to unseen categories.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- No Representation Rules Them All in Category DiscoverySagar Vaze, Andrea Vedaldi, Andrew ZissermanNeurIPS 2023 · 被引用 79 次
- Multimodal Compatibility Modeling via Exploring the Consistent and Complementary CorrelationsWeili Guan, Haokun Wen, Xuemeng Song, Chung-Hsing Yeh 等ACM MM 2021 · 被引用 33 次
- Learning Tuple Compatibility for Conditional Outfit RecommendationXuewen Yang, Dongliang Xie, Xin Wang, Jiangbo Yuan 等ACM MM 2020 · 被引用 25 次
- Collocation and Try-on Network: Whether an Outfit is CompatibleNa Zheng, Xuemeng Song, Qingying Niu, Xue Dong 等ACM MM 2021 · 被引用 23 次
- Complementary Factorization towards Outfit Compatibility ModelingTianyu Su, Xuemeng Song, Na Zheng, Weili Guan 等ACM MM 2021 · 被引用 18 次
相关 Paper
- Towards Latent Attribute Discovery From Triplet SimilaritiesIshan Nigam, Pavel Tokmakov, Deva RamananICCV 2019 · 被引用 11 次
- Identifying Ambiguous Similarity Conditions via Semantic MatchingHan-Jia Ye, Yi Shi, De-Chuan ZhanCVPR 2022 · 被引用 5 次
- Effectively Leveraging Attributes for Visual SimilaritySamarth Mishra, Zhongping Zhang, Yuan Shen, Ranjitha Kumar 等ICCV 2021 · 被引用 12 次
- FaD-VLP: Fashion Vision-and-Language Pre-training towards Unified Retrieval and CaptioningSuvir Mirchandani, Licheng Yu, Mengjiao Wang, Animesh Sinha 等EMNLP 2022 · 被引用 9 次
- Open-Set Representation Learning through Combinatorial EmbeddingGeeho Kim, Junoh Kang, Bohyung HanCVPR 2023
