Identifying Ambiguous Similarity Conditions via Semantic Matching
Han-Jia Ye, Yi Shi, De-Chuan Zhan
Abstract
Rich semantics inside an image result in its ambiguous relationship with others, i.e., two images could be similar in one condition but dissimilar in another. Given triplets like “aircraft” is similar to “bird” than “train”, Weakly Supervised Conditional Similarity Learning (WS-CSL) learns multiple embeddings to match semantic conditions without explicit condition labels such as “can fly”. However, similarity relationships in a triplet are uncertain except providing a condition. For example, the previous comparison becomes invalid once the conditional label changes to “is vehicle”. To this end, we introduce a novel evaluation criterion by predicting the comparison's correctness after assigning the learned embeddings to their optimal conditions, which measures how much WS-CSL could cover latent semantics as the supervised model. Furthermore, we propose the Distance Induced Semantic COndition VER-ification Network (DiscoverNet), which characterizes the instance-instance and triplets-condition relations in a “decompose-and-fuse” manner. To make the learned embeddings cover all semantics, Discovernet utilizes a set module or an additional regularizer over the correspondence between a triplet and a condition. DiscoverNet achieves state-of-the-art performance on benchmarks like UT-Zappos-50k and Celeb-A w.r.t. different criteria.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Forward Compatible Few-Shot Class-Incremental LearningDa-Wei Zhou, Fu-Yun Wang, Han-Jia Ye, Liang Ma et al.CVPR 2022 · 259 citations
- C-STS: Conditional Semantic Textual SimilarityAmeet Deshpande, Carlos E. Jimenez, Howard Chen, Vishvak Murahari et al.EMNLP 2023 · 9 citations
- Faith: An Efficient Framework for Transformer Verification on GPUsBoyuan Feng, Tianqi Tang, Yuke Wang, Zhaodong Chen et al.USENIX ATC 2022
Builds on4
- Learning Similarity Conditions Without Explicit SupervisionReuben Tan, Mariya I. Vasileva, Kate Saenko, Bryan A. PlummerICCV 2019 · 90 citations
- Learning Attribute-driven Disentangled Representations for Interactive Fashion RetrievalYuxin Hou, Eleonora Vig, Michael Donoser, Loris BazzaniICCV 2021 · 58 citations
- Towards Latent Attribute Discovery From Triplet SimilaritiesIshan Nigam, Pavel Tokmakov, Deva RamananICCV 2019 · 11 citations
- Fashion Outfit Complementary Item RetrievalYen-Liang Lin, Son Dinh Tran, Larry S. DavisCVPR 2020
Related papers
- Probabilistic Warp Consistency for Weakly-Supervised Semantic CorrespondencesPrune Truong, Martin Danelljan, Fisher Yu, Luc Van GoolCVPR 2022 · 25 citations
- Task-Independent Knowledge Makes for Transferable Representations for Generalized Zero-Shot LearningChaoqun Wang, Xuejin Chen, Shaobo Min, Xiaoyan Sun et al.AAAI 2021 · 22 citations
- Instance Similarity Learning for Unsupervised Feature RepresentationZiwei Wang, Yunsong Wang, Ziyi Wu, Jiwen Lu et al.ICCV 2021 · 8 citations
- Joint Learning of Semantic Alignment and Object Landmark DetectionSangryul Jeon, Dongbo Min, Seungryong Kim, Kwanghoon SohnICCV 2019 · 18 citations
- Cross-Image-Attention for Conditional Embeddings in Deep Metric LearningDmytro Kotovenko, Pingchuan Ma, Timo Milbich, Björn OmmerCVPR 2023
