Ladder Loss for Coherent Visual-Semantic Embedding
Mo Zhou, Zhenxing Niu, Le Wang, Zhanning Gao, Qilin Zhang, Gang Hua
摘要
For visual-semantic embedding, the existing methods normally treat the relevance between queries and candidates in a bipolar way – relevant or irrelevant, and all “irrelevant” candidates are uniformly pushed away from the query by an equal margin in the embedding space, regardless of their various proximity to the query. This practice disregards relatively discriminative information and could lead to suboptimal ranking in the retrieval results and poorer user experience, especially in the long-tail query scenario where a matching candidate may not necessarily exist. In this paper, we introduce a continuous variable to model the relevance degree between queries and multiple candidates, and propose to learn a coherent embedding space, where candidates with higher relevance degrees are mapped closer to the query than those with lower relevance degrees. In particular, the new ladder loss is proposed by extending the triplet loss inequality to a more general inequality chain, which implements variable push-away margins according to respective relevance degrees. In addition, a proper Coherent Score metric is proposed to better measure the ranking results including those “irrelevant” candidates. Extensive experiments on multiple datasets validate the efficacy of our proposed method, which achieves significant improvement over existing state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Negative-Aware Attention Framework for Image-Text MatchingKun Zhang, Zhendong Mao, Quan Wang, Yongdong ZhangCVPR 2022 · 被引用 185 次
- CoNT: Contrastive Neural Text GenerationChenxin An, Jiangtao Feng, Kai Lv, Lingpeng Kong 等NeurIPS 2022 · 被引用 37 次
- Practical Relative Order Attack in Deep RankingMo Zhou, Le Wang, Zhenxing Niu, Qilin Zhang 等ICCV 2021 · 被引用 19 次
- Enhancing Adversarial Robustness for Deep Metric LearningMo Zhou, Vishal M. PatelCVPR 2022 · 被引用 17 次
- Structure and Semantics Preserving Document RepresentationsNatraj Raman, Sameena Shah, Manuela VelosoSIGIR 2022 · 被引用 5 次
相关 Paper
- RankCSE: Unsupervised Sentence Representations Learning via Learning to RankJiduan Liu, Jiahao Liu, Qifan Wang, Jingang Wang 等ACL 2023 · 被引用 30 次
- Scene Graph Embeddings Using Relative Similarity SupervisionParidhi Maheshwari, Ritwick Chaudhry, Vishwa VinayAAAI 2021 · 被引用 17 次
- Supervised Metric Learning to Rank for Retrieval via Contextual Similarity OptimizationChristopher Liao, Theodoros Tsiligkaridis, Brian KulisICML 2023 · 被引用 10 次
- Continual Learning for Visual Search with Backward Consistent Feature EmbeddingTimmy S. T. Wan, Jun-Cheng Chen, Tzer-Yi Wu, Chu-Song ChenCVPR 2022 · 被引用 25 次
- ARRA: Absolute-Relative Ranking Attack against Image RetrievalSiyuan Li, Xing Xu, Zailei Zhou, Yang Yang 等ACM MM 2022 · 被引用 5 次
