Your Negative May not Be True Negative: Boosting Image-Text Matching with False Negative Elimination
Haoxuan Li, Yi Bin, Junrong Liao, Yang Yang, Heng Tao Shen
摘要
Most existing image-text matching methods adopt triplet loss as the optimization objective, and choosing a proper negative sample for the triplet of is important for effectively training the model, e.g., hard negatives make the model learn efficiently and effectively. However, we observe that existing methods mainly employ the most similar samples as hard negatives, which may not be true negatives. In other words, the samples with high similarity but not paired with the anchor may reserve positive semantic associations, and we call them false negatives. Repelling these false negatives in triplet loss would mislead the semantic representation learning and result in inferior retrieval performance. In this paper, we propose a novel False Negative Elimination (FNE) strategy to select negatives via sampling, which could alleviate the problem introduced by false negatives. Specifically, we first construct the distributions of positive and negative samples separately via their similarities with the anchor, based on the features extracted from image and text encoders. Then we calculate the false negative probability of a given sample based on its similarity with the anchor and the above distributions via the Bayes' rule, which is employed as the sampling weight during negative sampling process. Since there may not exist any false negative in a small batch size, we design a memory module with momentum to retain a large negative buffer and implement our negative sampling strategy spanning over the buffer. In addition, to make the model focus on hard negatives, we reassign the sampling weights for the simple negatives with a cut-down strategy. The extensive experiments are conducted on Flickr30K and MS-COCO, and the results demonstrate the superiority of our proposed false negative elimination strategy. The code is available at https://github.com/LuminosityX/FNE.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Empowering Collaborative Filtering with Principled Adversarial Contrastive LossAn Zhang, Leheng Sheng, Zhibo Cai, Xiang Wang 等NeurIPS 2023 · 被引用 56 次
- Unifying Two-Stream Encoders with Transformers for Cross-Modal RetrievalYi Bin, Haoxuan Li, Yahui Xu, Xing Xu 等ACM MM 2023 · 被引用 33 次
- Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken GenerationYongqi Li, Hongru Cai, Wenjie Wang, Leigang Qu 等SIGIR 2025 · 被引用 6 次
- MM-Forecast: A Multimodal Approach to Temporal Event Forecasting with Large Language ModelsHaoxuan Li, Zhengmao Yang, Yunshan Ma, Yi Bin 等ACM MM 2024 · 被引用 4 次
- Leveraging Weak Cross-Modal Guidance for Coherence Modelling via Iterative LearningYi Bin, Junrong Liao, Yujuan Ding, Haoxuan Li 等ACM MM 2024 · 被引用 3 次
它引用的顶会 Paper16
- Visual Semantic Reasoning for Image-Text MatchingKunpeng Li, Yulun Zhang, Kai Li, Yuanyuan Li 等ICCV 2019 · 被引用 598 次
- CAMP: Cross-Modal Adaptive Message Passing for Text-Image RetrievalZihao Wang, Xihui Liu, Hongsheng Li, Lu Sheng 等ICCV 2019 · 被引用 349 次
- Dynamic Modality Interaction Modeling for Image-Text RetrievalLeigang Qu, Meng Liu, Jianlong Wu, Zan Gao 等SIGIR 2021 · 被引用 187 次
- Negative-Aware Attention Framework for Image-Text MatchingKun Zhang, Zhendong Mao, Quan Wang, Yongdong ZhangCVPR 2022 · 被引用 185 次
- Context-Aware Multi-View Summarization Network for Image-Text MatchingLeigang Qu, Meng Liu, Da Cao, Liqiang Nie 等ACM MM 2020 · 被引用 159 次
相关 Paper
- TriSim: Tri-Dimensional Similarity Modeling with Extreme Value Theory for False-Negative Mitigation in Remote Sensing Image-Text RetrievalChengyu Zheng, Hanzhang Lu, Jie Nie, Shan DuCVPR 2026
- Expanding the Scope of Negatives: Boosting Image-Text Matching with Negatives Distribution Guided LearningZhao Zhou, Weizhong Zhang, Xiangcheng Du, Yingbin Zheng 等AAAI 2025
- FALCON: False-Negative Aware Learning of Contrastive Negatives in Vision-Language AlignmentMyunsoo Kim, Seong-Woong Shim, Byung-Jun LeeCVPR 2026 · 被引用 2 次
- Difficulty-Based Sampling for Debiased Contrastive Representation LearningTaeuk Jang, Xiaoqian WangCVPR 2023
- Discovering Global False Negatives On the Fly for Self-supervised Contrastive LearningVicente Balmaseda, Bokun Wang, Ching-Long Lin, Tianbao YangICML 2025
