DiCA: Disambiguated Contrastive Alignment for Cross-Modal Retrieval with Partial Labels
Chao Su, Huiming Zheng, Dezhong Peng, Xu Wang
Abstract
Cross-modal retrieval aims to retrieve relevant data across different modalities. Driven by costly massive labeled data, existing cross-modal retrieval methods achieve encouraging results. To reduce annotation costs while maintaining performance, this paper focuses on an untouched but challenging problem, i.e., cross-modal retrieval with partial labels (PLCMR). PLCMR faces the dual challenges of annotation ambiguity and modality gap. To address these challenges, we propose a novel method termed disambiguated contrastive alignment (DiCA) for cross-modal retrieval with partial labels. Specifically, DiCA proposes a novel non-candidate boosted disambiguation learning mechanism (NBDL), which elaborately balances the trade-off between the losses on candidate and non-candidate labels that eliminate label ambiguity and narrow the modality gap. Moreover, DiCA presents an instance-prototype representation learning mechanism (IPRL) to enhance the model by further eliminating the modality gap at both the instance and prototype levels. Thanks to NBDL and IPRL, our DiCA effectively addresses the issues of annotation ambiguity and modality gap for cross-modal retrieval with partial labels. Experiments on four benchmarks validate the effectiveness of our proposed method, which demonstrates enhanced performance over existing state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d30950c5-21b3-4aa5-a8d4-efb1c7d7cb73Cited by top-tier papers14
- MDReID: Modality-Decoupled Learning for Any-to-Any Multi-Modal Object Re-IdentificationYingying Feng, Jie Li, Jie Hu, Yukang Zhang et al.NeurIPS 2025 · 13 citations
- Neighbor-aware Contrastive Disambiguation for Cross-Modal Hashing with Redundant AnnotationsChao Su, Likang Peng, Yuan Sun, Dezhong Peng et al.NeurIPS 2025 · 12 citations
- Interactive Cross-modal Learning for Text-3D Scene RetrievalYanglin Feng, Yongxiang Li, Yuan Sun, Yang Qin et al.NeurIPS 2025 · 9 citations
- Learning Source-Free Domain Adaptation for Visible-Infrared Person Re-IdentificationYongxiang Li, Yanglin Feng, Yuan Sun, Dezhong Peng et al.NeurIPS 2025 · 4 citations
- FedAFD: Multimodal Federated Learning via Adversarial Fusion and DistillationMin Tan, Junchao Ma, Yinfu FENG, Jiajun Ding et al.CVPR 2026 · 1 citation
Builds on13
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Progressive Identification of True Labels for Partial-Label LearningJiaqi Lv, Miao Xu, Lei Feng, Gang Niu et al.ICML 2020 · 220 citations
- Provably Consistent Partial-Label LearningLei Feng, Jiaqi Lv, Bo Han, Miao Xu et al.NeurIPS 2020 · 188 citations
- PiCO: Contrastive Label Disambiguation for Partial Label LearningHaobo Wang, Ruixuan Xiao, Yixuan Li, Lei Feng et al.ICLR 2022 · 169 citations
- Leveraged Weighted Loss for Partial Label LearningHongwei Wen, Jingyi Cui, Hanyuan Hang, Jiabin Liu et al.ICML 2021 · 119 citations
Related papers
- DiDA: Disambiguated Domain Alignment for Cross-Domain Retrieval with Partial LabelsHaoran Liu, Ying Ma, Ming Yan, Yingke Chen et al.AAAI 2024 · 13 citations
- Ambiguity-Tolerant Cross-Modal Hashing with Partial LabelsChao Su, Yanan Li, Xu Wang, Yingke Chen et al.AAAI 2026 · 1 citation
- C3CMR: Cross-Modality Cross-Instance Contrastive Learning for Cross-Media RetrievalJunsheng Wang, Tiantian Gong, Zhixiong Zeng, Changchang Sun et al.ACM MM 2022 · 12 citations
- Semi-supervised Prototype Semantic Association Learning for Robust Cross-modal RetrievalJunsheng Wang, Tiantian Gong, Yan YanSIGIR 2024 · 3 citations
- Alleviating the Inconsistency of Multimodal Data in Cross-Modal RetrievalTieying Li, Xiaochun Yang, Yiping Ke, Bin Wang et al.ICDE 2024 · 8 citations
