Not All Pixels Are Matched: Dense Contrastive Learning for Cross-Modality Person Re-Identification
Hanzhe Sun, Jun Liu, Zhizhong Zhang, Chengjie Wang, Yanyun Qu, Yuan Xie, Lizhuang Ma
Abstract
Visible-Infrared Person Re-Identification (VI-ReID) has become an emerging task for night-time surveillance systems. In order to reduce the cross-modality discrepancy, previous works either align the features via metric learning or generate synthesized cross-modality images by Generative Adversary Network. However, feature-level alignment ignores the heterogeneous data itself while generative framework suffers from the low generation quality, limiting their applications. In this paper, we propose a dense contrastive learning framework (DCLNet), which performs pixel-to-pixel dense alignment acting on the intermediate representations, rather than the final deep feature. It is a new loss function that brings views of positive pixels with same semantic information closer in shallow representation space, whilst pushing views of negative pixels apart. It naturally provides additional dense supervision and captures fine-grained pixel correspondence, reducing the modality gap from a new perspective. To implement it, a Part Aware Parsing (PAP) module and a Semantic Rectification Module (SRM) are introduced to learn and refine a semantic-guided mask, allowing us to efficiently find positive pairs only requiring instance-level supervision. Extensive experiments on the public SYSU-MM01 and RegDB datasets demonstrate the superiority of our pipeline over state-of-the-arts. Code is available at https://github.com/sunhz0117/DCLNet.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 425cd930-50b5-4995-b5f0-ef779e899fdfCited by top-tier papers7
- Beyond Domain Gap: Exploiting Subjectivity in Sketch-Based Person RetrievalKejun Lin, Zhixiang Wang, Zheng Wang, Yinqiang Zheng et al.ACM MM 2023 · 16 citations
- DINOv2 Driven Gait Representation Learning for Video-Based Visible-Infrared Person Re-identificationYujie Yang, Shuang Li, Jun Ye, Neng Dong et al.ACM MM 2025 · 10 citations
- ChatReID: Open-Ended Interactive Person Retrieval via Hierarchical Progressive Tuning for Vision Language ModelsKe Niu, Haiyang Yu, Mengyang Zhao, Teng Fu et al.ICCV 2025 · 5 citations
- Optimal Transport-based Labor-free Text Prompt Modeling for Sketch Re-identificationRui Li, Tingting Ren, Jie Wen, Jinxing LiNeurIPS 2024 · 3 citations
- Cross-Category Subjectivity Generalization for Style-Adaptive Sketch Re-IDZechao Hu, Zhengwei Yang, Hao Li, Zheng Wang et al.ICCV 2025 · 1 citation
Related papers
- Learning by Aligning: Visible-Infrared Person Re-identification using Cross-Modal CorrespondencesHyunjong Park, Sanghoon Lee, Junghyup Lee, Bumsub HamICCV 2021 · 248 citations
- Joint Color-irrelevant Consistency Learning and Identity-aware Modality Adaptation for Visible-infrared Cross Modality Person Re-identificationZhiwei Zhao, Bin Liu, Qi Chu, Yan Lu et al.AAAI 2021 · 92 citations
- Efficient Bilateral Cross-Modality Cluster Matching for Unsupervised Visible-Infrared Person ReIDDe Cheng, Lingfeng He, Nannan Wang, Shizhou Zhang et al.ACM MM 2023 · 36 citations
- Learning Concordant Attention via Target-aware Alignment for Visible-Infrared Person Re-identificationJianbing Wu, Hong Liu, Yuxin Su, Wei Shi et al.ICCV 2023 · 45 citations
- Syncretic Modality Collaborative Learning for Visible Infrared Person Re-IdentificationZiyu Wei, Xi Yang, Nannan Wang, Xinbo GaoICCV 2021 · 173 citations
