Towards a Unified Middle Modality Learning for Visible-Infrared Person Re-Identification
Yukang Zhang, Yan Yan, Yang Lu, Hanzi Wang
Abstract
Visible-infrared person re-identification (VI-ReID) aims to search identities of pedestrians across different spectra. In this task, one of the major challenges is the modality discrepancy between the visible (VIS) and infrared (IR) images. Some state-of-the-art methods try to design complex networks or generative methods to mitigate the modality discrepancy while ignoring the highly non-linear relationship between the two modalities of VIS and IR. In this paper, we propose a non-linear middle modality generator (MMG), which helps to reduce the modality discrepancy. Our MMG can effectively project VIS and IR images into a unified middle modality image (UMMI) space to generate middle-modality (M-modality) images. The generated M-modality images and the original images are fed into the backbone network to reduce the modality discrepancy.Furthermore, in order to pull together the two types of M-modality images generated from the VIS and IR images in the UMMI space, we propose a distribution consistency loss (DCL) to make the modality distribution of the generated M-modalities images as consistent as possible. Finally, we propose a middle modality network (MMN) to further enhance the discrimination and richness of features in an explicit manner. Extensive experiments have been conducted to validate the superiority of MMN for VI-ReID over some state-of-the-art methods on two challenging datasets. The gain of MMN is more than 11.1% and 8.4% in terms of Rank-1 and mAP, respectively, even compared with the latest state-of-the-art methods on the SYSU-MM01 dataset.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 5ba0a4d8-1573-4509-9d38-cc52f043c174Cited by top-tier papers27
- Dynamic Prototype Mask for Occluded Person Re-IdentificationLei Tan, Pingyang Dai, Rongrong Ji, Yongjian WuACM MM 2022 · 94 citations
- High-Order Structure Based Middle-Feature Learning for Visible-Infrared Person Re-identificationLiuxiang Qiu, Si Chen, Yan Yan, Jing-Hao Xue et al.AAAI 2024 · 63 citations
- MRCN: A Novel Modality Restitution and Compensation Network for Visible-Infrared Person Re-identificationYukang Zhang, Yan Yan, Jie Li, Hanzi WangAAAI 2023 · 55 citations
- Towards Grand Unified Representation Learning for Unsupervised Visible-Infrared Person Re-IdentificationBin Yang, Jun Chen, Mang YeICCV 2023 · 53 citations
- Empowering Visible-Infrared Person Re-Identification with Large Foundation ModelsZhangyi Hu, Bin Yang, Mang YeNeurIPS 2024 · 45 citations
Related papers
- Semi-supervised Visible-Infrared Person Re-identification via Modality Unification and Confidence GuidanceXiying Zheng, Yukang Zhang, Yang Lu, Hanzi WangACM MM 2024 · 3 citations
- Syncretic Modality Collaborative Learning for Visible Infrared Person Re-IdentificationZiyu Wei, Xi Yang, Nannan Wang, Xinbo GaoICCV 2021 · 173 citations
- Modality Unifying Network for Visible-Infrared Person Re-IdentificationHao Yu, Xu Cheng, Wei Peng, Weihao Liu et al.ICCV 2023 · 75 citations
- Joint Color-irrelevant Consistency Learning and Identity-aware Modality Adaptation for Visible-infrared Cross Modality Person Re-identificationZhiwei Zhao, Bin Liu, Qi Chu, Yan Lu et al.AAAI 2021 · 92 citations
- RGB-Infrared Cross-Modality Person Re-Identification via Joint Pixel and Feature AlignmentGuan'an Wang, Tianzhu Zhang, Jian Cheng, Si Liu et al.ICCV 2019 · 464 citations
