BIT: Matching-based Bi-directional Interaction Transformation Network for Visible-Infrared Person Re-Identification
Haoxuan Xu, Guanglin Niu
Abstract
Visible-Infrared Person Re-Identification (VI-ReID) is a challenging retrieval task due to the substantial modality gap between visible and infrared images. While existing methods attempt to bridge this gap by learning modality-invariant features within a shared embedding space, they often overlook the complex and implicit correlations between modalities. This limitation becomes more severe under distribution shifts, where infrared samples are often far fewer than visible ones. To address these challenges, we propose a novel network termed Bi-directional Interaction Transformation (BIT). Instead of relying on rigid feature alignment, BIT adopts a matching-based strategy that explicitly models the interaction between visible and infrared image pairs. Specifically, BIT employs an encoder-decoder architecture where the encoder extracts preliminary feature representations, and the decoder performs bi-directional feature integration and query aware scoring to enhance cross-modality correspondence. To our best knowledge, BIT is the first to introduce such pairwise matching-driven interaction in VI-ReID. Extensive experiments on several benchmarks demonstrate that our BIT achieves state-of-the-art performance, highlighting its effectiveness in the VI-ReID task.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e55b1f1-07b7-4752-8db2-4a0d6a5e7a64Builds on40
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty et al.NeurIPS 2021 · 2,985 citations
- RGB-Infrared Cross-Modality Person Re-Identification via Joint Pixel and Feature AlignmentGuan'an Wang, Tianzhu Zhang, Jian Cheng, Si Liu et al.ICCV 2019 · 464 citations
- Infrared-Visible Cross-Modal Person Re-Identification with an X ModalityDiangang Li, Xing Wei, Xiaopeng Hong, Yihong GongAAAI 2020 · 419 citations
- Cross-Modality Paired-Images Generation for RGB-Infrared Person Re-IdentificationGuan'an Wang, Tianzhu Zhang, Yang Yang, Jian Cheng et al.AAAI 2020 · 364 citations
- Channel Augmented Joint Learning for Visible-Infrared RecognitionMang Ye, Weijian Ruan, Bo Du, Mike Zheng ShouICCV 2021 · 310 citations
Related papers
- Learning by Aligning: Visible-Infrared Person Re-identification using Cross-Modal CorrespondencesHyunjong Park, Sanghoon Lee, Junghyup Lee, Bumsub HamICCV 2021 · 248 citations
- Enhancing Unsupervised Visible-Infrared Person Re-Identification with Bidirectional-Consistency Gradual MatchingXiao Teng, Xingyu Shen, Kele Xu, Long LanACM MM 2024 · 16 citations
- Modality-Aware Bias Mitigation and Invariance Learning for Unsupervised Visible-Infrared Person Re-IdentificationMenglin Wang, Xiaojin Gong, Jiachen Li, Genlin JiAAAI 2026
- Modality Unifying Network for Visible-Infrared Person Re-IdentificationHao Yu, Xu Cheng, Wei Peng, Weihao Liu et al.ICCV 2023 · 75 citations
- Co-Attentive Lifting for Infrared-Visible Person Re-IdentificationXing Wei, Diangang Li, Xiaopeng Hong, Wei Ke et al.ACM MM 2020 · 61 citations
