Learning Modal-Invariant and Temporal-Memory for Video-based Visible-Infrared Person Re-Identification
Xinyu Lin, Jinxing Li, Zeyu Ma, Huafeng Li, Shuang Li, Kaixiong Xu, Guangming Lu, David Zhang
摘要
Thanks for the cross-modal retrieval techniques, visible-infrared (RGB-IR) person re-identification (Re-ID) is achieved by projecting them into a common space, allowing person Re-ID in 24-hour surveillance systems. However, with respect to the probe-to- gallery, almost all existing RGB-IR based cross-modal person Re-ID methods focus on image-to-image matching, while the video-to-video matching which contains much richer spatial- and temporal-information remains under-explored. In this paper, we primarily study the video-based cross-modal per-son Re-ID method. To achieve this task, a video-based RGB-IR dataset is constructed, in which 927 valid identities with 463,259 frames and 21,863 tracklets captured by 12 RGB/IR cameras are collected. Based on our constructed dataset, we prove that with the increase of frames in a tracklet, the performance does meet more enhancement, demonstrating the significance of video-to-video matching in RGB-IR person Re-ID. Additionally, a novel method is further proposed, which not only projects two modalities to a modal-invariant subspace, but also extracts the temporal-memory for motion-invariant. Thanks to these two strategies, much better results are achieved on our video-based cross-modal person Re-ID. The code and dataset are released at: https://github.com/VCM-project233/MITML.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- M3Net: Multi-view Encoding, Matching, and Fusion for Few-shot Fine-grained Action RecognitionHao Tang, Jun Liu, Shuanglin Yan, Rui Yan 等ACM MM 2023 · 被引用 78 次
- DINOv2 Driven Gait Representation Learning for Video-Based Visible-Infrared Person Re-identificationYujie Yang, Shuang Li, Jun Ye, Neng Dong 等ACM MM 2025 · 被引用 10 次
- TokenMatcher: Diverse Tokens Matching for Unsupervised Visible-Infrared Person Re-IdentificationXiao Wang, Lekai Liu, Bin Yang, Mang Ye 等AAAI 2025 · 被引用 8 次
- Multi-Modal Multi-Platform Person Re-Identification: Benchmark and MethodRuiyang Ha, Songyi Jiang, Bin Li, Bikang Pan 等ICCV 2025 · 被引用 4 次
- X-ReID: Multi-granularity Information Interaction for Video-Based Visible-Infrared Person Re-IdentificationChenyang Yu, Xuehu Liu, Pingping Zhang, Huchuan LuAAAI 2026 · 被引用 3 次
它引用的顶会 Paper15
- RGB-Infrared Cross-Modality Person Re-Identification via Joint Pixel and Feature AlignmentGuan'an Wang, Tianzhu Zhang, Jian Cheng, Si Liu 等ICCV 2019 · 被引用 464 次
- Infrared-Visible Cross-Modal Person Re-Identification with an X ModalityDiangang Li, Xing Wei, Xiaopeng Hong, Yihong GongAAAI 2020 · 被引用 419 次
- Channel Augmented Joint Learning for Visible-Infrared RecognitionMang Ye, Weijian Ruan, Bo Du, Mike Zheng ShouICCV 2021 · 被引用 310 次
- Learning by Aligning: Visible-Infrared Person Re-identification using Cross-Modal CorrespondencesHyunjong Park, Sanghoon Lee, Junghyup Lee, Bumsub HamICCV 2021 · 被引用 248 次
- Global-Local Temporal Representations for Video Person Re-IdentificationJianing Li, Shiliang Zhang, Jingdong Wang, Wen Gao 等ICCV 2019 · 被引用 241 次
相关 Paper
- Cross-Modality Person Re-identification with Memory-Based Contrastive EmbeddingDe Cheng, Xiaolong Wang, Nannan Wang, Zhen Wang 等AAAI 2023 · 被引用 22 次
- Similarity Metric Learning For RGB-Infrared Group Re-IdentificationJianghao Xiong, Jianhuang LaiCVPR 2023
- Co-Attentive Lifting for Infrared-Visible Person Re-IdentificationXing Wei, Diangang Li, Xiaopeng Hong, Wei Ke 等ACM MM 2020 · 被引用 61 次
- TVPR: Text-to-Video Person Retrieval and a New BenchmarkXu Zhang, Fan Ni, Guannan Dong, Aichun Zhu 等ACM MM 2024 · 被引用 2 次
- Robust Multi-Modality Person Re-identificationAihua Zheng, Zi Wang, Zi-Han Chen, Chenglong Li 等AAAI 2021 · 被引用 79 次
