Heterogeneous Test-Time Training for Multi-Modal Person Re-identification
Zi Wang, Huaibo Huang, Aihua Zheng, Ran He
摘要
Multi-modal person re-identification (ReID) seeks to mitigate challenging lighting conditions by incorporating diverse modalities. Most existing multi-modal ReID methods concentrate on leveraging complementary multi-modal information via fusion or interaction. However, the relationships among heterogeneous modalities and the domain traits of unlabeled test data are rarely explored. In this paper, we propose a Heterogeneous Test-time Training (HTT) framework for multi-modal person ReID. We first propose a Cross-identity Inter-modal Margin (CIM) loss to amplify the differentiation among distinct identity samples. Moreover, we design a Multi-modal Test-time Training (MTT) strategy to enhance the generalization of the model by leveraging the relationships in the heterogeneous modalities and the information existing in the test data. Specifically, in the training stage, we utilize the CIM loss to further enlarge the distance between anchor and negative by forcing the inter-modal distance to maintain the margin, resulting in an enhancement of the discriminative capacity of the ultimate descriptor. Subsequently, since the test data contains characteristics of the target domain, we adapt the MTT strategy to optimize the network before the inference by using self-supervised tasks designed based on relationships among modalities. Experimental results on benchmark multi-modal ReID datasets RGBNT201, Market1501-MM, RGBN300, and RGBNT100 validate the effectiveness of the proposed method. The codes can be found at https://github.com/ziwang1121/HTT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- DeMo: Decoupled Feature-Based Mixture of Experts for Multi-Modal Object Re-IdentificationYuhao Wang, Yang Liu, Aihua Zheng, Pingping ZhangAAAI 2025 · 被引用 31 次
- MambaPro: Multi-Modal Object Re-identification with Mamba Aggregation and Synergistic PromptYuhao Wang, Xuehu Liu, Tianyu Yan, Yang Liu 等AAAI 2025 · 被引用 30 次
- MDReID: Modality-Decoupled Learning for Any-to-Any Multi-Modal Object Re-IdentificationYingying Feng, Jie Li, Jie Hu, Yukang Zhang 等NeurIPS 2025 · 被引用 13 次
- UGG-ReID: Uncertainty-Guided Graph Model for Multi-Modal Object Re-IdentificationXixi Wan, Aihua Zheng, Bo Jiang, Beibei Wang 等NeurIPS 2025 · 被引用 4 次
- Unbiased Prototype Consistency Learning for Multi-Modal and Multi-Task Object Re-IdentificationZhongao Zhou, Bin Yang, Wenke Huang, Jun Chen 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper26
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen 等ICLR 2021 · 被引用 1,731 次
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller 等ICML 2020 · 被引用 1,220 次
- Omni-Scale Feature Learning for Person Re-IdentificationKaiyang Zhou, Yongxin Yang, Andrea Cavallaro, Tao XiangICCV 2019 · 被引用 997 次
- ABD-Net: Attentive but Diverse Person Re-IdentificationTianlong Chen, Shaojin Ding, Jingyi Xie, Ye Yuan 等ICCV 2019 · 被引用 544 次
相关 Paper
- Interact, Embed, and EnlargE: Boosting Modality-Specific Representations for Multi-Modal Person Re-identificationZi Wang, Chenglong Li, Aihua Zheng, Ran He 等AAAI 2022 · 被引用 61 次
- ProxyTTT: Proxy-driven Test-Time Training for Multi-modal Re-identificationAihua Zheng, Zhaojun Liu, Xixi Wan, Chenglong Li 等AAAI 2026
- Robust Multi-Modality Person Re-identificationAihua Zheng, Zi Wang, Zi-Han Chen, Chenglong Li 等AAAI 2021 · 被引用 79 次
- Infrared-Visible Cross-Modal Person Re-Identification with an X ModalityDiangang Li, Xing Wei, Xiaopeng Hong, Yihong GongAAAI 2020 · 被引用 419 次
- Cross-Modality Person Re-identification with Memory-Based Contrastive EmbeddingDe Cheng, Xiaolong Wang, Nannan Wang, Zhen Wang 等AAAI 2023 · 被引用 22 次
