Unbiased Prototype Consistency Learning for Multi-Modal and Multi-Task Object Re-Identification
Zhongao Zhou, Bin Yang, Wenke Huang, Jun Chen, Mang Ye
Abstract
In object re-identification (ReID) task, both cross-modal and multi-modal retrieval methods have achieved notable progress. However, existing approaches are designed for specific modality and category (person or vehicle) retrieval task, lacking generalizability to others. Acquiring multiple task-specific models would result in wasteful allocation of both training and deployment resources. To address the practical requirements for unified retrieval, we introduce Multi-Modal and Multi-Task object ReID (M 3 T-ReID). The M 3 T-ReID task aims to utilize a unified model to simultaneously achieve retrieval tasks across different modalities and different categories. Specifically, to tackle the challenges of modality distibution divergence and category semantics discrepancy posed in M 3 T-ReID, we design a novel Unbi-ased Prototype Consistency Learning (UPCL) framework, which consists of two main modules: Unbiased Prototypes-guided Modality Enhancement (UPME) and Cluster Prototype Consistency Regularization (CPCR). UPME leverages modality-unbiased prototypes to simultaneously enhance cross-modal shared features and multi-modal fused features. Additionally, CPCR regulates discriminative semantics learning with category-consistent information through prototypes clustering. Under the collaborative operation of these two modules, our model can simultaneously learn robust cross-modal shared feature and multi-modal fused feature spaces, while also exhibiting strong category-discriminative capabilities. Extensive experiments on multi-modal datasets RGBNT201 and RGBNT100 demonstrates our UPCL framework showcasing exceptional performance for M 3 T-ReID. The code is available at https://github.com/ZhouZhongao/UPCL .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0c07a5f0-db68-4898-9a25-49dc0f1f69ebCited by top-tier papers2
- WHU-MARS: A Multispectral Aerial-Ground Benchmark Towards Any-Scenario Person Re-IdentificationYuxuan Zhao, Zhongao Zhou, Bin Yang, He Li et al.CVPR 2026
- Robust Pedestrian Detection with Uncertain ModalityQian Bie, Xiao Wang, Bin Yang, Zhixi Yu et al.AAAI 2026
Builds on40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- TransReID: Transformer-based Object Re-IdentificationShuting He, Hao Luo, Pichao Wang, Fan Wang et al.ICCV 2021 · 1,172 citations
- FedProto: Federated Prototype Learning across Heterogeneous ClientsYue Tan, Guodong Long, Lu Liu, Tianyi Zhou et al.AAAI 2022 · 851 citations
- No Fear of Heterogeneity: Classifier Calibration for Federated Learning with Non-IID DataMi Luo, Fei Chen, Dapeng Hu, Yifan Zhang et al.NeurIPS 2021 · 510 citations
- Prototypical Contrastive Learning of Unsupervised RepresentationsJunnan Li, Pan Zhou, Caiming Xiong, Steven C. H. HoiICLR 2021 · 484 citations
Related papers
- X-ReID: Multi-granularity Information Interaction for Video-Based Visible-Infrared Person Re-IdentificationChenyang Yu, Xuehu Liu, Pingping Zhang, Huchuan LuAAAI 2026 · 3 citations
- Object-Generalized Re-Identification: A Step Towards Universal Instance PerceptionShuoyi Chen, Yurui Wu, Mang YeCVPR 2026 · 1 citation
- Towards Modality-Agnostic Person Re-identification with Descriptive QueryCuiqun Chen, Mang Ye, Ding JiangCVPR 2023
- Multi-Modal Object Re-identification via Sparse Mixture-of-ExpertsYingying Feng, Jie Li, Chi Xie, Lei Tan et al.ICML 2025
- ProxyTTT: Proxy-driven Test-Time Training for Multi-modal Re-identificationAihua Zheng, Zhaojun Liu, Xixi Wan, Chenglong Li et al.AAAI 2026
