All in One Framework for Multimodal Re-Identification in the Wild
He Li, Mang Ye, Ming Zhang, Bo Du
Abstract
In Re-identification (ReID), recent advancements yield noteworthy progress in both unimodal and cross-modal retrieval tasks. However, the challenge persists in developing a unified framework that could effectively handle varying multimodal data, including RGB, infrared, sketches, and textual information. Additionally, the emergence of large-scale models shows promising performance in various vision tasks but the foundation model in ReID is still blank. In response to these challenges, a novel multimodal learning paradigm for ReID is introduced, referred to as All-in-One (AIO), which harnesses a frozen pre-trained big model as an encoder, enabling effective multimodal retrieval without additional fine-tuning. The diverse multimodal data in AIO are seamlessly tokenized into a unified space, allowing the modality-shared frozen encoder to extract identity-consistent features comprehensively across all modalities. Furthermore, a meticulously crafted ensemble of cross-modality heads is designed to guide the learning trajectory. AIO is the first framework to perform allin-one ReID, encompassing four commonly used modalities. Experiments on cross-modal and multimodal ReID reveal that AIO not only adeptly handles various modal data but also excels in challenging contexts, showcasing exceptional performance in zero-shot and domain generalization scenarios. Code will be available at: https: //github.com/lihe404/AIO .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5abf018e-a2b0-4ffe-9234-d95a2662f6c2Cited by top-tier papers15
- Cross-video Identity Correlating for Person Re-identification Pre-trainingJialong Zuo, Ying Nie, Hanyu Zhou, Huaxin Zhang et al.NeurIPS 2024 · 15 citations
- ReID5o: Achieving Omni Multi-modal Person Re-identification in a Single ModelJialong Zuo, Yongtai Deng, Mengdan Tan, Rui Jin et al.NeurIPS 2025 · 11 citations
- Cross-Modal Ship Re-Identification via Optical and SAR Imagery: A Novel Dataset and MethodHan Wang, Shengyang Li, Jian Yang, Yuxuan Liu et al.ICCV 2025 · 10 citations
- A Theory-Inspired Framework for Few-Shot Cross-Modal Sketch Person Re-IdentificationYunpeng Gong, Yongjie Hou, Jiangming Shi, Kim Long Diep et al.AAAI 2026 · 7 citations
- Optimal Transport-based Labor-free Text Prompt Modeling for Sketch Re-identificationRui Li, Tingting Ren, Jie Wen, Jinxing LiNeurIPS 2024 · 3 citations
Builds on46
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty et al.NeurIPS 2021 · 2,985 citations
Related papers
- Towards Modality-Agnostic Person Re-identification with Descriptive QueryCuiqun Chen, Mang Ye, Ding JiangCVPR 2023
- Empowering Visible-Infrared Person Re-Identification with Large Foundation ModelsZhangyi Hu, Bin Yang, Mang YeNeurIPS 2024 · 45 citations
- FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-IdentificationZhen Sun, Lei Tan, Yunhang Shen, Chengmao Cai et al.ICML 2025
- Unbiased Prototype Consistency Learning for Multi-Modal and Multi-Task Object Re-IdentificationZhongao Zhou, Bin Yang, Wenke Huang, Jun Chen et al.NeurIPS 2025 · 2 citations
- Object-Generalized Re-Identification: A Step Towards Universal Instance PerceptionShuoyi Chen, Yurui Wu, Mang YeCVPR 2026 · 1 citation
