Diffusion-based Synthetic Data Generation for Visible-Infrared Person Re-Identification
Wenbo Dai, Lijing Lu, Zhihang Li
Abstract
The performance of models is intricately linked to the abundance of training data. In Visible-Infrared person Re-IDentification (VI-ReID) tasks, collecting and annotating large-scale images of each individual under various cameras and modalities is tedious, time-expensive, costly and must comply with data protection laws, posing a severe challenge in meeting dataset requirements. Current research investigates the generation of synthetic data as an efficient and privacy-ensuring alternative to collecting real data in the field. However, a specific data synthesis technique tailored for VI-ReID models has yet to be explored. In this paper, we present a novel data generation framework, dubbed Diffusion-based VI-ReID data Expansion (DiVE), that automatically obtain massive RGB-IR paired images with identity preserving by decoupling identity and modality to improve the performance of VI-ReID models. Specifically, identity representation is acquired from a set of samples sharing the same ID, whereas the modality of images is learned by fine-tuning the Stable Diffusion (SD) on modality-specific data. DiVE extend the text-driven image synthesis to identity-preserving RGB-IR multimodal image synthesis. This approach significantly reduces data collection and annotation costs by directly incorporating synthetic data into ReID model training. Experiments have demonstrated that VI-ReID models trained on synthetic data produced by DiVE consistently exhibit notable enhancements. In particular, the state-of-the-art method, CAJ, trained with synthetic images, achieves an improvement of about 9% in mAP over the baseline on the LLCM dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fe7ac42b-139d-43aa-b409-134a995da50aCited by top-tier papers4
- BIT: Matching-based Bi-directional Interaction Transformation Network for Visible-Infrared Person Re-IdentificationHaoxuan Xu, Guanglin NiuCVPR 2026 · 3 citations
- DiffCrossGait: Trajectory-Level Alignment for 2D-3D Cross-Modal Gait Recognition via Latent DiffusionZhiyang Lu, Ming ChengICML 2026
- Learning What to Generate: A Reinforcement Learning-based Closed-Loop Augmentation Framework for Person Re-identificationXincheng Shi, Changxiao Ma, Yongfei Zhang, Yuzhuo Ma et al.ICML 2026
- Revisiting Attention in the Dark for Low-Light Person Re-IdentiffcationXiang Guo, Ruimin Hu, Dongliang Zhu, Mei WangAAAI 2026
Builds on25
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 StepsCheng Lu, Yuhao Zhou, Fan Bao, Jianfei Chen et al.NeurIPS 2022 · 2,653 citations
- ILVR: Conditioning Method for Denoising Diffusion Probabilistic ModelsJooyoung Choi, Sungwon Kim, Yonghyun Jeong, Youngjune Gwon et al.ICCV 2021 · 933 citations
Related papers
- Diverse Embedding Expansion Network and Low-Light Cross-Modality Benchmark for Visible-Infrared Person Re-identificationYukang Zhang, Hanzi WangCVPR 2023
- Viperson: Flexibly Generating Virtual Identity for Person Re-IdentificationXiao-Wen Zhang, Delong Zhang, Yi-Xing Peng, Zhi Ouyang et al.ICCV 2025 · 2 citations
- DiffTV: Identity-Preserved Thermal-to-Visible Face Translation via Feature Alignment and Dual-Stage ConditionsJingyu Lin, Guiqin Zhao, Jing Xu, Guoli Wang et al.ACM MM 2024 · 9 citations
- Prior-Free Augmentation for Cloth-Changing Person Re-IdentificationJiajun Zhang, Xin Li, Si Wu, Yong Xu et al.ACM MM 2025
- Empowering Visible-Infrared Person Re-Identification with Large Foundation ModelsZhangyi Hu, Bin Yang, Mang YeNeurIPS 2024 · 45 citations
