Try Harder: Hard Sample Generation and Learning for Cloth-Changing Person Re-ID
Hankun Liu, Yujian Zhao, Guanglin Niu
摘要
Hard samples pose a significant challenge in person re-identification (ReID) tasks, particularly in clothing-changing person Re-ID (CC-ReID). Their inherent ambiguity or similarity, coupled with the lack of explicit definitions, makes them a fundamental bottleneck. These issues not only limit the design of targeted learning strategies but also diminish the model's robustness under clothing or viewpoint changes. In this paper, we propose a novel multimodal-guided Hard Sample Generation and Learning (HSGL) framework, which is the first effort to unify textual and visual modalities to explicitly define, generate, and optimize hard samples within a unified paradigm. HSGL comprises two core components: (1) Dual-Granularity Hard Sample Generation (DGHSG), which leverages multimodal cues to synthesize semantically consistent samples, including both coarse- and fine-grained hard positives and negatives for effectively increasing the hardness and diversity of the training data. (2) Hard Sample Adaptive Learning (HSAL), which introduces a hardness-aware optimization strategy that adjusts feature distances based on textual semantic labels, encouraging the separation of hard positives and drawing hard negatives closer in the embedding space to enhance the model's discriminative capability and robustness to hard samples. Extensive experiments on multiple CC-ReID benchmarks demonstrate the effectiveness of our approach and highlight the potential of multimodal-guided hard sample generation and learning for robust CC-ReID. Notably, HSAL significantly accelerates the convergence of the targeted learning procedure and achieves state-of-the-art performance on both PRCC and LTCC datasets. The code is available at https://github.com/undooo/TryHarder-ACMMM25.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- MOS: Mitigating Optical-SAR Modality Gap for Cross-Modal Ship Re-IdentificationYujian Zhao, Hankun Liu, Guanglin NiuCVPR 2026 · 被引用 3 次
- BIT: Matching-based Bi-directional Interaction Transformation Network for Visible-Infrared Person Re-IdentificationHaoxuan Xu, Guanglin NiuCVPR 2026 · 被引用 3 次
它引用的顶会 Paper33
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Palette: Image-to-Image Diffusion ModelsChitwan Saharia, William Chan, Huiwen Chang, Chris A. Lee 等SIGGRAPH 2022 · 被引用 1,638 次
- VideoBERT: A Joint Model for Video and Language Representation LearningChen Sun, Austin Myers, Carl Vondrick, Kevin Murphy 等ICCV 2019 · 被引用 1,396 次
相关 Paper
- Identity-Clothing Similarity Modeling for Unsupervised Clothing Change Person Re-IdentificationZhiqi Pang, Junjie Wang, Lingling Zhao, Chunyu WangCVPR 2025
- Multigranular Visual-Semantic Embedding for Cloth-Changing Person Re-identificationZan Gao, Hongwei Wei, Weili Guan, Weizhi Nie 等ACM MM 2022 · 被引用 29 次
- Disentangling Identity Features from Interference Factors for Cloth-Changing Person Re-identificationYubo Li, De Cheng, Chaowei Fang, Changzhe Jiao 等ACM MM 2024 · 被引用 7 次
- Semantic-aware Consistency Network for Cloth-changing Person Re-IdentificationPeini Guo, Hong Liu, Jianbing Wu, Guoquan Wang 等ACM MM 2023 · 被引用 37 次
- Learning Concordant Attention via Target-aware Alignment for Visible-Infrared Person Re-identificationJianbing Wu, Hong Liu, Yuxin Su, Wei Shi 等ICCV 2023 · 被引用 45 次
