A Theory-Inspired Framework for Few-Shot Cross-Modal Sketch Person Re-Identification
Yunpeng Gong, Yongjie Hou, Jiangming Shi, Kim Long Diep, Min Jiang
摘要
Sketch-based person re-identification aims to match handdrawn sketches with RGB surveillance images, but remains challenging due to severe modality gaps and limited labeled data. To address this, we propose KTCAA, a theoretically inspired framework for few-shot cross-modal generalization. Drawing on generalization bounds, we identify two key factors affecting target risk: (1) domain discrepancy, reflecting the alignment difficulty between source and target distributions; and (2) perturbation invariance, measuring the model's robustness to modality shifts. Accordingly, we design: (1) Alignment Augmentation (AA), which applies localized sketch-style transformations to simulate target distributions and guide progressive alignment; and (2) Knowledge Transfer Catalyst (KTC), which enhances perturbation invariance by introducing worst-case modality perturbations and enforcing consistency. These modules are jointly optimized within a meta-learning paradigm that transfers alignment knowledge from data-abundant RGB domains to sketch scenarios. Experiments on multiple benchmarks show that KTCAA achieves state-of-theart performance, particularly under data-scarce conditions. The code will be available at https://github.com/ finger-monkey/REID_KTCAA .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Statistical Characteristic-Guided Denoising for Rapid High-Resolution Transmission Electron Microscopy ImagingHesong Li, Ziqi Wu, Ruiwen Shao, Ying FuCVPR 2026 · 被引用 4 次
- VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel OptimizationJiajing Lin, Shu Jiang, Qingyuan Zeng, Zhenzhong Wang 等ICLR 2026 · 被引用 4 次
- Re-evaluating Continual VQA: Toward Fair and Robust Evaluation for Multimodal Continual LearningZijian Gao, Zicheng Sun, Xingxing Zhang, Kele Xu 等CVPR 2026
- Correspondence Cognitive Learning for Multi-Modal Object Re-IdentificationChao Su, Shuying Li, Ruitao Pu, Dezhong Peng 等ICML 2026
它引用的顶会 Paper17
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Learning with Twin Noisy Labels for Visible-Infrared Person Re-IdentificationMouxing Yang, Zhenyu Huang, Peng Hu, Taihao Li 等CVPR 2022 · 被引用 248 次
- Augmented Dual-Contrastive Aggregation Learning for Unsupervised Visible-Infrared Person Re-IdentificationBin Yang, Mang Ye, Jun Chen, Zesen WuACM MM 2022 · 被引用 113 次
- Cross-Modality Perturbation Synergy Attack for Person Re-identificationYunpeng Gong, Zhun Zhong, Yansong Qu, Zhiming Luo 等NeurIPS 2024 · 被引用 67 次
- Shallow-Deep Collaborative Learning for Unsupervised Visible-Infrared Person Re-IdentificationBin Yang, Jun Chen, Mang YeCVPR 2024 · 被引用 52 次
相关 Paper
- Cross-Category Subjectivity Generalization for Style-Adaptive Sketch Re-IDZechao Hu, Zhengwei Yang, Hao Li, Zheng Wang 等ICCV 2025 · 被引用 1 次
- SG-FSL: Cross-Domain Few-Shot Learning with Style-Decoupled Augmentation and Gradient-Conflict AdjustmentYunyu Zou, Yishu Liu, Jun Liang, Bingzhi ChenACM MM 2025
- Doodle Your Keypoints: Sketch-Based Few-Shot Keypoint DetectionSubhajit Maity, Ayan Kumar Bhunia, Subhadeep Koley, Pinaki Nath Chowdhury 等ICCV 2025
- Meta Distribution Alignment for Generalizable Person Re-IdentificationHao Ni, Jingkuan Song, Xiaopeng Luo, Feng Zheng 等CVPR 2022 · 被引用 77 次
- PIRN: Prototypical-based Intra-modal Reconstruction with Normality Communication for Multi-modal Anomaly Detection.YITING LI, Xulei Yang, Jing Zhang, Sichao Tian 等ICLR 2026
