On the Difficulty of Unpaired Infrared-to-Visible Video Translation: Fine-Grained Content-Rich Patches Transfer
Zhenjie Yu, Shuang Li, Yirui Shen, Chi Harold Liu, Shuigen Wang
摘要
Explicit visible videos can provide sufficient visual information and facilitate vision applications. Unfortunately, the image sensors of visible cameras are sensitive to light conditions like darkness or overexposure. To make up for this, recently, infrared sensors capable of stable imaging have received increasing attention in autonomous driving and monitoring. However, most prosperous vision models are still trained on massive clear visible data, facing huge visual gaps when deploying to infrared imaging scenarios. In such cases, transferring the infrared video to a distinct visible one with fine-grained semantic patterns is a worthwhile endeavor. Previous works improve the outputs by equally optimizing each patch on the translated visible results, which is unfair for enhancing the details on content-rich patches due to the long-tail effect of pixel distribution. Here we propose a novel CPTrans framework to tackle the challenge via balancing gradients of different patches, achieving the fine-grained Content-rich Patches Transferring. Specifically, the content-aware optimization module encourages model optimization along gradients of target patches, ensuring the improvement of visual details. Additionally, the content-aware temporal normalization module enforces the generator to be robust to the motions of target patches. Moreover, we extend the existing dataset In-fraredCity to more challenging adverse weather conditions (rain and snow), dubbed as InfraredCity-Adverse 1 . Extensive experiments show that the proposed CPTrans achieves state-of-the-art performance under diverse scenes while requiring less training time than competitive methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Bridging Human Evaluation to Infrared and Visible Image FusionJinyuan Liu, Xingyuan Li, Qingyun Mei, HaoYuan Xu 等CVPR 2026 · 被引用 4 次
- DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image FusionJinyuan Liu, Bowei Zhang, Qingyun Mei, Xingyuan Li 等CVPR 2025
- Thermal-Physics Guided Infrared Image Super-Resolution with Dynamic High-Frequency AmplificationMingxuan Zhou, Yirui Shen, Shuang Li, Jing Geng 等AAAI 2026
- Thermal Diffusion Matters: Infrared Spatial-Temporal Video Super-Resolution through Heat Conduction PriorsMingxuan Zhou, Shuang Li, Yutang Zhang, Jing Geng 等CVPR 2026
它引用的顶会 Paper15
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Balanced Contrastive Learning for Long-Tailed Visual RecognitionJianggang Zhu, Zheng Wang, Jingjing Chen, Yi-Ping Phoebe Chen 等CVPR 2022 · 被引用 194 次
- Exploring Patch-wise Semantic Relation for Contrastive Learning in Image-to-Image Translation TasksChanyong Jung, Gihyun Kwon, Jong Chul YeCVPR 2022 · 被引用 103 次
相关 Paper
- I2V-GAN: Unpaired Infrared-to-Visible Video TranslationShuang Li, Bingfeng Han, Zhenjie Yu, Chi Harold Liu 等ACM MM 2021 · 被引用 58 次
- ROMA: Cross-Domain Region Similarity Matching for Unpaired Nighttime Infrared to Daytime Visible Video TranslationZhenjie Yu, Kai Chen, Shuang Li, Bingfeng Han 等ACM MM 2022 · 被引用 20 次
- Style Transfer Meets Super-Resolution: Advancing Unpaired Infrared-to-Visible Image Translation with Detail EnhancementYirui Shen, Jingxuan Kang, Shuang Li, Zhenjie Yu 等ACM MM 2023 · 被引用 11 次
- UniRGB-IR: A Unified Framework for Visible-Infrared Semantic Tasks via Adapter TuningMaoxun Yuan, Bo Cui, Tianyi Zhao, Jiayi Wang 等ACM MM 2025 · 被引用 21 次
- CDUPatch: Color-Driven Universal Adversarial Patch Attack for Dual-Modal Visible-Infrared DetectorsJiahuan Long, Wen Yao, Tingsong Jiang, Jiacheng Hou 等ACM MM 2025 · 被引用 7 次
