Rethinking the Paradigm of Content Constraints in Unpaired Image-to-Image Translation
Xiuding Cai, Yaoyao Zhu, Dong Miao, Linjie Fu, Yu Yao
摘要
In an unpaired setting, lacking sufficient content constraints for image-to-image translation (I2I) tasks, GAN-based approaches are usually prone to model collapse. Current solutions can be divided into two categories, reconstructionbased and Siamese network-based. The former requires that the transformed or transforming image can be perfectly converted back to the original image, which is sometimes too strict and limits the generative performance. The latter involves feeding the original and generated images into a feature extractor and then matching their outputs. This is not efficient enough, and a universal feature extractor is not easily available. In this paper, we propose EnCo, a simple but efficient way to maintain the content by constraining the representational similarity in the latent space of patch-level features from the same stage of the Encoder and deCoder of the generator. For the similarity function, we use a simple MSE loss instead of contrastive loss, which is currently widely used in I2I tasks. Benefits from the design, EnCo training is extremely efficient, while the features from the encoder produce a more positive effect on the decoding, leading to more satisfying generations. In addition, we rethink the role played by discriminators in sampling patches and propose a discriminative attention-guided (DAG) patch sampling strategy to replace random sampling. DAG is parameter-free and only requires negligible computational overhead, while significantly improving the performance of the model. Extensive experiments on multiple datasets demonstrate the effectiveness and advantages of EnCo, and we achieve multiple state-of-theart compared to previous methods. Our code is available at https://github.com/XiudingCai/EnCo-pytorch .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- UDAPose: Unsupervised Domain Adaptation for Low-Light Human Pose EstimationHaopeng Chen, Yihao Ai, Kabeen Kim, Robby T. Tan 等CVPR 2026 · 被引用 1 次
- Graph-Semantic Guided Learning for Virtual Immunohistochemistry Staining on Consecutive Histology SectionsFanhao Qiu, Yangyang Zhang, Zhengxia WangAAAI 2026
它引用的顶会 Paper13
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Contrastive Learning with Hard Negative SamplesJoshua David Robinson, Ching-Yao Chuang, Suvrit Sra, Stefanie JegelkaICLR 2021 · 被引用 999 次
- FD-GAN: Generative Adversarial Networks with Fusion-Discriminator for Single Image DehazingYu Dong, Yihao Liu, He Zhang, Shifeng Chen 等AAAI 2020 · 被引用 307 次
- QS-Attn: Query-Selected Attention for Contrastive Learning in I2I TranslationXueqi Hu, Xinyue Zhou, Qiusheng Huang, Zhengyi Shi 等CVPR 2022 · 被引用 103 次
相关 Paper
- Self-Supervised Dense Consistency Regularization for Image-to-Image TranslationMinsu Ko, Eunju Cha, Sungjoo Suh, Huijin Lee 等CVPR 2022 · 被引用 25 次
- Zero-Shot Contrastive Loss for Text-Guided Diffusion Image Style TransferSerin Yang, Hyunmin Hwang, Jong Chul YeICCV 2023 · 被引用 94 次
- Unpaired Image-to-Image Translation via Latent Energy TransportYang Zhao, Changyou ChenCVPR 2021
- Large Scale Image Completion via Co-Modulated Generative Adversarial NetworksShengyu Zhao, Jonathan Cui, Yilun Sheng, Yue Dong 等ICLR 2021 · 被引用 348 次
- ReMix: Towards Image-to-Image Translation With Limited DataJie Cao, Luanxuan Hou, Ming-Hsuan Yang, Ran He 等CVPR 2021
