Maximum Spatial Perturbation Consistency for Unpaired Image-to-Image Translation
Yanwu Xu, Shaoan Xie, Wenhao Wu, Kun Zhang, Mingming Gong, Kayhan Batmanghelich
Abstract
Unpaired image-to-image translation (I2I) is an ill-posed problem, as an infinite number of translation functions can map the source domain distribution to the target distribution. Therefore, much effort has been put into designing suitable constraints, e.g., cycle consistency (CycleGAN), geometry consistency (GCGAN), and contrastive learning-based constraints (CUTGAN), that help better pose the problem. However, these well-known constraints have limitations: (1) they are either too restrictive or too weak for specific I2I tasks; (2) these methods result in content distortion when there is a significant spatial variation between the source and target domains. This paper proposes a universal regularization technique called maximum spatial perturbation consistency (MSPC), which enforces a spatial perturbation function <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> and the translation operator <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> to be commutative (i.e., <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> . In addition, we introduce two adversarial training components for learning the spatial perturbation function. The first one lets <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> compete with <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> to achieve maximum perturbation. The second one lets <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> and <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> compete with discriminators to align the spatial variations caused by the change of object size, object distortion, background interruptions, etc. Our method outperforms the state-of-the-art methods on most I2I benchmarks. We also introduce a new benchmark, namely the front face to profile face dataset, to emphasize the underlying challenges of I2I for real-world applications. We finally perform ablation experiments to study the sensitivity of our method to the severity of spatial perturbation and its effectiveness for distribution alignment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0dbe2668-3d5a-4dbe-a160-38d63a40ee6fCited by top-tier papers6
- Unsupervised Image-to-Image Translation with Density Changing RegularizationShaoan Xie, Qirong Ho, Kun ZhangNeurIPS 2022 · 38 citations
- JPEG Compression-aware Image Forgery LocalizationMenglu Wang, Xueyang Fu, Jiawei Liu, Zheng-Jun ZhaACM MM 2022 · 22 citations
- Towards Identifiable Unsupervised Domain Translation: A Diversified Distribution Matching ApproachSagar Shrestha, Xiao FuICLR 2024 · 6 citations
- Domain Transfer Becomes Identifiable via a Single AlignmentSagar Shrestha, Subash Timilsina, Hoang-Son Nguyen, Xiao FuICML 2026
- FreeDrag: Feature Dragging for Reliable Point-Based Image EditingPengyang Ling, Lin Chen, Pan Zhang, Huaian Chen et al.CVPR 2024
Builds on13
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras et al.ICCV 2019 · 668 citations
- U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationJunho Kim, Minjae Kim, Hyeonwoo Kang, Kwanghee LeeICLR 2020 · 632 citations
- Rethinking the Truly Unsupervised Image-to-Image TranslationKyungjune Baek, Yunjey Choi, Youngjung Uh, Jaejun Yoo et al.ICCV 2021 · 115 citations
- Instance-wise Hard Negative Example Generation for Contrastive Learning in Unpaired Image-to-Image TranslationWeilun Wang, Wengang Zhou, Jianmin Bao, Dong Chen et al.ICCV 2021 · 100 citations
Related papers
- Kernel of CycleGAN as a principal homogeneous spaceNikita Moriakov, Jonas Adler, Jonas TeuwenICLR 2020 · 6 citations
- On Translation and Reconstruction Guarantees of the Cycle-Consistent Generative Adversarial NetworksAnish Chakrabarty, Swagatam DasNeurIPS 2022 · 5 citations
- Breaking the Dilemma of Medical Image-to-image TranslationLingke Kong, Chenyu Lian, Detian Huang, Zhenjiang Li et al.NeurIPS 2021 · 234 citations
- Robustness via Uncertainty-aware Cycle ConsistencyUddeshya Upadhyay, Yanbei Chen, Zeynep AkataNeurIPS 2021 · 28 citations
- Semantically Robust Unpaired Image Translation for Data with Unmatched Semantics StatisticsZhiwei Jia, Bodi Yuan, Kangkang Wang, Hong Wu et al.ICCV 2021 · 26 citations
