Maximum Spatial Perturbation Consistency for Unpaired Image-to-Image Translation
Yanwu Xu, Shaoan Xie, Wenhao Wu, Kun Zhang, Mingming Gong, Kayhan Batmanghelich
摘要
Unpaired image-to-image translation (I2I) is an ill-posed problem, as an infinite number of translation functions can map the source domain distribution to the target distribution. Therefore, much effort has been put into designing suitable constraints, e.g., cycle consistency (CycleGAN), geometry consistency (GCGAN), and contrastive learning-based constraints (CUTGAN), that help better pose the problem. However, these well-known constraints have limitations: (1) they are either too restrictive or too weak for specific I2I tasks; (2) these methods result in content distortion when there is a significant spatial variation between the source and target domains. This paper proposes a universal regularization technique called maximum spatial perturbation consistency (MSPC), which enforces a spatial perturbation function <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> and the translation operator <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> to be commutative (i.e., <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> . In addition, we introduce two adversarial training components for learning the spatial perturbation function. The first one lets <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> compete with <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> to achieve maximum perturbation. The second one lets <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> and <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> compete with discriminators to align the spatial variations caused by the change of object size, object distortion, background interruptions, etc. Our method outperforms the state-of-the-art methods on most I2I benchmarks. We also introduce a new benchmark, namely the front face to profile face dataset, to emphasize the underlying challenges of I2I for real-world applications. We finally perform ablation experiments to study the sensitivity of our method to the severity of spatial perturbation and its effectiveness for distribution alignment.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Unsupervised Image-to-Image Translation with Density Changing RegularizationShaoan Xie, Qirong Ho, Kun ZhangNeurIPS 2022 · 被引用 38 次
- JPEG Compression-aware Image Forgery LocalizationMenglu Wang, Xueyang Fu, Jiawei Liu, Zheng-Jun ZhaACM MM 2022 · 被引用 22 次
- Towards Identifiable Unsupervised Domain Translation: A Diversified Distribution Matching ApproachSagar Shrestha, Xiao FuICLR 2024 · 被引用 6 次
- Domain Transfer Becomes Identifiable via a Single AlignmentSagar Shrestha, Subash Timilsina, Hoang-Son Nguyen, Xiao FuICML 2026
- FreeDrag: Feature Dragging for Reliable Point-Based Image EditingPengyang Ling, Lin Chen, Pan Zhang, Huaian Chen 等CVPR 2024
它引用的顶会 Paper13
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras 等ICCV 2019 · 被引用 668 次
- U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationJunho Kim, Minjae Kim, Hyeonwoo Kang, Kwanghee LeeICLR 2020 · 被引用 632 次
- Rethinking the Truly Unsupervised Image-to-Image TranslationKyungjune Baek, Yunjey Choi, Youngjung Uh, Jaejun Yoo 等ICCV 2021 · 被引用 115 次
- Instance-wise Hard Negative Example Generation for Contrastive Learning in Unpaired Image-to-Image TranslationWeilun Wang, Wengang Zhou, Jianmin Bao, Dong Chen 等ICCV 2021 · 被引用 100 次
相关 Paper
- Kernel of CycleGAN as a principal homogeneous spaceNikita Moriakov, Jonas Adler, Jonas TeuwenICLR 2020 · 被引用 6 次
- On Translation and Reconstruction Guarantees of the Cycle-Consistent Generative Adversarial NetworksAnish Chakrabarty, Swagatam DasNeurIPS 2022 · 被引用 5 次
- Breaking the Dilemma of Medical Image-to-image TranslationLingke Kong, Chenyu Lian, Detian Huang, Zhenjiang Li 等NeurIPS 2021 · 被引用 234 次
- Robustness via Uncertainty-aware Cycle ConsistencyUddeshya Upadhyay, Yanbei Chen, Zeynep AkataNeurIPS 2021 · 被引用 28 次
- Semantically Robust Unpaired Image Translation for Data with Unmatched Semantics StatisticsZhiwei Jia, Bodi Yuan, Kangkang Wang, Hong Wu 等ICCV 2021 · 被引用 26 次
