A Style-aware Discriminator for Controllable Image Translation
Kunhee Kim, Sanghun Park, Eunyeong Jeon, Taehun Kim, Daijin Kim
Abstract
Current image-to-image translations do not control the output domain beyond the classes used during training, nor do they interpolate between different domains well, leading to implausible results. This limitation largely arises because labels do not consider the semantic distance. To mitigate such problems, we propose a style-aware discriminator that acts as a critic as well as a style encoder to provide conditions. The style-aware discriminator learns a controllable style space using prototype-based self-supervised learning and simultaneously guides the generator. Experiments on multiple datasets verify that the proposed model outperforms current state-of-the-art image-to-image translation methods. In contrast with current methods, the proposed approach supports various applications, including style interpolation, content transplantation, and local image translation. The code is available at github.com/ kunheek/style-aware-discriminator.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Image Translation as Diffusion Visual ProgrammersCheng Han, James Chenhao Liang, Qifan Wang, Majid Rabbani et al.ICLR 2024 · 19 citations
- Conditional 360-degree Image Synthesis for Immersive Indoor Scene DecorationKa-Chun Shum, Hong-Wing Pang, Binh-Son Hua, Duc Thanh Nguyen et al.ICCV 2023 · 16 citations
- AnyTouch: Learning Unified Static-Dynamic Representation across Multiple Visuo-tactile SensorsRuoxuan Feng, Jiangyu Hu, Wenke Xia, Tianci Gao et al.ICLR 2025 · 1 citation
- Plug-and-Play Diffusion Features for Text-Driven Image-to-Image TranslationNarek Tumanyan, Michal Geyer, Shai Bagon, Tali DekelCVPR 2023
- 3D-Aware Multi-Class Image-to-Image Translation with NeRFsSenmao Li, Joost van de Weijer, Yaxing Wang, Fahad Shahbaz Khan et al.CVPR 2023
Builds on18
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or et al.ICCV 2021 · 1,437 citations
Related papers
- Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image TranslationYahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao et al.CVPR 2021
- Style-Guided and Disentangled Representation for Robust Image-to-Image TranslationJaewoong Choi, Dae Ha Kim, Byung Cheol SongAAAI 2022 · 9 citations
- StEP: Style-Based Encoder Pre-Training for Multi-Modal Image SynthesisMoustafa Meshry, Yixuan Ren, Larry S. Davis, Abhinav ShrivastavaCVPR 2021
- SPatchGAN: A Statistical Feature Based Discriminator for Unsupervised Image-to-Image TranslationXuning Shao, Weidong ZhangICCV 2021 · 34 citations
- Memory-Guided Unsupervised Image-to-Image TranslationSomi Jeong, Youngjung Kim, Eungbean Lee, Kwanghoon SohnCVPR 2021
