A Style-aware Discriminator for Controllable Image Translation
Kunhee Kim, Sanghun Park, Eunyeong Jeon, Taehun Kim, Daijin Kim
摘要
Current image-to-image translations do not control the output domain beyond the classes used during training, nor do they interpolate between different domains well, leading to implausible results. This limitation largely arises because labels do not consider the semantic distance. To mitigate such problems, we propose a style-aware discriminator that acts as a critic as well as a style encoder to provide conditions. The style-aware discriminator learns a controllable style space using prototype-based self-supervised learning and simultaneously guides the generator. Experiments on multiple datasets verify that the proposed model outperforms current state-of-the-art image-to-image translation methods. In contrast with current methods, the proposed approach supports various applications, including style interpolation, content transplantation, and local image translation. The code is available at github.com/ kunheek/style-aware-discriminator.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Image Translation as Diffusion Visual ProgrammersCheng Han, James Chenhao Liang, Qifan Wang, Majid Rabbani 等ICLR 2024 · 被引用 19 次
- Conditional 360-degree Image Synthesis for Immersive Indoor Scene DecorationKa-Chun Shum, Hong-Wing Pang, Binh-Son Hua, Duc Thanh Nguyen 等ICCV 2023 · 被引用 16 次
- AnyTouch: Learning Unified Static-Dynamic Representation across Multiple Visuo-tactile SensorsRuoxuan Feng, Jiangyu Hu, Wenke Xia, Tianci Gao 等ICLR 2025 · 被引用 1 次
- Plug-and-Play Diffusion Features for Text-Driven Image-to-Image TranslationNarek Tumanyan, Michal Geyer, Shai Bagon, Tali DekelCVPR 2023
- 3D-Aware Multi-Class Image-to-Image Translation with NeRFsSenmao Li, Joost van de Weijer, Yaxing Wang, Fahad Shahbaz Khan 等CVPR 2023
它引用的顶会 Paper18
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
相关 Paper
- Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image TranslationYahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao 等CVPR 2021
- Style-Guided and Disentangled Representation for Robust Image-to-Image TranslationJaewoong Choi, Dae Ha Kim, Byung Cheol SongAAAI 2022 · 被引用 9 次
- StEP: Style-Based Encoder Pre-Training for Multi-Modal Image SynthesisMoustafa Meshry, Yixuan Ren, Larry S. Davis, Abhinav ShrivastavaCVPR 2021
- SPatchGAN: A Statistical Feature Based Discriminator for Unsupervised Image-to-Image TranslationXuning Shao, Weidong ZhangICCV 2021 · 被引用 34 次
- Memory-Guided Unsupervised Image-to-Image TranslationSomi Jeong, Youngjung Kim, Eungbean Lee, Kwanghoon SohnCVPR 2021
