Style-Guided and Disentangled Representation for Robust Image-to-Image Translation
Jaewoong Choi, Dae Ha Kim, Byung Cheol Song
Abstract
Recently, various image-to-image translation (I2I) methods have improved mode diversity and visual quality in terms of neural networks or regularization terms. However, conventional I2I methods relies on a static decision boundary and the encoded representations in those methods are entangled with each other, so they often face with ‘mode collapse’ phenomenon. To mitigate mode collapse, 1) we design a so-called style-guided discriminator that guides an input image to the target image style based on the strategy of flexible decision boundary. 2) Also, we make the encoded representations include independent domain attributes. Based on two ideas, this paper proposes Style-Guided and Disentangled Representation for Robust Image-to-Image Translation (SRIT). SRIT showed outstanding FID by 8%, 22.8%, and 10.1% for CelebA-HQ, AFHQ, and Yosemite datasets, respectively. The translated images of SRIT reflect the styles of target domain successfully. This indicates that SRIT shows better mode diversity than previous works.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e6a9a393-aa9e-4e8e-9944-2a315861e9ebCited by top-tier papers1
Ask how each one uses itBuilds on6
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras et al.ICCV 2019 · 668 citations
- Rethinking the Truly Unsupervised Image-to-Image TranslationKyungjune Baek, Yunjey Choi, Youngjung Uh, Jaejun Yoo et al.ICCV 2021 · 115 citations
- Towards a Better Global Loss Landscape of GANsRuoyu Sun, Tiantian Fang, Alexander G. SchwingNeurIPS 2020 · 39 citations
- Unpaired Image-to-Image Translation via Latent Energy TransportYang Zhao, Changyou ChenCVPR 2021
- StarGAN v2: Diverse Image Synthesis for Multiple DomainsYunjey Choi, Youngjung Uh, Jaejun Yoo, Jung-Woo HaCVPR 2020
Related papers
- Image-to-Image Translation via Hierarchical Style DisentanglementXinyang Li, Shengchuan Zhang, Jie Hu, Liujuan Cao et al.CVPR 2021
- Retrieval Guided Unsupervised Multi-domain Image to Image TranslationRaul Gomez, Yahui Liu, Marco De Nadai, Dimosthenis Karatzas et al.ACM MM 2020 · 7 citations
- A Style-aware Discriminator for Controllable Image TranslationKunhee Kim, Sanghun Park, Eunyeong Jeon, Taehun Kim et al.CVPR 2022 · 31 citations
- Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image TranslationYahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao et al.CVPR 2021
- Harnessing the Conditioning Sensorium for Improved Image TranslationCooper Nederhood, Nicholas I. Kolkin, Deqing Fu, Jason SalavonICCV 2021 · 6 citations
