Bridging the Gap between Label- and Reference-based Synthesis in Multi-attribute Image-to-Image Translation
Qiusheng Huang, Zhilin Zheng, Xueqi Hu, Li Sun, Qingli Li
Abstract
The image-to-image translation (I2TT) model takes a target label or a reference image as the input, and changes a source into the specified target domain. The two types of synthesis, either label- or reference-based, have substantial differences. Particularly, the label-based synthesis reflects the common characteristics of the target domain, and the reference-based shows the specific style similar to the reference. This paper intends to bridge the gap between them in the task of multi-attribute I2TT. We design the label- and reference-based encoding modules (LEM and REM) to compare the domain differences. They first transfer the source image and target label (or reference) into a common embedding space, by providing the opposite directions through the attribute difference vector. Then the two embeddings are simply fused together to form the latent code Srand (or Sref), reflecting the domain style differences, which is injected into each layer of the generator by SPADE. To link LEM and REM, so that two types of results benefit each other, we encourage the two latent codes to be close, and set up the cycle consistency between the forward and backward translations on them. Moreover, the interpolation between the Srand and Sref is also used to synthesize an extra image. Experiments show that label- and reference-based synthesis are indeed mutually promoted, so that we can have the diverse results from LEM, and high quality results with the similar style of the reference. Code will be available at https://github.com/huangqiusheng/BridgeGAN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5dba7daa-6d3e-45a9-9bd0-73edd46e8d58Cited by top-tier papers2
- Style Transformer for Image Inversion and EditingXueqi Hu, Qiusheng Huang, Zhengyi Shi, Siyuan Li et al.CVPR 2022 · 58 citations
- Unpaired Image-to-Image Translation with Shortest Path RegularizationShaoan Xie, Yanwu Xu, Mingming Gong, Kun ZhangCVPR 2023
Builds on5
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras et al.ICCV 2019 · 668 citations
- RelGAN: Multi-Domain Image-to-Image Translation via Relative AttributesYu-Jing Lin, Po-Wei Wu, Che-Han Chang, Edward Y. Chang et al.ICCV 2019 · 158 citations
- Rethinking the Truly Unsupervised Image-to-Image TranslationKyungjune Baek, Yunjey Choi, Youngjung Uh, Jaejun Yoo et al.ICCV 2021 · 115 citations
- Gradient Origin NetworksSam Bond-Taylor, Chris G. WillcocksICLR 2021 · 2 citations
- StarGAN v2: Diverse Image Synthesis for Multiple DomainsYunjey Choi, Youngjung Uh, Jaejun Yoo, Jung-Woo HaCVPR 2020
Related papers
- StEP: Style-Based Encoder Pre-Training for Multi-Modal Image SynthesisMoustafa Meshry, Yixuan Ren, Larry S. Davis, Abhinav ShrivastavaCVPR 2021
- Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image TranslationYahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao et al.CVPR 2021
- Style-Guided and Disentangled Representation for Robust Image-to-Image TranslationJaewoong Choi, Dae Ha Kim, Byung Cheol SongAAAI 2022 · 9 citations
- Unpaired Image-to-Image Translation via Latent Energy TransportYang Zhao, Changyou ChenCVPR 2021
- One-Shot Generative Domain AdaptationCeyuan Yang, Yujun Shen, Zhiyi Zhang, Yinghao Xu et al.ICCV 2023 · 69 citations
