Diagonal Attention and Style-based GAN for Content-Style Disentanglement in Image Generation and Translation
Gihyun Kwon, Jong Chul Ye
Abstract
One of the important research topics in image generative models is to disentangle the spatial contents and styles for their separate control. Although StyleGAN can generate content feature vectors from random noises, the resulting spatial content control is primarily intended for minor spatial variations, and the disentanglement of global content and styles is by no means complete. Inspired by a mathematical understanding of normalization and attention, here we present a novel hierarchical adaptive Diagonal spatial ATtention (DAT) layers to separately manipulate the spatial contents from styles in a hierarchical manner. Using DAT and AdaIN, our method enables coarse-to-fine level disentanglement of spatial contents and styles. In addition, our generator can be easily integrated into the GAN inversion framework so that the content and style of translated images from multi-domain image translation tasks can be flexibly controlled. By using various datasets, we confirm that the proposed method not only outperforms the existing models in disentanglement scores, but also provides more flexible control over spatial features in the generated images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ead5ac31-e89e-487a-a961-610142c71875Cited by top-tier papers14
- StyleDiffusion: Controllable Disentangled Style Transfer via Diffusion ModelsZhizhong Wang, Lei Zhao, Wei XingICCV 2023 · 219 citations
- Pastiche Master: Exemplar-Based High-Resolution Portrait Style TransferShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyCVPR 2022 · 130 citations
- TransEditor: Transformer-Based Dual-Space GAN for Highly Controllable Facial EditingYanbo Xu, Yueqin Yin, Liming Jiang, Qianyi Wu et al.CVPR 2022 · 53 citations
- Style Equalization: Unsupervised Learning of Controllable Generative Sequence ModelsJen-Hao Rick Chang, Ashish Shrivastava, Hema Koppula, Xiaoshuai Zhang et al.ICML 2022 · 21 citations
- Exploring Negatives in Contrastive Learning for Unpaired Image-to-Image TranslationYupei Lin, Sen Zhang, Tianshui Chen, Yongyi Lu et al.ACM MM 2022 · 19 citations
Builds on5
- Content and Style Disentanglement for Artistic Style TransferDmytro Kotovenko, Artsiom Sanakoyeu, Sabine Lang, Björn OmmerICCV 2019 · 187 citations
- Progressive Learning and Disentanglement of Hierarchical RepresentationsZhiyuan Li, Jaideep Vitthal Murkute, Prashnna Kumar Gyawali, Linwei WangICLR 2020 · 47 citations
- StarGAN v2: Diverse Image Synthesis for Multiple DomainsYunjey Choi, Youngjung Uh, Jaejun Yoo, Jung-Woo HaCVPR 2020
- VSGNet: Spatial Attention Network for Detecting Human Object Interactions Using Graph ConvolutionsOytun Ulutan, A. S. M. Iftekhar, B. S. ManjunathCVPR 2020
- Disentangled Image Generation Through Structured Noise InjectionYazeed Alharbi, Peter WonkaCVPR 2020
Related papers
- Attribute-specific Control Units in StyleGAN for Fine-grained Image ManipulationRui Wang, Jian Chen, Gang Yu, Li Sun et al.ACM MM 2021 · 13 citations
- StylePrompter: All Styles Need Is AttentionChenyi Zhuang, Pan Gao, Aljosa SmolicACM MM 2023 · 1 citation
- SemanticStyleGAN: Learning Compositional Generative Priors for Controllable Image Synthesis and EditingYichun Shi, Xiao Yang, Yangyue Wan, Xiaohui ShenCVPR 2022 · 88 citations
- Style Transformer for Image Inversion and EditingXueqi Hu, Qiusheng Huang, Zhengyi Shi, Siyuan Li et al.CVPR 2022 · 58 citations
- Image-to-Image Translation via Hierarchical Style DisentanglementXinyang Li, Shengchuan Zhang, Jie Hu, Liujuan Cao et al.CVPR 2021
