Diagonal Attention and Style-based GAN for Content-Style Disentanglement in Image Generation and Translation
Gihyun Kwon, Jong Chul Ye
摘要
One of the important research topics in image generative models is to disentangle the spatial contents and styles for their separate control. Although StyleGAN can generate content feature vectors from random noises, the resulting spatial content control is primarily intended for minor spatial variations, and the disentanglement of global content and styles is by no means complete. Inspired by a mathematical understanding of normalization and attention, here we present a novel hierarchical adaptive Diagonal spatial ATtention (DAT) layers to separately manipulate the spatial contents from styles in a hierarchical manner. Using DAT and AdaIN, our method enables coarse-to-fine level disentanglement of spatial contents and styles. In addition, our generator can be easily integrated into the GAN inversion framework so that the content and style of translated images from multi-domain image translation tasks can be flexibly controlled. By using various datasets, we confirm that the proposed method not only outperforms the existing models in disentanglement scores, but also provides more flexible control over spatial features in the generated images.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- StyleDiffusion: Controllable Disentangled Style Transfer via Diffusion ModelsZhizhong Wang, Lei Zhao, Wei XingICCV 2023 · 被引用 219 次
- Pastiche Master: Exemplar-Based High-Resolution Portrait Style TransferShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyCVPR 2022 · 被引用 130 次
- TransEditor: Transformer-Based Dual-Space GAN for Highly Controllable Facial EditingYanbo Xu, Yueqin Yin, Liming Jiang, Qianyi Wu 等CVPR 2022 · 被引用 53 次
- Style Equalization: Unsupervised Learning of Controllable Generative Sequence ModelsJen-Hao Rick Chang, Ashish Shrivastava, Hema Koppula, Xiaoshuai Zhang 等ICML 2022 · 被引用 21 次
- Exploring Negatives in Contrastive Learning for Unpaired Image-to-Image TranslationYupei Lin, Sen Zhang, Tianshui Chen, Yongyi Lu 等ACM MM 2022 · 被引用 19 次
它引用的顶会 Paper5
- Content and Style Disentanglement for Artistic Style TransferDmytro Kotovenko, Artsiom Sanakoyeu, Sabine Lang, Björn OmmerICCV 2019 · 被引用 187 次
- Progressive Learning and Disentanglement of Hierarchical RepresentationsZhiyuan Li, Jaideep Vitthal Murkute, Prashnna Kumar Gyawali, Linwei WangICLR 2020 · 被引用 47 次
- StarGAN v2: Diverse Image Synthesis for Multiple DomainsYunjey Choi, Youngjung Uh, Jaejun Yoo, Jung-Woo HaCVPR 2020
- VSGNet: Spatial Attention Network for Detecting Human Object Interactions Using Graph ConvolutionsOytun Ulutan, A. S. M. Iftekhar, B. S. ManjunathCVPR 2020
- Disentangled Image Generation Through Structured Noise InjectionYazeed Alharbi, Peter WonkaCVPR 2020
相关 Paper
- Attribute-specific Control Units in StyleGAN for Fine-grained Image ManipulationRui Wang, Jian Chen, Gang Yu, Li Sun 等ACM MM 2021 · 被引用 13 次
- StylePrompter: All Styles Need Is AttentionChenyi Zhuang, Pan Gao, Aljosa SmolicACM MM 2023 · 被引用 1 次
- SemanticStyleGAN: Learning Compositional Generative Priors for Controllable Image Synthesis and EditingYichun Shi, Xiao Yang, Yangyue Wan, Xiaohui ShenCVPR 2022 · 被引用 88 次
- Style Transformer for Image Inversion and EditingXueqi Hu, Qiusheng Huang, Zhengyi Shi, Siyuan Li 等CVPR 2022 · 被引用 58 次
- Image-to-Image Translation via Hierarchical Style DisentanglementXinyang Li, Shengchuan Zhang, Jie Hu, Liujuan Cao 等CVPR 2021
