SWAGAN: a style-based wavelet-driven generative model
Rinon Gal, Dana Cohen Hochberg, Amit Bermano, Daniel Cohen-Or
摘要
In recent years, considerable progress has been made in the visual quality of Generative Adversarial Networks (GANs). Even so, these networks still suffer from degradation in quality for high-frequency content, stemming from a spectrally biased architecture, and similarly unfavorable loss functions. To address this issue, we present a novel general-purpose Style and WAvelet based GAN (SWAGAN) that implements progressive generation in the frequency domain. SWAGAN incorporates wavelets throughout its generator and discriminator architectures, enforcing a frequency-aware latent representation at every step of the way. This approach, designed to directly tackle the spectral bias of neural networks, yields an improvement in the ability to generate medium and high frequency content, including structures which other networks fail to learn. We demonstrate the advantage of our method by integrating it into the SyleGAN2 framework, and verifying that content generation in the wavelet domain leads to more realistic high-frequency content, even when trained for fewer iterations. Furthermore, we verify that our model's latent space retains the qualities that allow StyleGAN to serve as a basis for a multitude of editing tasks, and show that our frequency-aware approach also induces improved high-frequency performance in downstream tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper28
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- Focal Frequency Loss for Image Reconstruction and SynthesisLiming Jiang, Bo Dai, Wayne Wu, Chen Change LoyICCV 2021 · 被引用 422 次
- StyleSwin: Transformer-based GAN for High-resolution Image GenerationBowen Zhang, Shuyang Gu, Bo Zhang, Jianmin Bao 等CVPR 2022 · 被引用 217 次
- On the Frequency Bias of Generative ModelsKatja Schwarz, Yiyi Liao, Andreas GeigerNeurIPS 2021 · 被引用 117 次
- Aligning Latent and Image Spaces to Connect the UnconnectableIvan Skorokhodov, Grigorii Sotnikov, Mohamed ElhoseinyICCV 2021 · 被引用 99 次
它引用的顶会 Paper11
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- Photorealistic Style Transfer via Wavelet TransformsJaejun Yoo, Youngjung Uh, Sanghyuk Chun, Byeongkyu Kang 等ICCV 2019 · 被引用 412 次
- Fourier Spectrum Discrepancies in Deep Network Generated ImagesTarik Dzanic, Karan Shah, Freddie D. WitherdenNeurIPS 2020 · 被引用 235 次
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt 等ICCV 2019 · 被引用 98 次
相关 Paper
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten 等CVPR 2020
- Content-Aware GAN CompressionYuchen Liu, Zhixin Shu, Yijun Li, Zhe Lin 等CVPR 2021
- AgileGAN: stylizing portraits by inversion-consistent transfer learningGuoxian Song, Linjie Luo, Jing Liu, Wan-Chun Ma 等SIGGRAPH 2021 · 被引用 80 次
- SSD-GAN: Measuring the Realness in the Spatial and Spectral DomainsYuanqi Chen, Ge Li, Cece Jin, Shan Liu 等AAAI 2021 · 被引用 61 次
- Wavelet Knowledge Distillation: Towards Efficient Image-to-Image TranslationLinfeng Zhang, Xin Chen, Xiaobing Tu, Pengfei Wan 等CVPR 2022 · 被引用 105 次
