Improving Few-shot Image Generation by Structural Discrimination and Textural Modulation
Mengping Yang, Zhe Wang, Wenyi Feng, Qian Zhang, Ting Xiao
摘要
Few-shot image generation, which aims to produce plausible and diverse images for one category given a few images from this category, has drawn extensive attention. Existing approaches either globally interpolate different images or fuse local representations with pre-defined coefficients. However, such an intuitive combination of images/features only exploits the most relevant information for generation, leading to poor diversity and coarse-grained semantic fusion. To remedy this, this paper proposes a novel textural modulation (TexMod) mechanism to inject external semantic signals into internal local representations. Parameterized by the feedback from the discriminator, our TexMod enables more fined-grained semantic injection while maintaining the synthesis fidelity. Moreover, a global structural discriminator (StructD) is developed to explicitly guide the model to generate images with reasonable layout and outline. Furthermore, the frequency awareness of the model is reinforced by encouraging the model to distinguish frequency signals. Together with these techniques, we build a novel and effective model for few-shot image generation. The effectiveness of our model is identified by extensive experiments on three popular datasets and various settings. Besides achieving state-of-the-art synthesis performance on these datasets, our proposed techniques could be seamlessly integrated into existing models for a further performance boost. Our code and models are available at ://github.com/kobeshegu/SDTM-GAN-ACMMM-2023 here.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Improving the Training of the GANs with Limited Data via Dual Adaptive Noise InjectionZhaoyu Zhang, Yang Hua, Guanxiong Sun, Hui Wang 等ACM MM 2024 · 被引用 3 次
- Improving the Training of Data-Efficient GANs via Quality Aware Dynamic Discriminator Rejection SamplingZhaoyu Zhang, Yang Hua, Guanxiong Sun, Hui Wang 等CVPR 2025
它引用的顶会 Paper25
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- Differentiable Augmentation for Data-Efficient GAN TrainingShengyu Zhao, Zhijian Liu, Ji Lin, Jun-Yan Zhu 等NeurIPS 2020 · 被引用 707 次
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras 等ICCV 2019 · 被引用 668 次
相关 Paper
- Semantic-Aware Generator and Low-level Feature Augmentation for Few-shot Image GenerationZhe Wang, Jiaoyan Guan, Mengping Yang, Ting Xiao 等ACM MM 2023 · 被引用 2 次
- F2GAN: Fusing-and-Filling GAN for Few-shot Image GenerationYan Hong, Li Niu, Jianfu Zhang, Weijie Zhao 等ACM MM 2020 · 被引用 93 次
- LoFGAN: Fusing Local Representations for Few-shot Image GenerationZheng Gu, Wenbin Li, Jing Huo, Lei Wang 等ICCV 2021 · 被引用 64 次
- Exact Fusion via Feature Distribution Matching for Few-Shot Image GenerationYingbo Zhou, Yutong Ye, Pengyu Zhang, Xian Wei 等CVPR 2024 · 被引用 8 次
- Few-shot Image Generation Using Discrete Content RepresentationYan Hong, Li Niu, Jianfu Zhang, Liqing ZhangACM MM 2022 · 被引用 11 次
