PetsGAN: Rethinking Priors for Single Image Generation
Zicheng Zhang, Yinglu Liu, Congying Han, Hailin Shi, Tiande Guo, Bowen Zhou
Abstract
Single image generation (SIG), described as generating diverse samples that have the same visual content as the given natural image, is first introduced by SinGAN, which builds a pyramid of GANs to progressively learn the internal patch distribution of the single image. It shows excellent performance in a wide range of image manipulation tasks. However, SinGAN has some limitations. Firstly, due to lack of semantic information, SinGAN cannot handle the object images well as it does on the scene and texture images. Secondly, the independent progressive training scheme is time-consuming and easy to cause artifacts accumulation. To tackle these problems, in this paper, we dig into the single image generation problem and improve SinGAN by fully-utilization of internal and external priors. The main contributions of this paper include: 1) We interpret single image generation from the perspective of the general generative task, that is, to learn a diverse distribution from the Dirac distribution composed of a single image. In order to solve this non-trivial problem, we construct a regularized latent variable model to formulate SIG. To the best of our knowledge, it is the first time to give a clear formulation and optimization goal of SIG, and all the existing methods for SIG can be regarded as special cases of this model. 2) We design a novel Prior-based end-to-end training GAN (PetsGAN), which is infused with internal prior and external prior to overcome the problems of SinGAN. For one thing, we employ the pre-trained GAN model to inject external prior for image generation, which can alleviate the problem of lack of semantic information and generate natural, reasonable and diverse samples, even for the object image. For another, we fully-utilize the internal prior by a differential Patch Matching module and an effective reconstruction network to generate consistent and realistic texture. 3) We construct abundant of qualitative and quantitative experiments on three datasets. The experimental results show our method surpasses other methods on both generated image quality, diversity, and training speed. Moreover, we apply our method to other image manipulation tasks (e.g., style transfer, harmonization) and the results further prove the effectiveness and efficiency of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Single Motion DiffusionSigal Raab, Inbal Leibovitch, Guy Tevet, Moab Arar et al.ICLR 2024 · 81 citations
- ChatEdit: Towards Multi-turn Interactive Facial Image Editing via DialogueXing Cui, Zekun Li, Pei Li, Yibo Hu et al.EMNLP 2023 · 4 citations
Builds on9
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 933 citations
- InGAN: Capturing and Retargeting the "DNA" of a Natural ImageAssaf Shocher, Shai Bagon, Phillip Isola, Michal IraniICCV 2019 · 146 citations
- Hierarchical Patch VAE-GAN: Generating Diverse Videos from a Single SampleShir Gur, Sagie Benaim, Lior WolfNeurIPS 2020 · 84 citations
- SinIR: Efficient General Image Manipulation with Single Image ReconstructionJihyeong Yoo, Qifeng ChenICML 2021 · 25 citations
- Image Processing Using Multi-Code GAN PriorJinjin Gu, Yujun Shen, Bolei ZhouCVPR 2020
Related papers
- Patchwise Generative ConvNet: Training Energy-Based Models From a Single Natural Image for Internal LearningZilong Zheng, Jianwen Xie, Ping LiCVPR 2021
- Hiding Images in Deep Probabilistic ModelsHaoyu Chen, Linqi Song, Zhenxing Qian, Xinpeng Zhang et al.NeurIPS 2022 · 20 citations
- Semi-Supervised Single-Stage Controllable GANs for Conditional Fine-Grained Image GenerationTianyi Chen, Yi Liu, Yunfei Zhang, Si Wu et al.ICCV 2021 · 11 citations
- Semantic Image Analogy with a Conditional Single-Image GANJiacheng Li, Zhiwei Xiong, Dong Liu, Xuejin Chen et al.ACM MM 2020 · 4 citations
- SinGRAF: Learning a 3D Generative Radiance Field for a Single SceneMinjung Son, Jeong Joon Park, Leonidas J. Guibas, Gordon WetzsteinCVPR 2023
