Drop the GAN: In Defense of Patches Nearest Neighbors as Single Image Generative Models
Niv Granot, Ben Feinstein, Assaf Shocher, Shai Bagon, Michal Irani
Abstract
Image manipulation dates back long before the deep learning era. The classical prevailing approaches were based on maximizing patch similarity between the input and generated output. Recently, single-image GANs were introduced as a superior and more sophisticated solution to image manipulation tasks. Moreover, they offered the opportunity not only to manipulate a given image, but also to generate a large and diverse set of different outputs from a single natural image. This gave rise to new tasks, which are considered “GAN-only”. However, despite their impressiveness, single-image GANs require long training time (usually hours) for each image and each task and often suffer from visual artifacts. In this paper we revisit the classical patch-based methods, and show that - unlike previously believed - classical methods can be adapted to tackle these novel “GAN-only” tasks. Moreover, they do so better and faster than single-image GAN-based methods. More specifically, we show that: (i) by introducing slight modifications, classical patch-based methods are able to unconditionally generate diverse images based on a single natural image; (ii) the generated output visual quality exceeds that of single-image GANs by a large margin (confirmed both quantitatively and qualitatively); (iii) they are orders of magnitude faster (runtime reduced from hours to seconds). <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2</sup> <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2</sup> This project received funding from the European Research Council (ERC) under the European Union's Horizon 2020 research and innovation programme (grant agreement No 788535), and the Carolito Stiftung. Dr Bagon is a Robin Chemers Neustein AI Fellow.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fbebadb4-1676-4818-9526-74715d8ba83cCited by top-tier papers21
- SinDDM: A Single Image Denoising Diffusion ModelVladimir Kulikov, Shahar Yadin, Matan Kleiner, Tomer MichaeliICML 2023 · 113 citations
- SinFusion: Training Diffusion Models on a Single Image or VideoYaniv Nikankin, Niv Haim, Michal IraniICML 2023 · 83 citations
- Single Motion DiffusionSigal Raab, Inbal Leibovitch, Guy Tevet, Moab Arar et al.ICLR 2024 · 81 citations
- Diffusion-based Image Translation using disentangled style and content representationGihyun Kwon, Jong Chul YeICLR 2023 · 46 citations
- Sin3DM: Learning a Diffusion Model from a Single 3D Textured ShapeRundi Wu, Ruoshi Liu, Carl Vondrick, Changxi ZhengICLR 2024 · 34 citations
Builds on4
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 933 citations
- Hierarchical Patch VAE-GAN: Generating Diverse Videos from a Single SampleShir Gur, Sagie Benaim, Lior WolfNeurIPS 2020 · 84 citations
- An Internal Learning Approach to Video InpaintingHaotian Zhang, Long Mai, Hailin Jin, Zhaowen Wang et al.ICCV 2019 · 77 citations
- Positional Encoding As Spatial Inductive Bias in GANsRui Xu, Xintao Wang, Kai Chen, Bolei Zhou et al.CVPR 2021
Related papers
- InGAN: Capturing and Retargeting the "DNA" of a Natural ImageAssaf Shocher, Shai Bagon, Phillip Isola, Michal IraniICCV 2019 · 146 citations
- SinIR: Efficient General Image Manipulation with Single Image ReconstructionJihyeong Yoo, Qifeng ChenICML 2021 · 25 citations
- PetsGAN: Rethinking Priors for Single Image GenerationZicheng Zhang, Yinglu Liu, Congying Han, Hailin Shi et al.AAAI 2022 · 27 citations
- Scaling up GANs for Text-to-Image SynthesisMinguk Kang, Jun-Yan Zhu, Richard Zhang, Jaesik Park et al.CVPR 2023
- Attack Deterministic Conditional Image Generative Models for Diverse and Controllable GenerationTianyi Chu, Wei Xing, Jiafu Chen, Zhizhong Wang et al.AAAI 2024 · 3 citations
