Drop the GAN: In Defense of Patches Nearest Neighbors as Single Image Generative Models
Niv Granot, Ben Feinstein, Assaf Shocher, Shai Bagon, Michal Irani
摘要
Image manipulation dates back long before the deep learning era. The classical prevailing approaches were based on maximizing patch similarity between the input and generated output. Recently, single-image GANs were introduced as a superior and more sophisticated solution to image manipulation tasks. Moreover, they offered the opportunity not only to manipulate a given image, but also to generate a large and diverse set of different outputs from a single natural image. This gave rise to new tasks, which are considered “GAN-only”. However, despite their impressiveness, single-image GANs require long training time (usually hours) for each image and each task and often suffer from visual artifacts. In this paper we revisit the classical patch-based methods, and show that - unlike previously believed - classical methods can be adapted to tackle these novel “GAN-only” tasks. Moreover, they do so better and faster than single-image GAN-based methods. More specifically, we show that: (i) by introducing slight modifications, classical patch-based methods are able to unconditionally generate diverse images based on a single natural image; (ii) the generated output visual quality exceeds that of single-image GANs by a large margin (confirmed both quantitatively and qualitatively); (iii) they are orders of magnitude faster (runtime reduced from hours to seconds). <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2</sup> <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2</sup> This project received funding from the European Research Council (ERC) under the European Union's Horizon 2020 research and innovation programme (grant agreement No 788535), and the Carolito Stiftung. Dr Bagon is a Robin Chemers Neustein AI Fellow.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- SinDDM: A Single Image Denoising Diffusion ModelVladimir Kulikov, Shahar Yadin, Matan Kleiner, Tomer MichaeliICML 2023 · 被引用 113 次
- SinFusion: Training Diffusion Models on a Single Image or VideoYaniv Nikankin, Niv Haim, Michal IraniICML 2023 · 被引用 83 次
- Single Motion DiffusionSigal Raab, Inbal Leibovitch, Guy Tevet, Moab Arar 等ICLR 2024 · 被引用 81 次
- Diffusion-based Image Translation using disentangled style and content representationGihyun Kwon, Jong Chul YeICLR 2023 · 被引用 46 次
- Sin3DM: Learning a Diffusion Model from a Single 3D Textured ShapeRundi Wu, Ruoshi Liu, Carl Vondrick, Changxi ZhengICLR 2024 · 被引用 34 次
它引用的顶会 Paper4
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 被引用 933 次
- Hierarchical Patch VAE-GAN: Generating Diverse Videos from a Single SampleShir Gur, Sagie Benaim, Lior WolfNeurIPS 2020 · 被引用 84 次
- An Internal Learning Approach to Video InpaintingHaotian Zhang, Long Mai, Hailin Jin, Zhaowen Wang 等ICCV 2019 · 被引用 77 次
- Positional Encoding As Spatial Inductive Bias in GANsRui Xu, Xintao Wang, Kai Chen, Bolei Zhou 等CVPR 2021
相关 Paper
- InGAN: Capturing and Retargeting the "DNA" of a Natural ImageAssaf Shocher, Shai Bagon, Phillip Isola, Michal IraniICCV 2019 · 被引用 146 次
- SinIR: Efficient General Image Manipulation with Single Image ReconstructionJihyeong Yoo, Qifeng ChenICML 2021 · 被引用 25 次
- PetsGAN: Rethinking Priors for Single Image GenerationZicheng Zhang, Yinglu Liu, Congying Han, Hailin Shi 等AAAI 2022 · 被引用 27 次
- Scaling up GANs for Text-to-Image SynthesisMinguk Kang, Jun-Yan Zhu, Richard Zhang, Jaesik Park 等CVPR 2023
- Attack Deterministic Conditional Image Generative Models for Diverse and Controllable GenerationTianyi Chu, Wei Xing, Jiafu Chen, Zhizhong Wang 等AAAI 2024 · 被引用 3 次
