InfinityGAN: Towards Infinite-Pixel Image Synthesis
Chieh Hubert Lin, Hsin-Ying Lee, Yen-Chi Cheng, Sergey Tulyakov, Ming-Hsuan Yang
摘要
We present a novel framework, InfinityGAN, for arbitrary-sized image generation. The task is associated with several key challenges. First, scaling existing models to an arbitrarily large image size is resource-constrained, in terms of both computation and availability of large-field-of-view training data. InfinityGAN trains and infers in a seamless patch-by-patch manner with low computational resources. Second, large images should be locally and globally consistent, avoid repetitive patterns, and look realistic. To address these, InfinityGAN disentangles global appearances, local structures, and textures. With this formulation, we can generate images with spatial size and level of details not attainable before. Experimental evaluation validates that InfinityGAN generates images with superior realism compared to baselines and features parallelizable inference. Finally, we show several applications unlocked by our approach, such as spatial style fusion, multi-modal outpainting, and image inbetweening. All applications can be operated with arbitrary input and output sizes. Please find the full version of the paper at https://openreview.net/forum?id=ufGMqIM0a4b .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Palette: Image-to-Image Diffusion ModelsChitwan Saharia, William Chan, Huiwen Chang, Chris A. Lee 等SIGGRAPH 2022 · 被引用 1,638 次
- Versatile Diffusion: Text, Images and Variations All in One Diffusion ModelXingqian Xu, Zhangyang Wang, Eric J. Zhang, Kai Wang 等ICCV 2023 · 被引用 265 次
- NUWA-Infinity: Autoregressive over Autoregressive Generation for Infinite Visual SynthesisJian Liang, Chenfei Wu, Xiaowei Hu, Zhe Gan 等NeurIPS 2022 · 被引用 105 次
- InfiniCity: Infinite-Scale City SynthesisChieh Hubert Lin, Hsin-Ying Lee, Willi Menapace, Menglei Chai 等ICCV 2023 · 被引用 86 次
- Single Motion DiffusionSigal Raab, Inbal Leibovitch, Guy Tevet, Moab Arar 等ICLR 2024 · 被引用 81 次
它引用的顶会 Paper22
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen 等ICCV 2019 · 被引用 1,990 次
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 被引用 933 次
- Infinite Nature: Perpetual View Generation of Natural Scenes from a Single ImageAndrew Liu, Ameesh Makadia, Richard Tucker, Noah Snavely 等ICCV 2021 · 被引用 260 次
相关 Paper
- COCO-GAN: Generation by Parts via Conditional CoordinatingChieh Hubert Lin, Chia-Che Chang, Yu-Sheng Chen, Da-Cheng Juan 等ICCV 2019 · 被引用 147 次
- Infinite-Canvas: Higher-Resolution Video Outpainting with Extensive Content GenerationQihua Chen, Yue Ma, Hongfa Wang, Junkun Yuan 等AAAI 2025 · 被引用 4 次
- Infinite-Story: A Training-Free Consistent Text-to-Image GenerationJihun Park, Kyoungmin Lee, Jongmin Gim, Hyeonseo Jo 等AAAI 2026 · 被引用 1 次
- InGAN: Capturing and Retargeting the "DNA" of a Natural ImageAssaf Shocher, Shai Bagon, Phillip Isola, Michal IraniICCV 2019 · 被引用 146 次
- Very Long Natural Scenery Image Prediction by OutpaintingZongxin Yang, Jian Dong, Ping Liu, Yi Yang 等ICCV 2019 · 被引用 97 次
