Aligning Latent and Image Spaces to Connect the Unconnectable
Ivan Skorokhodov, Grigorii Sotnikov, Mohamed Elhoseiny
摘要
In this work, we develop a method to generate infinite high-resolution images with diverse and complex content. It is based on a perfectly equivariant patch-wise generator with synchronous interpolations in the image and latent spaces. Latent codes, when sampled, are positioned on the coordinate grid, and each pixel is computed from an interpolation of the neighboring codes. We modify the AdaIN mechanism to work in such a setup and train a GAN model to generate images positioned between any two latent vectors. At test time, this allows for generating infinitely large images of diverse scenes that transition naturally from one into another. Apart from that, we introduce LHQ: a new dataset of 90k high-resolution nature landscapes. We test the approach on LHQ, LSUN Tower and LSUN Bridge and outperform the baselines by at least 4 times in terms of quality and diversity of the produced infinite images. The project website is located at https://universome.github.io/alis.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper35
- Drag Your GAN: Interactive Point-based Manipulation on the Generative Image ManifoldXingang Pan, Ayush Tewari, Thomas Leimkühler, Lingjie Liu 等SIGGRAPH 2023 · 被引用 206 次
- StyleGAN-V: A Continuous Video Generator with the Price, Image Quality and Perks of StyleGAN2Ivan Skorokhodov, Sergey Tulyakov, Mohamed ElhoseinyCVPR 2022 · 被引用 167 次
- EpiGRAF: Rethinking training of 3D GANsIvan Skorokhodov, Sergey Tulyakov, Yiqun Wang, Peter WonkaNeurIPS 2022 · 被引用 145 次
- Frido: Feature Pyramid Diffusion for Complex Scene Image SynthesisWan-Cyuan Fan, Yen-Chun Chen, Dongdong Chen, Yu Cheng 等AAAI 2023 · 被引用 118 次
- NUWA-Infinity: Autoregressive over Autoregressive Generation for Infinite Visual SynthesisJian Liang, Chenfei Wu, Xiaowei Hu, Zhe Gan 等NeurIPS 2022 · 被引用 105 次
它引用的顶会 Paper27
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 被引用 933 次
相关 Paper
- InfinityGAN: Towards Infinite-Pixel Image SynthesisChieh Hubert Lin, Hsin-Ying Lee, Yen-Chi Cheng, Sergey Tulyakov 等ICLR 2022 · 被引用 84 次
- LT3SD: Latent Trees for 3D Scene DiffusionQuan Meng, Lei Li, Matthias Nießner, Angela DaiCVPR 2025
- Progressive Semantic-Aware Style Transformation for Blind Face RestorationChaofeng Chen, Xiaoming Li, Lingbo Yang, Xianhui Lin 等CVPR 2021
- High-resolution Face Swapping via Latent Semantics DisentanglementYangyang Xu, Bailin Deng, Junle Wang, Yanqing Jing 等CVPR 2022 · 被引用 93 次
- Patched Denoising Diffusion Models For High-Resolution Image SynthesisZheng Ding, Mengqi Zhang, Jiajun Wu, Zhuowen TuICLR 2024 · 被引用 55 次
