Learning Images Across Scales Using Adversarial Training
Krzysztof Wolski, Adarsh Djeacoumar, Alireza Javanmardi, Hans-Peter Seidel, Christian Theobalt, Guillaume Cordonnier, Karol Myszkowski, George Drettakis, Xingang Pan, Thomas Leimkühler
Abstract
The real world exhibits rich structure and detail across many scales of observation. It is difficult, however, to capture and represent a broad spectrum of scales using ordinary images. We devise a novel paradigm for learning a representation that captures an orders-of-magnitude variety of scales from an unstructured collection of ordinary images. We treat this collection as a distribution of scale-space slices to be learned using adversarial training, and additionally enforce coherency across slices. Our approach relies on a multiscale generator with carefully injected procedural frequency content, which allows to interactively explore the emerging continuous scale space. Training across vastly different scales poses challenges regarding stability, which we tackle using a supervision scheme that involves careful sampling of scales. We show that our generator can be used as a multiscale generative model, and for reconstructions of scale spaces from unstructured patches. Significantly outperforming the state of the art, we demonstrate zoom-in factors of up to 256x at high quality and scale consistency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference AlignmentBryan Sangwoo Kim, Jeongsol Kim, Jong Chul YeNeurIPS 2025 · 10 citations
- Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image SynthesisYu Yuan, Xijun Wang, Yichen Sheng, Prateek Chennuri et al.CVPR 2025
Builds on27
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 933 citations
- StyleGAN-XL: Scaling StyleGAN to Large Diverse DatasetsAxel Sauer, Katja Schwarz, Andreas GeigerSIGGRAPH 2022 · 326 citations
- Learning temporal coherence via self-supervision for GAN-based video generationMengyu Chu, You Xie, Jonas Mayer, Laura Leal-Taixé et al.SIGGRAPH 2020 · 198 citations
Related papers
- Arbitrary-Scale Image SynthesisEvangelos Ntavelis, Mohamad Shahbazi, Iason Kastanis, Radu Timofte et al.CVPR 2022 · 17 citations
- Generative Powers of TenXiaojuan Wang, Janne Kontkanen, Brian Curless, Steven M. Seitz et al.CVPR 2024 · 3 citations
- WonderZoom: Multi-Scale 3D World GenerationJin Cao, Hong-Xing Yu, Jiajun WuCVPR 2026 · 1 citation
- MSG-GAN: Multi-Scale Gradients for Generative Adversarial NetworksAnimesh Karnewar, Oliver WangCVPR 2020
- Forward Super-Resolution: How Can GANs Learn Hierarchical Generative Models for Real-World DistributionsZeyuan Allen-Zhu, Yuanzhi LiICLR 2023 · 4 citations
