Arbitrary-Scale Image Synthesis
Evangelos Ntavelis, Mohamad Shahbazi, Iason Kastanis, Radu Timofte, Martin Danelljan, Luc Van Gool
Abstract
Positional encodings have enabled recent works to train a single adversarial network that can generate images of different scales. However, these approaches are either limited to a set of discrete scales or struggle to maintain good perceptual quality at the scales for which the model is not trained explicitly. We propose the design of scale-consistent positional encodings invariant to our generator's layers transformations. This enables the generation of arbitrary-scale images even at scales unseen during training. Moreover, we incorporate novel inter-scale augmentations into our pipeline and partial generation training to facilitate the synthesis of consistent images at arbitrary scales. Lastly, we show competitive results for a continuum of scales on various commonly used datasets for image synthesis.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- PixNerd: Pixel Neural Field DiffusionShuai Wang, Ziteng Gao, Chenhui Zhu, Weilin Huang et al.ICLR 2026 · 78 citations
- Continuous Field Reconstruction from Sparse Observations with Implicit Neural NetworksXihaier Luo, Wei Xu, Balu Nadiga, Yihui Ren et al.ICLR 2024 · 23 citations
- DDMI: Domain-agnostic Latent Diffusion Models for Synthesizing High-Quality Implicit Neural RepresentationsDogyun Park, Sihyeon Kim, Sojin Lee, Hyunwoo J. KimICLR 2024 · 19 citations
- UnitedHuman: Harnessing Multi-Source Data for High-Resolution Human GenerationJianglin Fu, Shikai Li, Yuming Jiang, Kwan-Yee Lin et al.ICCV 2023 · 18 citations
- Probabilistic Precision and Recall Towards Reliable Evaluation of Generative ModelsDogyun Park, Suhyun KimICCV 2023 · 12 citations
Builds on16
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or et al.ICCV 2021 · 1,437 citations
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 1,195 citations
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 421 citations
- You Only Need Adversarial Supervision for Semantic Image SynthesisEdgar Schönfeld, Vadim Sushko, Dan Zhang, Juergen Gall et al.ICLR 2021 · 219 citations
Related papers
- Toward Spatially Unbiased Generative ModelsJooyoung Choi, Jungbeom Lee, Yonghyun Jeong, Sungroh YoonICCV 2021 · 17 citations
- Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional EncodingsLiang Hou, Cong Liu, Mingwu Zheng, Xin Tao et al.AAAI 2026 · 2 citations
- Learning Images Across Scales Using Adversarial TrainingKrzysztof Wolski, Adarsh Djeacoumar, Alireza Javanmardi, Hans-Peter Seidel et al.SIGGRAPH 2024 · 2 citations
- Efficient Scale-Invariant Generator with Column-Row Entangled Pixel SynthesisThuan Hoang Nguyen, Thanh Van Le, Anh TranCVPR 2023
- OPE-SR: Orthogonal Position Encoding for Designing a Parameter-free Upsampling Module in Arbitrary-scale Image Super-ResolutionGaochao Song, Qian Sun, Luo Zhang, Ran Su et al.CVPR 2023
