Patched Denoising Diffusion Models For High-Resolution Image Synthesis
Zheng Ding, Mengqi Zhang, Jiajun Wu, Zhuowen Tu
Abstract
We propose an effective denoising diffusion model for generating high-resolution images (e.g., 1024512), trained on small-size image patches (e.g., 6464). We name our algorithm Patch-DM, in which a new feature collage strategy is designed to avoid the boundary artifact when synthesizing large-size images. Feature collage systematically crops and combines partial features of the neighboring patches to predict the features of a shifted image patch, allowing the seamless generation of the entire image due to the overlap in the patch feature space. Patch-DM produces high-quality image synthesis results on our newly collected dataset of nature images (1024512), as well as on standard benchmarks of smaller sizes (256256), including LSUN-Bedroom, LSUN-Church, and FFHQ. We compare our method with previous patch-based generation methods and achieve state-of-the-art FID scores on all four datasets. Further, Patch-DM also reduces memory complexity compared to the classic diffusion models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc5109ec-aed0-4478-96ac-aff0c8f0e326Cited by top-tier papers19
- Patch Diffusion: Faster and More Data-Efficient Training of Diffusion ModelsZhendong Wang, Yifan Jiang, Huangjie Zheng, Peihao Wang et al.NeurIPS 2023 · 205 citations
- DiP: Taming Diffusion Models in Pixel SpaceZhennan Chen, Junwei Zhu, Xu Chen, Jiangning Zhang et al.CVPR 2026 · 46 citations
- Learning Image Priors Through Patch-Based Diffusion Models for Solving Inverse ProblemsJason Hu, Bowen Song, Xiaojian Xu, Liyue Shen et al.NeurIPS 2024 · 32 citations
- Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion ModelingHuangjie Zheng, Zhendong Wang, Jianbo Yuan, Guanghan Ning et al.ICLR 2024 · 17 citations
- Is One GPU Enough? Pushing Image Generation at Higher-Resolutions with Foundation ModelsAthanasios Tragakis, Marco Aversa, Chaitanya Kaul, Roderick Murray-Smith et al.NeurIPS 2024 · 16 citations
Builds on20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Fixed Point Diffusion ModelsXingjian Bai, Luke Melas-KyriaziCVPR 2024
- Relay Diffusion: Unifying diffusion process across resolutions for image synthesisJiayan Teng, Wendi Zheng, Ming Ding, Wenyi Hong et al.ICLR 2024 · 84 citations
- Factorized Diffusion Architectures for Unsupervised Image Generation and SegmentationXin Yuan, Michael MaireNeurIPS 2024 · 4 citations
- ElasticDiffusion: Training-Free Arbitrary Size Image Generation Through Global-Local Content SeparationMoayed Haji-Ali, Guha Balakrishnan, Vicente OrdonezCVPR 2024
- Diffusion Model Patching via Mixture-of-PromptsSeokil Ham, Sangmin Woo, Jin-Young Kim, Hyojun Go et al.AAAI 2025 · 9 citations
