Patched Denoising Diffusion Models For High-Resolution Image Synthesis
Zheng Ding, Mengqi Zhang, Jiajun Wu, Zhuowen Tu
摘要
We propose an effective denoising diffusion model for generating high-resolution images (e.g., 1024512), trained on small-size image patches (e.g., 6464). We name our algorithm Patch-DM, in which a new feature collage strategy is designed to avoid the boundary artifact when synthesizing large-size images. Feature collage systematically crops and combines partial features of the neighboring patches to predict the features of a shifted image patch, allowing the seamless generation of the entire image due to the overlap in the patch feature space. Patch-DM produces high-quality image synthesis results on our newly collected dataset of nature images (1024512), as well as on standard benchmarks of smaller sizes (256256), including LSUN-Bedroom, LSUN-Church, and FFHQ. We compare our method with previous patch-based generation methods and achieve state-of-the-art FID scores on all four datasets. Further, Patch-DM also reduces memory complexity compared to the classic diffusion models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Patch Diffusion: Faster and More Data-Efficient Training of Diffusion ModelsZhendong Wang, Yifan Jiang, Huangjie Zheng, Peihao Wang 等NeurIPS 2023 · 被引用 205 次
- DiP: Taming Diffusion Models in Pixel SpaceZhennan Chen, Junwei Zhu, Xu Chen, Jiangning Zhang 等CVPR 2026 · 被引用 46 次
- Learning Image Priors Through Patch-Based Diffusion Models for Solving Inverse ProblemsJason Hu, Bowen Song, Xiaojian Xu, Liyue Shen 等NeurIPS 2024 · 被引用 32 次
- Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion ModelingHuangjie Zheng, Zhendong Wang, Jianbo Yuan, Guanghan Ning 等ICLR 2024 · 被引用 17 次
- Is One GPU Enough? Pushing Image Generation at Higher-Resolutions with Foundation ModelsAthanasios Tragakis, Marco Aversa, Chaitanya Kaul, Roderick Murray-Smith 等NeurIPS 2024 · 被引用 16 次
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- Fixed Point Diffusion ModelsXingjian Bai, Luke Melas-KyriaziCVPR 2024
- Relay Diffusion: Unifying diffusion process across resolutions for image synthesisJiayan Teng, Wendi Zheng, Ming Ding, Wenyi Hong 等ICLR 2024 · 被引用 84 次
- Factorized Diffusion Architectures for Unsupervised Image Generation and SegmentationXin Yuan, Michael MaireNeurIPS 2024 · 被引用 4 次
- ElasticDiffusion: Training-Free Arbitrary Size Image Generation Through Global-Local Content SeparationMoayed Haji-Ali, Guha Balakrishnan, Vicente OrdonezCVPR 2024
- Diffusion Model Patching via Mixture-of-PromptsSeokil Ham, Sangmin Woo, Jin-Young Kim, Hyojun Go 等AAAI 2025 · 被引用 9 次
