Unifying Layout Generation with a Decoupled Diffusion Model
Mude Hui, Zhizheng Zhang, Xiaoyi Zhang, Wenxuan Xie, Yuwang Wang, Yan Lu
摘要
Layout generation aims to synthesize realistic graphic scenes consisting of elements with different attributes including category, size, position, and between-element relation. It is a crucial task for reducing the burden on heavyduty graphic design works for formatted scenes, e.g., publications, documents, and user interfaces (UIs). Diverse application scenarios impose a big challenge in unifying various layout generation subtasks, including conditional and unconditional generation. In this paper, we propose a Layout Diffusion Generative Model (LDGM) to achieve such unification with a single decoupled diffusion model. LDGM views a layout of arbitrary missing or coarse element attributes as an intermediate diffusion status from a completed layout. Since different attributes have their individual semantics and characteristics, we propose to decouple the diffusion processes for them to improve the diversity of training samples and learn the reverse process jointly to exploit global-scope contexts for facilitating generation. As a result, our LDGM can generate layouts either from scratch or conditional on arbitrary available attributes. Extensive qualitative and quantitative experiments demonstrate our proposed LDGM outperforms existing layout generation models in both functionality and performance. * This work was done when Mude Hui was an intern at MSRA.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- LayoutPrompter: Awaken the Design Ability of Large Language ModelsJiawei Lin, Jiaqi Guo, Shizhao Sun, Zijiang Yang 等NeurIPS 2023 · 被引用 71 次
- Towards Aligned Layout Generation via Diffusion Model with Aesthetic ConstraintsJian Chen, Ruiyi Zhang, Yufan Zhou, Changyou ChenICLR 2024 · 被引用 32 次
- LayoutNUWA: Revealing the Hidden Layout Expertise of Large Language ModelsZecheng Tang, Chenfei Wu, Juntao Li, Nan DuanICLR 2024 · 被引用 25 次
- Visual Layout Composer: Image-Vector Dual Diffusion Model for Design Layout GenerationMohammad Amin Shabani, Zhaowen Wang, Difan Liu, Nanxuan Zhao 等CVPR 2024 · 被引用 7 次
- PosterVerse: A Full-Workflow Framework for Commercial-Grade Poster Generation with HTML-Based Scalable TypographyJunle Liu, Peirong Zhang, Yuyi Zhang, Pengyu Yan 等AAAI 2026 · 被引用 3 次
它引用的顶会 Paper13
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
- DiffWave: A Versatile Diffusion Model for Audio SynthesisZhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao 等ICLR 2021 · 被引用 1,902 次
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang 等NeurIPS 2022 · 被引用 1,546 次
相关 Paper
- LayoutDiffusion: Improving Graphic Layout Generation by Discrete Diffusion Probabilistic ModelsJunyi Zhang, Jiaqi Guo, Shizhao Sun, Jian-Guang Lou 等ICCV 2023 · 被引用 58 次
- LayoutDM: Discrete Diffusion Model for Controllable Layout GenerationNaoto Inoue, Kotaro Kikuchi, Edgar Simo-Serra, Mayu Otani 等CVPR 2023
- DLT: Conditioned layout generation with Joint Discrete-Continuous Diffusion Layout TransformerElad Levi, Eli Brosh, Mykola Mykhailych, Meir PerezICCV 2023 · 被引用 28 次
- LayoutDM: Transformer-based Diffusion Model for Layout GenerationShang Chai, Liansheng Zhuang, Fengying YanCVPR 2023
- PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language ModelsRunze He, Bo Cheng, Yuhang Ma, Qingxiang Jia 等ICCV 2025 · 被引用 1 次
