DLT: Conditioned layout generation with Joint Discrete-Continuous Diffusion Layout Transformer
Elad Levi, Eli Brosh, Mykola Mykhailych, Meir Perez
摘要
Generating visual layouts is an essential ingredient of graphic design. The ability to condition layout generation on a partial subset of component attributes is critical to real-world applications that involve user interaction. Recently, diffusion models have demonstrated high-quality generative performances in various domains. However, it is unclear how to apply diffusion models to the natural representation of layouts which consists of a mix of discrete (class) and continuous (location, size) attributes. To address the conditioning layout generation problem, we introduce DLT, a joint discrete-continuous diffusion model. DLT is a transformer-based model which has a flexible conditioning mechanism that allows for conditioning on any given subset of all the layout component classes, locations, and sizes. Our method outperforms state-of-the-art generative models on various layout generation datasets with respect to different metrics and conditioning settings. Additionally, we validate the effectiveness of our proposed conditioning mechanism and the joint continuous-diffusion process. This joint process can be incorporated into a wide range of mixed discrete-continuous generative tasks. More information can be found on our project webpage: https://wix-incubator.github.io/DLT
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Remix-DiT: Mixing Diffusion Transformers for Multi-Expert DenoisingGongfan Fang, Xinyin Ma, Xinchao WangNeurIPS 2024 · 被引用 7 次
- JointDiff: Bridging Continuous and Discrete in Multi-Agent Trajectory GenerationGuillem Capellera, Luis Ferraz, Antonio Romano, Alexandre Alahi 等ICLR 2026 · 被引用 6 次
- Dimension-free Score Matching and Time Bootstrapping for Diffusion ModelsSyamantak Kumar, Dheeraj Nagaraj, Purnamrita SarkarNeurIPS 2025 · 被引用 2 次
- Agentic Design Review SystemSayan Nag, K. J. Joseph, Koustava Goswami, Vlad I. Morariu 等AAAI 2026 · 被引用 1 次
- Step-by-step Layered Design GenerationFaizan Farooq Khan, K. J. Joseph, Koustava Goswami, Mohamed Elhoseiny 等AAAI 2026
它引用的顶会 Paper19
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
相关 Paper
- Towards Aligned Layout Generation via Diffusion Model with Aesthetic ConstraintsJian Chen, Ruiyi Zhang, Yufan Zhou, Changyou ChenICLR 2024 · 被引用 32 次
- LayoutDM: Transformer-based Diffusion Model for Layout GenerationShang Chai, Liansheng Zhuang, Fengying YanCVPR 2023
- LayoutDiffusion: Improving Graphic Layout Generation by Discrete Diffusion Probabilistic ModelsJunyi Zhang, Jiaqi Guo, Shizhao Sun, Jian-Guang Lou 等ICCV 2023 · 被引用 58 次
- LayoutDM: Discrete Diffusion Model for Controllable Layout GenerationNaoto Inoue, Kotaro Kikuchi, Edgar Simo-Serra, Mayu Otani 等CVPR 2023
- Infinite-Precision Autoregressive Modeling for Vector Graphics and LayoutsYeonsang Shin, Insoo Kim, Bongkeun Kim, Keonwoo Bae 等ICML 2026
