DLT: Conditioned layout generation with Joint Discrete-Continuous Diffusion Layout Transformer
Elad Levi, Eli Brosh, Mykola Mykhailych, Meir Perez
Abstract
Generating visual layouts is an essential ingredient of graphic design. The ability to condition layout generation on a partial subset of component attributes is critical to real-world applications that involve user interaction. Recently, diffusion models have demonstrated high-quality generative performances in various domains. However, it is unclear how to apply diffusion models to the natural representation of layouts which consists of a mix of discrete (class) and continuous (location, size) attributes. To address the conditioning layout generation problem, we introduce DLT, a joint discrete-continuous diffusion model. DLT is a transformer-based model which has a flexible conditioning mechanism that allows for conditioning on any given subset of all the layout component classes, locations, and sizes. Our method outperforms state-of-the-art generative models on various layout generation datasets with respect to different metrics and conditioning settings. Additionally, we validate the effectiveness of our proposed conditioning mechanism and the joint continuous-diffusion process. This joint process can be incorporated into a wide range of mixed discrete-continuous generative tasks. More information can be found on our project webpage: https://wix-incubator.github.io/DLT
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9c4e664f-44dc-4028-867a-31fd31b46526Cited by top-tier papers6
- Remix-DiT: Mixing Diffusion Transformers for Multi-Expert DenoisingGongfan Fang, Xinyin Ma, Xinchao WangNeurIPS 2024 · 7 citations
- JointDiff: Bridging Continuous and Discrete in Multi-Agent Trajectory GenerationGuillem Capellera, Luis Ferraz, Antonio Romano, Alexandre Alahi et al.ICLR 2026 · 6 citations
- Dimension-free Score Matching and Time Bootstrapping for Diffusion ModelsSyamantak Kumar, Dheeraj Nagaraj, Purnamrita SarkarNeurIPS 2025 · 2 citations
- Agentic Design Review SystemSayan Nag, K. J. Joseph, Koustava Goswami, Vlad I. Morariu et al.AAAI 2026 · 1 citation
- Step-by-step Layered Design GenerationFaizan Farooq Khan, K. J. Joseph, Koustava Goswami, Mohamed Elhoseiny et al.AAAI 2026
Builds on19
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
Related papers
- Towards Aligned Layout Generation via Diffusion Model with Aesthetic ConstraintsJian Chen, Ruiyi Zhang, Yufan Zhou, Changyou ChenICLR 2024 · 32 citations
- LayoutDM: Transformer-based Diffusion Model for Layout GenerationShang Chai, Liansheng Zhuang, Fengying YanCVPR 2023
- LayoutDiffusion: Improving Graphic Layout Generation by Discrete Diffusion Probabilistic ModelsJunyi Zhang, Jiaqi Guo, Shizhao Sun, Jian-Guang Lou et al.ICCV 2023 · 58 citations
- LayoutDM: Discrete Diffusion Model for Controllable Layout GenerationNaoto Inoue, Kotaro Kikuchi, Edgar Simo-Serra, Mayu Otani et al.CVPR 2023
- Infinite-Precision Autoregressive Modeling for Vector Graphics and LayoutsYeonsang Shin, Insoo Kim, Bongkeun Kim, Keonwoo Bae et al.ICML 2026
