Spot the Error: Non-autoregressive Graphic Layout Generation with Wireframe Locator
Jieru Lin, Danqing Huang, Tiejun Zhao, Dechen Zhan, Chin-Yew Lin
Abstract
Layout generation is a critical step in graphic design to achieve meaningful compositions of elements. Most previous works view it as a sequence generation problem by concatenating element attribute tokens (i.e., category, size, position). So far the autoregressive approach (AR) has achieved promising results, but is still limited in global context modeling and suffers from error propagation since it can only attend to the previously generated tokens. Recent non-autoregressive attempts (NAR) have shown competitive results, which provides a wider context range and the flexibility to refine with iterative decoding. However, current works only use simple heuristics to recognize erroneous tokens for refinement which is inaccurate. This paper first conducts an in-depth analysis to better understand the difference between the AR and NAR framework. Furthermore, based on our observation that pixel space is more sensitive in capturing spatial patterns of graphic layouts (e.g., overlap, alignment), we propose a learning-based locator to detect erroneous tokens which takes the wireframe image rendered from the generated layout sequence as input. We show that it serves as a complementary modality to the element sequence in object space and contributes greatly to the overall performance. Experiments on two public datasets show that our approach outperforms both AR and NAR baselines. Extensive studies further prove the effectiveness of different modules with interesting findings. Our code will be available at https://github.com/ffffatgoose/SpotError.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d86cb6a5-cc91-44a4-b239-06ea51808c7bCited by top-tier papers3
- Desigen: A Pipeline for Controllable Design Template GenerationHaohan Weng, Danqing Huang, Yu Qiao, Zheng Hu et al.CVPR 2024 · 10 citations
- Dancing With Chains: Ideating Under Constraints With UIDEC in UI/UX DesignAtefeh Shokrizadeh, Boniface Bahati Tadjuidje, Shivam Kumar, Sohan Kamble et al.CHI 2025 · 6 citations
- BannerAgency: Advertising Banner Design with Multimodal LLM AgentsHeng Wang, Yotaro Shimose, Shingo TakamatsuEMNLP 2025 · 2 citations
Builds on13
- LayoutVAE: Stochastic Scene Layout Generation From a Label SetAkash Abdu Jyothi, Thibaut Durand, Jiawei He, Leonid Sigal et al.ICCV 2019 · 194 citations
- Improving Non-Autoregressive Translation Models Without DistillationXiao Shi Huang, Felipe Pérez, Maksims VolkovsICLR 2022 · 60 citations
- LayoutDiffusion: Improving Graphic Layout Generation by Discrete Diffusion Probabilistic ModelsJunyi Zhang, Jiaqi Guo, Shizhao Sun, Jian-Guang Lou et al.ICCV 2023 · 58 citations
- Coarse-to-Fine Generative Modeling for Graphic LayoutsZhaoyun Jiang, Shizhao Sun, Jihua Zhu, Jian-Guang Lou et al.AAAI 2022 · 54 citations
- Geometry Aligned Variational Transformer for Image-conditioned Layout GenerationYunning Cao, Ye Ma, Min Zhou, Chuanbin Liu et al.ACM MM 2022 · 37 citations
Related papers
- Layout Generation as Intermediate Action Sequence PredictionHuiting Yang, Danqing Huang, Chin-Yew Lin, Shengfeng HeAAAI 2023 · 2 citations
- Seeing is Improving: Visual Feedback for Iterative Text Layout RefinementJunrong Guo, Shancheng Fang, Yadong Qu, Hongtao XieCVPR 2026 · 2 citations
- Graphic Design with Large Multimodal ModelYutao Cheng, Zhao Zhang, Maoke Yang, Hui Nie et al.AAAI 2025 · 3 citations
- Order Matters: Learning Element Ordering for Graphic Design GenerationBo Yang, Ying CaoSIGGRAPH 2025 · 3 citations
- LoCo: Training-Free Layout-to-Image Synthesis with Localized ConstraintsPeiang Zhao, Han Li, Ruiyang Jin, S. Kevin ZhouACM MM 2025 · 2 citations
