Tell2Design: A Dataset for Language-Guided Floor Plan Generation
Sicong Leng, Yang Zhou, Mohammed Haroon Dupty, Wee Sun Lee, Sam Joyce, Wei Lu
Abstract
We consider the task of generating designs directly from natural language descriptions, and consider floor plan generation as the initial research area. Language conditional generative models have recently been very successful in generating high-quality artistic images. However, designs must satisfy different constraints that are not present in generating artistic images, particularly spatial and relational constraints. We make multiple contributions to initiate research on this task. First, we introduce a novel dataset, Tell2Design (T2D), which contains more than 80k floor plan designs associated with natural language instructions. Second, we propose a Sequence-to-Sequence model that can serve as a strong baseline for future research. Third, we benchmark this task with several text-conditional image generation models. We conclude by conducting human evaluations on the generated samples and providing an analysis of human performance. We hope our contributions will propel the research on language-guided design generation forward 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 81017eac-a75f-4c1c-95c8-b341cc03835aCited by top-tier papers9
- FloorPlan-LLaMa: Aligning Architects' Feedback and Domain Knowledge in Architectural Floor Plan GenerationJun Yin, Pengyu Zeng, Haoyuan Sun, Yuqin Dai et al.ACL 2025 · 11 citations
- MANSION: Multi-floor lANguage-to-3D Scene generatIOn for loNg-horizon tasksLirong Che, Shuo Wen, Shan Huang, Chuang Wang et al.CVPR 2026 · 8 citations
- CARD: Cross-modal Agent Framework for Generative and Editable Residential DesignPengyu Zeng, Jun Yin, Miao Zhang, Yuqin Dai et al.EMNLP 2025 · 6 citations
- Text-to-Code Generation for Modular Building Layouts in Building Information ModelingYinyi Wei, Xiao LiNeurIPS 2025 · 2 citations
- Tokenization Allows Multimodal Large Language Models to Understand, Generate and Edit Architectural Floor PlansSizhong Qin, Ramon Elias Weber, Xinzheng LuCVPR 2026 · 1 citation
Builds on15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
Related papers
- Intelligent Home 3D: Automatic 3D-House Design From Linguistic Descriptions OnlyQi Chen, Qi Wu, Rui Tang, Yuhan Wang et al.CVPR 2020
- HouseTune: Two-Stage Floorplan Generation with LLM AssistanceZiyang Zong, Guanying Chen, Zhaohuan Zhan, Fengcheng Yu et al.AAAI 2026
- Unified Vector Floorplan Generation via Markup RepresentationKaede Shiohara, Toshihiko YamasakiCVPR 2026
- Placeit3d: Language-Guided Object Placement in Real 3D ScenesAhmed Abdelreheem, Filippo Aleotti, Jamie Watson, Zawar Qureshi et al.ICCV 2025 · 11 citations
- CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image GenerationHui Zhang, Dexiang Hong, Yitong Wang, Jie Shao et al.ICCV 2025 · 7 citations
