Generating Physically Stable and Buildable Brick Structures from Text
Ava Pun, Kangle Deng, Ruixuan Liu, Deva Ramanan, Changliu Liu, Jun-Yan Zhu
摘要
We introduce BRICKGPT, the first approach for generating physically stable interconnecting brick assembly models from text prompts. To achieve this, we construct a large-scale, physically stable dataset of brick structures, along with their associated captions, and train an autoregressive large language model to predict the next brick to add via next-token prediction. To improve the stability of the resulting designs, we employ an efficient validity check and physics-aware rollback during autoregressive inference, which prunes infeasible token predictions using physics laws and assembly constraints. Our experiments show that BRICKGPT produces stable, diverse, and aesthetically pleasing brick structures that align closely with the input text prompts. We also develop a text-based brick texturing method to gen-* Indicates equal contribution.
erate colored and textured designs. We show that our designs can be assembled manually by humans and automatically by robotic arms. We release our new dataset, Stable-Text2Brick, containing over 47,000 brick structures of over 28,000 unique 3D objects accompanied by detailed captions, along with our code and models at the project website: https://avalovelace1.github.io/BrickGPT/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- LottieGPT: Tokenizing Vector Animation for Autoregressive GenerationJunhao Chen, Kejun Gao, Yuehan Cui, Mingze Sun 等CVPR 2026 · 被引用 10 次
- BrickNet: Graph-Backed Generative Brick AssemblyPeter Kulits, Cordelia SchmidCVPR 2026 · 被引用 5 次
- Voxify3D: Pixel Art Meets Volumetric RenderingYi-Chuan Huang, Jiewen Chan, Hao-Jen Chien, Yu-Lun LiuCVPR 2026 · 被引用 3 次
- CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language ModelsJunming Huang, Chi Wang, Letian Li, Guangkai Xu 等ICML 2026 · 被引用 2 次
- AssemblyBench: Physics-Aware Assembly of Complex Industrial ObjectsDanrui Li, Jiahao Zhang, Bernhard Egger, Moitreya Chatterjee 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper46
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao 等NeurIPS 2023 · 被引用 1,498 次
相关 Paper
- BuildingGPT: Auto-Regressive Building Wireframe Reconstruction Model with Reinforcement LearningYuzhou Liu, Lingjie Zhu, Hanqiao Ye, Yujun Liu 等CVPR 2026
- ArtLLM: Generating Articulated Assets via 3D LLMPenghao Wang, Siyuan Xie, Hongyu Yan, Xianghui Yang 等CVPR 2026 · 被引用 7 次
- CADMate: Generating CAD Assembly Plan with Geometric Chain-of-Thought and Spatial Physical RewardsJiali Chen, DingBa Fu, Xusen Hei, Yuhang Liu 等ACL 2026
- BuildingBlock: A Hybrid Approach for Structured Building GenerationJunming Huang, Chi Wang, Letian Li, Changxin Huang 等SIGGRAPH 2025
- DressCode: Autoregressively Sewing and Generating Garments from Text GuidanceKai He, Kaixin Yao, Qixuan Zhang, Jingyi Yu 等SIGGRAPH 2024 · 被引用 43 次
