Syncity: Training-Free Generation of 3D Worlds
Paul Engstler, Aleksandar Shtedritski, Iro Laina, Christian Rupprecht, Andrea Vedaldi
Abstract
We address the challenge of generating 3D worlds from textual descriptions. We propose SynCity, a training- and optimization-free approach, which leverages the geometric precision of pre-trained 3D generative models and the artistic versatility of 2D image generators to create large, high-quality 3D spaces. While most 3D generative models are object-centric and cannot generate large-scale worlds, we show how 3D and 2D generators can be combined to generate ever-expanding scenes. Through a tile-based approach, we allow fine-grained control over the layout and the appearance of scenes. The world is generated tile-by-tile, and each new tile is generated within its world-context and then fused with the scene. SynCity generates compelling and immersive scenes that are rich in detail and diversity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- LATTICE: Democratize High-Fidelity 3D Generation at ScaleZeqiang Lai, Yunfei Zhao, Zibo Zhao, Haolin Liu et al.CVPR 2026 · 46 citations
- AutoPartGen: Autoregressive 3D Part Generation and DiscoveryMinghao Chen, Jianyuan Wang, Roman Shapovalov, Tom Monnier et al.NeurIPS 2025 · 29 citations
- WorldGrow: Generating Infinite 3D WorldSikuang Li, Chen Yang, Jiemin Fang, Taoran Yi et al.AAAI 2026 · 10 citations
- Yo'City: Personalized and Boundless 3D Realistic City Scene Generation via Self-Critic ExpansionKeyang Lu, Sifan Zhou, Hongbin Xu, Gang Xu et al.CVPR 2026 · 9 citations
- Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video GeneratorHyojun Go, Dominik Narnhofer, Goutam Bhat, Prune Truong et al.ICLR 2026 · 9 citations
Builds on32
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao et al.NeurIPS 2023 · 1,498 citations
- MVDream: Multi-view Diffusion for 3D GenerationYichun Shi, Peng Wang, Jianglong Ye, Long Mai et al.ICLR 2024 · 973 citations
- 2D Gaussian Splatting for Geometrically Accurate Radiance FieldsBinbin Huang, Zehao Yu, Anpei Chen, Andreas Geiger et al.SIGGRAPH 2024 · 660 citations
Related papers
- ArtiScene: Language-Driven Artistic 3D Scene Generation Through Image IntermediaryZeqi Gu, Yin Cui, Zhaoshuo Li, Fangyin Wei et al.CVPR 2025
- Towards Text-guided 3D Scene CompositionQihang Zhang, Chaoyang Wang, Aliaksandr Siarohin, Peiye Zhuang et al.CVPR 2024 · 16 citations
- Training-Free Consistent Text-to-Image GenerationYoad Tewel, Omri Kaduri, Rinon Gal, Yoni Kasten et al.SIGGRAPH 2024 · 57 citations
- WorldGen: From Text to Traversable and Interactive 3D WorldsDilin Wang, Hyunyoung Jung, Tom Monnier, Kihyuk Sohn et al.CVPR 2026 · 24 citations
- A Recipe for Generating 3D Worlds from a Single ImageKatja Schwarz, Denis Rozumny, Samuel Rota Bulò, Lorenzo Porzi et al.ICCV 2025 · 5 citations
