WorldSmith: Iterative and Expressive Prompting for World Building with a Generative AI
Hai Dang, Frederik Brudy, George W. Fitzmaurice, Fraser Anderson
摘要
Crafting a rich and unique environment is crucial for fictional world-building, but can be difficult to achieve since illustrating a world from scratch requires time and significant skill. We investigate the use of recent multi-modal image generation systems to enable users iteratively visualize and modify elements of their fictional world using a combination of text input, sketching, and region-based filling. WorldSmith enables novice world builders to quickly visualize a fictional world with layered edits and hierarchical compositions. Through a formative study (4 participants) and first-use study (13 participants) we demonstrate that WorldSmith offers more expressive interactions with prompt-based models. With this work, we explore how creatives can be empowered to leverage prompt-based generative AI as a tool in their creative process, beyond current "click-once" prompting UI paradigms.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- GenQuery: Supporting Expressive Visual Search with Generative ModelsKihoon Son, DaEun Choi, Tae Soo Kim, Young-Ho Kim 等CHI 2024 · 被引用 48 次
- Rambler: Supporting Writing With Speech via LLM-Assisted Gist ManipulationSusan Lin, Jeremy Warner, J. D. Zamfirescu-Pereira, Matthew G. Lee 等CHI 2024 · 被引用 38 次
- Patchview: LLM-powered Worldbuilding with Generative Dust and Magnet VisualizationJohn Joon Young Chung, Max KreminskiUIST 2024 · 被引用 29 次
- SpaceBlender: Creating Context-Rich Collaborative Spaces Through Generative 3D Scene BlendingNels Numan, Shwetha Rajaram, Balasaravanan Thoravi Kumaravel, Nicolai Marquardt 等UIST 2024 · 被引用 17 次
- Paratrouper: Exploratory Creation of Character Cast Visuals Using Generative AIJoanne Leong, David Ledo, Thomas Driscoll, Tovi Grossman 等CHI 2025 · 被引用 16 次
它引用的顶会 Paper16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- BARTScore: Evaluating Generated Text as Text GenerationWeizhe Yuan, Graham Neubig, Pengfei LiuNeurIPS 2021 · 被引用 1,143 次
- Design Guidelines for Prompt Engineering Text-to-Image Generative ModelsVivian Liu, Lydia B. ChiltonCHI 2022 · 被引用 586 次
相关 Paper
- Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language ModelsStephen Brade, Bryan Wang, Maurício Sousa, Sageev Oore 等UIST 2023 · 被引用 179 次
- WorldGen: From Text to Traversable and Interactive 3D WorldsDilin Wang, Hyunyoung Jung, Tom Monnier, Kihyuk Sohn 等CVPR 2026 · 被引用 24 次
- LayerCraft: Enhancing Text-to-Image Generation with CoT Reasoning and Layered Object IntegrationYuyao Zhang, Jinghao Li, Yu-Wing TaiNeurIPS 2025 · 被引用 21 次
- DesignWeaver: Dimensional Scaffolding for Text-to-Image Product DesignSirui Tao, Ivan Liang, Cindy Peng, Zhiqing Wang 等CHI 2025 · 被引用 19 次
- The Impact of Sketch-guided vs. Prompt-guided 3D Generative AIs on the Design Exploration ProcessSeung Won Lee, Tae Hee Jo, Semin Jin, Jiin Choi 等CHI 2024 · 被引用 48 次
