LLMR: Real-time Prompting of Interactive Worlds using Large Language Models
Fernanda De La Torre, Cathy Mengying Fang, Han Huang, Andrzej Banburski-Fahey, Judith Amores Fernandez, Jaron Lanier
摘要
We present Large Language Model for Mixed Reality (LLMR), a framework for the real-time creation and modification of interactive Mixed Reality experiences using LLMs. LLMR leverages novel strategies to tackle difficult cases where ideal training data is scarce, or where the design goal requires the synthesis of internal dynamics, intuitive analysis, or advanced interactivity. Our framework relies on text interaction and the Unity game engine. By incorporating techniques for scene understanding, task planning, self-debugging, and memory management, LLMR outperforms the standard GPT-4 by 4x in average error rate. We demonstrate LLMR’s cross-platform interoperability with several example worlds, and evaluate it on a variety of creation and modification tasks to show that it can produce and edit diverse objects, tools, and scenes. Finally, we conducted a usability study (N=11) with a diverse set that revealed participants had positive experiences with the system and would use it again.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper35
- Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature ReviewRock Yuren Pang, Hope Schroeder, Kynnedy Simone Smith, Solon Barocas 等CHI 2025 · 被引用 51 次
- AIdeation: Designing a Human-AI Collaborative Ideation System for Concept DesignersWen-Fan Wang, Chien-Ting Lu, Nil Ponsa Campanyà, Bing-Yu Chen 等CHI 2025 · 被引用 46 次
- DreamCodeVR: Towards Democratizing Behavior Design in Virtual Reality with Speech-Driven ProgrammingDaniele Giunchi, Nels Numan, Elia Gatti, Anthony SteedIEEE VR 2024 · 被引用 43 次
- Vision-Based Multimodal Interfaces: A Survey and Taxonomy for Enhanced Context-Aware System DesignYongquan 'Owen' Hu, Jingyu Tang, Xinya Gong, Zhongyi Zhou 等CHI 2025 · 被引用 37 次
- How CO2STLY Is CHI? The Carbon Footprint of Generative AI in HCI Research and What We Should Do About ItNanna Inie, Jeanette Falk, Raghavendra SelvanCHI 2025 · 被引用 33 次
它引用的顶会 Paper26
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch 等ICML 2023 · 被引用 2,601 次
- 3D-LLM: Injecting the 3D World into Large Language ModelsYining Hong, Haoyu Zhen, Peihao Chen, Shuhong Zheng 等NeurIPS 2023 · 被引用 662 次
- Instruct-NeRF2NeRF: Editing 3D Scenes with InstructionsAyaan Haque, Matthew Tancik, Alexei A. Efros, Aleksander Holynski 等ICCV 2023 · 被引用 544 次
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 被引用 463 次
相关 Paper
- LLMER: Crafting Interactive Extended Reality Worlds with JSON Data Generated by Large Language ModelsJiangong Chen, Xiaoyi Wu, Tian Lan, Bin LiIEEE VR 2025 · 被引用 20 次
- VirtualEnv: A Platform for Embodied AI ResearchKabir Swain, Sijie Han, Ayush Raina, Jin Zhang 等AAAI 2026
- Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D ScenesJunlong Chen, Jens Grubert, Per Ola KristenssonIEEE VR 2025 · 被引用 9 次
- LIVE-GS: LLM Powers Interactive VR Experience with Physics-Aware Gaussian SplattingHaotian Mao, Hangyu Zhou, Zhuoxiong Xu, Siyue Wei 等IEEE VR 2026 · 被引用 1 次
- Text2VRScene: Exploring the Framework of Automated Text-driven Generation System for VR ExperienceZhizhuo Yin, Yuyang Wang, Theodoros Papatheodorou, Pan HuiIEEE VR 2024 · 被引用 25 次
