Augmented Object Intelligence with XR-Objects
Mustafa Doga Dogan, Eric J. Gonzalez, Karan Ahuja, Ruofei Du, Andrea Colaço, Johnny Lee, Mar González-Franco, David Kim
摘要
Seamless integration of physical objects as interactive digital entities remains a challenge for spatial computing. This paper explores Augmented Object Intelligence (AOI) in the context of XR, an interaction paradigm that aims to blur the lines between digital and physical by equipping real-world objects with the ability to interact as if they were digital, where every object has the potential to serve as a portal to digital functionalities. Our approach utilizes real-time object segmentation and classification, combined with the power of Multimodal Large Language Models (MLLMs), to facilitate these interactions without the need for object pre-registration. We implement the AOI concept in the form of XR-Objects, an open-source prototype system that provides a platform for users to engage with their physical environment in contextually relevant ways using object-based context menus. This system enables analog objects to not only convey information but also to initiate digital actions, such as querying for details or executing tasks. Our contributions are threefold: (1) we define the AOI concept and detail its advantages over traditional AI assistants, (2) detail the XR-Objects system’s open-source design and implementation, and (3) show its versatility through various use cases and a user study.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- AiGet: Transforming Everyday Moments into Hidden Knowledge Discovery with AI Assistance on Smart GlassesRunze Cai, Nuwan Janaka, Hyeongcheol Kim, Yang Chen 等CHI 2025 · 被引用 26 次
- Sensible Agent: A Framework for Unobtrusive Interaction with Proactive AR AgentsGeonsun Lee, Min Xia, Nels Numan, Xun Qian 等UIST 2025 · 被引用 15 次
- InteRecon: Towards Reconstructing Interactivity of Personal Memorable Items in Mixed RealityZisu Li, Jiawei Li, Zeyu Xiong, Shumeng Zhang 等CHI 2025 · 被引用 14 次
- Guided Reality: Generating Visually-Enriched AR Task Guidance with LLMs and Vision ModelsAda Yi Zhao, Aditya Gunturu, Ellen Yi-Luen Do, Ryo SuzukiUIST 2025 · 被引用 12 次
- Draw2Cut: Direct On-Material Annotations for CNC MillingXinyue Gui, Ding Xia, Wang Gao, Mustafa Doga Dogan 等CHI 2025 · 被引用 9 次
它引用的顶会 Paper30
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- DepthLab: Real-time 3D Interaction with Depth Maps for Mobile Augmented RealityRuofei Du, Eric Turner, Maksym Dzitsiuk, Luca Prasso 等UIST 2020 · 被引用 145 次
- Luminate: Structured Generation and Exploration of Design Space with Large Language Models for Human-AI Co-CreationSangho Suh, Meng Chen, Bryan Min, Toby Jia-Jun Li 等CHI 2024 · 被引用 143 次
- SemanticAdapt: Optimization-based Adaptation of Mixed Reality Layouts Leveraging Virtual-Physical Semantic ConnectionsYifei Cheng, Yukang Yan, Xin Yi, Yuanchun Shi 等UIST 2021 · 被引用 143 次
- The Dark Side of Perceptual Manipulations in Virtual RealityWen-Jie Tseng, Elise Bonnail, Mark McGill, Mohamed Khamis 等CHI 2022 · 被引用 118 次
相关 Paper
- Reality Proxy: Fluid Interactions with Real-World Objects in MR via Abstract RepresentationsXiaoan Liu, Difan Jia, Xianhao Carton Liu, Mar González-Franco 等UIST 2025 · 被引用 2 次
- OmniActions: Predicting Digital Actions in Response to Real-World Multimodal Sensory Inputs with LLMsJiahao Nick Li, Yan Xu, Tovi Grossman, Stephanie Santosa 等CHI 2024 · 被引用 25 次
- Explainable XR: Understanding User Behaviors of XR Environments Using LLM-Assisted Analytics FrameworkYoonsang Kim, Zainab Aamir, Mithilesh Kumar Singh, Saeed Boorboor 等IEEE VR 2025 · 被引用 27 次
- HAMMER: Harnessing MLLMs via Cross-Modal Integration for Intention-Driven 3D Affordance GroundingLei Yao, Yong Chen, Yuejiao Su, Yi Wang 等CVPR 2026
- GenAssist: Interactive Prompt-Driven XR Program GenerationSruti Srinidhi, Akul Singh, Edward Lu, Anthony RoweIEEE VR 2026
