EmBARDiment: an Embodied AI Agent for Productivity in XR
Riccardo Bovo, Steven Abreu, Karan Ahuja, Eric J. Gonzalez, Li-Te Cheng, Mar González-Franco
摘要
XR devices running chat-bots powered by Large Language Models (LLMs) have the to become always-on agents that enable much better productivity scenarios. Current screen based chat-bots do not take advantage of the the full-suite of natural inputs available in XR, including inward facing sensor data, instead they over-rely on explicit voice or text prompts, sometimes paired with multi-modal data dropped as part of the query. We propose a solution that leverages an attention framework that derives context implicitly from user actions, eye-gaze, and contextual memory within the XR environment. Our work minimizes the need for engineered explicit prompts, fostering grounded and intuitive interactions that glean user insights for the chat-bot.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Exploring Collaborative GenAI Agents in Synchronous Group Settings: Eliciting Team Perceptions and Design Considerations for the Future of WorkJanet G. Johnson, Macarena Peralta, Mansanjam Kaur, Ruijie Sophia Huang 等CSCW 2025 · 被引用 16 次
- Reality Proxy: Fluid Interactions with Real-World Objects in MR via Abstract RepresentationsXiaoan Liu, Difan Jia, Xianhao Carton Liu, Mar González-Franco 等UIST 2025 · 被引用 2 次
- Gaze and Speech in Multimodal Human-Computer Interaction: A Scoping ReviewAnam Ahmad Khan, Florian Weidner, Jungwoo Rhee, Yasmeen Abdrabou 等CHI 2026 · 被引用 1 次
- AgentHands: Generating Interactive Hand Gestures for Spatially Grounded Agent Conversations in XRZiyi Liu, David Li, Zhongyi Zhou, David Kim 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper6
- Robust Speech Recognition via Large-Scale Weak SupervisionAlec Radford, Jong Wook Kim, Tao Xu, Greg Brockman 等ICML 2023 · 被引用 6,966 次
- Do we still need physical monitors? An evaluation of the usability of AR virtual monitors for productivity workLeonardo Pavanatto, Chris North, Doug A. Bowman, Carmen Badea 等IEEE VR 2021 · 被引用 90 次
- Enhancing Mobile Voice Assistants with WorldGazeSven Mayer, Gierad Laput, Chris HarrisonCHI 2020 · 被引用 65 次
- Direction-of-Voice (DoV) Estimation for Intuitive Speech Interaction with Smart Devices EcosystemsKaran Ahuja, Andy Kong, Mayank Goel, Chris HarrisonUIST 2020 · 被引用 29 次
- Improving Automatic Summarization for Browsing Longform Spoken DialogDaniel Li, Thomas Chen, Alec Zadikian, Albert Tung 等CHI 2023 · 被引用 12 次
相关 Paper
- Explainable XR: Understanding User Behaviors of XR Environments Using LLM-Assisted Analytics FrameworkYoonsang Kim, Zainab Aamir, Mithilesh Kumar Singh, Saeed Boorboor 等IEEE VR 2025 · 被引用 27 次
- Exploring Large Language Model-Driven Agents for Environment-Aware Spatial Interactions and Conversations in Virtual Reality Role-Play ScenariosZiming Li, Huadong Zhang, Chao Peng, Roshan L. PeirisIEEE VR 2025 · 被引用 18 次
- InteractGuide: LLM-Enhanced Multimodal Reasoning for User-Centric Interaction Recommendations in AR-HRI AuthoringYunqiang Pei, Hongrong Yang, Kaiyue Zhang, Guoqing Wang 等ACM MM 2025 · 被引用 1 次
- Online-EYE: Multimodal Implicit Eye Tracking Calibration for XRBaosheng James Hou, Lucy Abramyan, Prasanthi Gurumurthy, Haley Adams 等CHI 2025 · 被引用 6 次
- Sensible Agent: A Framework for Unobtrusive Interaction with Proactive AR AgentsGeonsun Lee, Min Xia, Nels Numan, Xun Qian 等UIST 2025 · 被引用 15 次
