AMMA: Adaptive Multimodal Assistants Through Automated State Tracking and User Model-Directed Guidance Planning
Jackie (Junrui) Yang, Leping Qiu, Emmanuel Angel Corona-Moreno, Louisa Shi, Hung Bui, Monica S. Lam, James A. Landay
摘要
Novel technologies such as augmented reality and computer perception lay the foundation for smart assistants that can guide us through real-world tasks, such as cooking or home repair. However, the nature of real-world interaction requires assistants that adapt to users’ mistakes, environments, and communication preferences. We propose Adaptive Multimodal Assistants (AMMA), a software architecture for task guidance with generated adaptive interfaces from step-by-step instructions. This is achieved through 1) an automatically generated user action state tracker and 2) a guidance planner that leverages a continuously trained user model. The assistant also adjusts its guidance and communication delivery methods based on observed user performance as well as implicit and explicit user feedback. We demonstrated the viability of AMMA by building an adaptive cooking assistant running in a high-fidelity virtual reality-based simulator. A user study of the cooking assistant showed that AMMA can reduce the task completion time and the number of manual communication methods changes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- ProMemAssist: Exploring Timely Proactive Assistance Through Working Memory Modeling in Multi-Modal Wearable DevicesKevin Pu, Ting Zhang, Naveen Sendhilnathan, Sebastian Freitag 等UIST 2025 · 被引用 9 次
- Scaling Context-Aware Task Assistants that Learn from Demonstration and Adapt through Mixed-Initiative DialogueRiku Arakawa, Prasoon Patidar, Will Page, Jill Lehman 等UIST 2025 · 被引用 3 次
- Seeing Eye to Eye: Enabling Cognitive Alignment Through Shared First-Person Perspective in Human-AI Collaboration: Seeing Eye to EyeZhuyu Teng, Pei Chen, Yichen Cai, Ruoqing Lu 等CHI 2026 · 被引用 2 次
- Gesturing Toward Abstraction: Multimodal Convention Formation in Collaborative Physical TasksKiyosu Maeda, William P. McCarthy, Ching-Yi Tsai, Jeffrey Mu 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper6
- AdapTutAR: An Adaptive Tutoring System for Machine Tasks in Augmented RealityGaoping Huang, Xun Qian, Tianyi Wang, Fagun Patel 等CHI 2021 · 被引用 93 次
- Reinforcement Learning for the Adaptive Scheduling of Educational ActivitiesJonathan Bassen, Bharathan Balaji, Michael Schaarschmidt, Candace Thille 等CHI 2020 · 被引用 73 次
- An Exploratory Study of Augmented Reality Presence for Tutoring Machine TasksYuanzhi Cao, Xun Qian, Tianyi Wang, Rachel Lee 等CHI 2020 · 被引用 73 次
- ScalAR: Authoring Semantically Adaptive Augmented Reality Experiences in Virtual RealityXun Qian, Fengming He, Xiyun Hu, Tianyi Wang 等CHI 2022 · 被引用 69 次
- HybridTrak: Adding Full-Body Tracking to VR Using an Off-the-Shelf WebcamJackie (Junrui) Yang, Tuochao Chen, Fang Qin, Monica S. Lam 等CHI 2022 · 被引用 39 次
相关 Paper
- Pro 2 Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural TasksLilin Xu, Bufang Yang, Siyang Jiang, Kaiwei Liu 等UbiComp 2026
- Satori 悟り: Towards Proactive AR Assistant with Belief-Desire-Intention User ModelingChenyi Li, Guande Wu, Gromit Yeuk-Yin Chan, Dishita G. Turakhia 等CHI 2025 · 被引用 49 次
- ARGESTUREAID: A Voice-Based, Adaptive, and Context-Aware Conversational Assistant for Supporting Mid-Air Gesture Discovery and ExecutionAnjali Khurana, Amy Karlson, Christopher Collins, Mengjie Yu 等UbiComp 2026
- "Mango Mango, How to Let The Lettuce Dry Without A Spinner?": Exploring User Perceptions of Using An LLM-Based Conversational Assistant Toward Cooking PartnerSzeyi Chan, Jiachen Li, Bingsheng Yao, Amama Mahmood 等CSCW 2025 · 被引用 4 次
- Identifying Multimodal Context Awareness Requirements for Supporting User Interaction with Procedural VideosGeorgianna Lin, Jin Yi Li, Afsaneh Fazly, Vladimir Pavlovic 等CHI 2023 · 被引用 12 次
