"Rewind to the Jiggling Meat Part": Understanding Voice Control of Instructional Videos in Everyday Tasks
Yaxi Zhao, Razan Jaber, Donald McMillan, Cosmin Munteanu
摘要
Voice interaction has long been envisioned as enabling users to transform physical interaction into hands-free, such as allowing fine-grained control of instructional videos without physically disengaging from the task at hand. While significant engineering advances have brought us closer to this ideal, we do not fully understand the user requirements for voice interactions that should be supported in such contexts. This paper presents an ecologically-valid wizard-of-oz elicitation study exploring realistic user requirements for an ideal instructional video playback control while cooking. Through the analysis of the issued commands and performed actions during this non-linear and complex task, we identify (1) patterns of command formulation, (2) challenges for design, and (3) how task and voice-based commands are interwoven in real-life. We discuss implications for the design and research of voice interactions for navigating instructional videos while performing complex tasks.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- TutoAI: a cross-domain framework for AI-assisted mixed-media tutorial creation on physical tasksYuexi Chen, Vlad I. Morariu, Anh Truong, Zhicheng LiuCHI 2024 · 被引用 16 次
- AQuA: Automated Question-Answering in Software Tutorial Videos with Visual AnchorsSaelyne Yang, Jo Vermeulen, George W. Fitzmaurice, Justin MatejkaCHI 2024 · 被引用 15 次
- Stargazer: An Interactive Camera Robot for Capturing How-To Videos Based on Subtle Instructor CuesJiannan Li, Maurício Sousa, Karthik Mahadevan, Bryan Wang 等CHI 2023 · 被引用 12 次
- Vid2Coach: Transforming How-To Videos into Task AssistantsMina Huh, Zihui Xue, Ujjaini Das, Kumar Ashutosh 等UIST 2025 · 被引用 9 次
- "Mango Mango, How to Let The Lettuce Dry Without A Spinner?": Exploring User Perceptions of Using An LLM-Based Conversational Assistant Toward Cooking PartnerSzeyi Chan, Jiachen Li, Bingsheng Yao, Amama Mahmood 等CSCW 2025 · 被引用 4 次
相关 Paper
- Identifying Multimodal Context Awareness Requirements for Supporting User Interaction with Procedural VideosGeorgianna Lin, Jin Yi Li, Afsaneh Fazly, Vladimir Pavlovic 等CHI 2023 · 被引用 12 次
- Cooking With Agents: Designing Context-aware Voice InteractionRazan Jaber, Sabrina Zhong, Sanna Kuoppamäki, Aida Hosseini 等CHI 2024 · 被引用 36 次
- Freehand Grasping: An Analysis of Grasping for Docking Tasks in Virtual RealityAndreea-Dalia Blaga, Maite Frutos-Pascual, Chris Creed, Ian WilliamsIEEE VR 2021 · 被引用 18 次
- Better to Ask Than Assume: Proactive Voice Assistants' Communication Strategies That Respect User Agency in a Smart Home EnvironmentJeesun Oh, Wooseok Kim, Sungbae Kim, Hyeonjeong Im 等CHI 2024 · 被引用 22 次
- On Pause: How Online Instructional Videos are Used to Achieve Practical TasksSylvaine Tuncer, Barry A. T. Brown, Oskar LindwallCHI 2020 · 被引用 30 次
