RubySlippers: Supporting Content-based Voice Navigation for How-to Videos
Minsuk Chang, Mina Huh, Juho Kim
摘要
Directly manipulating the timeline, such as scrubbing for thumbnails, is the standard way of controlling how-to videos. However, when how-to videos involve physical activities, people inconveniently alternate between controlling the video and performing the tasks. Adopting a voice user interface allows people to control the video with voice while performing the tasks with hands. However, naively translating timeline manipulation into voice user interfaces (VUI) results in temporal referencing (e.g. “rewind 20 seconds”), which requires a different mental model for navigation and thereby limiting users’ ability to peek into the content. We present RubySlippers, a system that supports efficient content-based voice navigation through keyword-based queries. Our computational pipeline automatically detects referenceable elements in the video, and finds the video segmentation that minimizes the number of needed navigational commands. Our evaluation (N=12) shows that participants could perform three representative navigation tasks with fewer commands and less frustration using RubySlippers than the conventional voice-enabled video interface.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper12
- A Literature Review of Video-Sharing Platform Research in HCIAva Bartolome, Shuo NiuCHI 2023 · 被引用 63 次
- AVscript: Accessible Video Editing with Audio-Visual ScriptsMina Huh, Saelyne Yang, Yi-Hao Peng, Xiang 'Anthony' Chen 等CHI 2023 · 被引用 44 次
- "It Feels Like Taking a Gamble": Exploring Perceptions, Practices, and Challenges of Using Makeup and Cosmetics for People with Visual ImpairmentsFranklin Mingzhe Li, Franchesca Spektor, Meng Xia, Mina Huh 等CHI 2022 · 被引用 39 次
- VideoDiff: Human-AI Video Co-Creation with AlternativesMina Huh, Ding Li, Kim Pimmel, Hijung Valentina Shin 等CHI 2025 · 被引用 26 次
- Synthesis-Assisted Video Prototyping From a DocumentPeggy Chi, Tao Dong, Christian Früh, Brian Colonna 等UIST 2022 · 被引用 18 次
相关 Paper
- Automatic Generation of Two-Level Hierarchical Tutorials from Instructional Makeup VideosAnh Truong, Peggy Chi, David Salesin, Irfan Essa 等CHI 2021 · 被引用 57 次
- "Rewind to the Jiggling Meat Part": Understanding Voice Control of Instructional Videos in Everyday TasksYaxi Zhao, Razan Jaber, Donald McMillan, Cosmin MunteanuCHI 2022 · 被引用 18 次
- Identifying Multimodal Context Awareness Requirements for Supporting User Interaction with Procedural VideosGeorgianna Lin, Jin Yi Li, Afsaneh Fazly, Vladimir Pavlovic 等CHI 2023 · 被引用 12 次
- Reactive Video: Adaptive Video Playback Based on User Motion for Supporting Physical ActivityChristopher Clarke, Doga Cavdir, Patrick Chiu, Laurent Denoue 等UIST 2020 · 被引用 30 次
- Automatic Instructional Video Creation from a Markdown-Formatted TutorialPeggy Chi, Nathan Frey, Katrina Panovich, Irfan EssaUIST 2021 · 被引用 27 次
