Viago: Exploring Visual-Audio Modality Transitions for Social Media Consumption on the Go
Ruei-Che Chang, Tovi Grossman, Carine Rognon, Michael Glueck, Christopher Collins, Amy Karlson, Hemant Bhaskar Surale
摘要
Figure 1: Viago supports visual-audio modality transitions for mobile users on the go. In a scenario where a user is waiting for an order and browsing a mobile app: (a) The user initiates Viago by tapping the orange button, which provides a visual and audio preview, preparing the user for audio interactions on the go. (b) By vaguely recalling the layout of the interface and the available audio elements, the user can navigate the screen with touch gestures and audio feedback while heading to pick up the order. The important action (e.g., submit) will be cached for further visual confirmation before issuing. (c) Upon returning to the seat and resuming visual use, Viago provides visual and audio review to refresh the user's memory about content they heard and actions they did, and automatically restores cached actions, allowing the user to complete them visually.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper38
- Enabling Conversational Interaction with Mobile UI using Large Language ModelsBryan Wang, Gang Li, Yang LiCHI 2023 · 被引用 149 次
- AuraRing: Precise Electromagnetic Finger TrackingFarshid Salemi Parizi, Eric Whitmire, Shwetak N. PatelUbiComp 2020 · 被引用 101 次
- Unblind your apps: predicting natural-language labels for mobile GUI components by deep learningJieshan Chen, Chunyang Chen, Zhenchang Xing, Xiwei Xu 等ICSE 2020 · 被引用 101 次
- Screen2Words: Automatic Mobile UI Summarization with Multimodal LearningBryan Wang, Gang Li, Xin Zhou, Zhourong Chen 等UIST 2021 · 被引用 97 次
- Enabling Hand Gesture Customization on Wrist-Worn DevicesXuhai Xu, Jun Gong, Carolina Brum, Lilian Liang 等CHI 2022 · 被引用 82 次
相关 Paper
- Exploring the Design of Human Speech Indicators to Enhance Waiting Experience in Voice User InterfaceWenan Li, Junnan Yu, Yehong Zhou, Jinlei Shi 等CHI 2025 · 被引用 3 次
- LSAR: Sparse Lexical Representation Learning for Efficient and Interpretable Audio RetrievalHaoyue Li, Yuzhe Bai, Li NiuKDD 2026
- AutoAD II: The Sequel - Who, When, and What in Movie Audio DescriptionTengda Han, Max Bain, Arsha Nagrani, Gül Varol 等ICCV 2023 · 被引用 55 次
- LSVP: Towards Effective On-the-go Video Learning Using Optical Head-Mounted DisplaysAshwin Ram, Shengdong ZhaoUbiComp 2021 · 被引用 23 次
- GhostUI: Unveiling Hidden Interactions in Mobile UIMinkyu Kweon, Seokhyeon Park, Soohyun Lee, You Been Lee 等CHI 2026 · 被引用 1 次
