Enabling Voice-Accompanying Hand-to-Face Gesture Recognition with Cross-Device Sensing
Zisu Li, Chen Liang, Yuntao Wang, Yue Qin, Chun Yu, Yukang Yan, Mingming Fan, Yuanchun Shi
Abstract
Gestures performed accompanying the voice are essential for voice interaction to convey complementary semantics for interaction purposes such as wake-up state and input modality. In this paper, we investigated voice-accompanying hand-to-face (VAHF) gestures for voice interaction. We targeted on hand-to-face gestures because such gestures relate closely to speech and yield significant acoustic features (e.g., impeding voice propagation). We conducted a user study to explore the design space of VAHF gestures, where we first gathered candidate gestures and then applied a structural analysis to them in different dimensions (e.g., contact position and type), outputting a total of 8 VAHF gestures with good usability and least confusion. To facilitate VAHF gesture recognition, we proposed a novel cross-device sensing method that leverages heterogeneous channels (vocal, ultrasound, and IMU) of data from commodity devices (earbuds, watches, and rings). Our recognition model achieved an accuracy of 97.3% for recognizing 3 gestures and 91.5% for recognizing 8 gestures (excluding the "empty" gesture), proving the high applicability. Quantitative analysis also shed light on the recognition capability of each sensor channel and their different combinations. In the end, we illustrated the feasible use cases and their design principles to demonstrate the applicability of our system in various scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bf3661c1-6820-4ee9-ae45-0596387cb35dCited by top-tier papers5
- UbiPhysio: Support Daily Functioning, Fitness, and Rehabilitation with Action Understanding and Feedback in Natural LanguageChongyang Wang, Yuan Feng, Lingxiao Zhong, Siyi Zhu et al.UbiComp 2024 · 20 citations
- MAF: Exploring Mobile Acoustic Field for Hand-to-Face Gesture InteractionsYongjie Yang, Tao Chen, Yujing Huang, Xiuzhen Guo et al.CHI 2024 · 15 citations
- Computing with Smart Rings: A Systematic Literature ReviewZeyu Wang, Ruotong Yu, Xiangyang Wang, Jiexin Ding et al.UbiComp 2025 · 14 citations
- LION-FS: Fast & Slow Video-Language Thinker as Online Video AssistantWei Li, Bing Hu, Rui Shao, Leyang Shen et al.CVPR 2025
- FlexiCamAR: Enhancing Everyday Camera Interactions on AR Glasses with a Flexible Additional ViewpointZiming Li, Hongji Li, Jialin Wang, Pan Hui et al.IEEE VR 2026
Builds on18
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- EarBuddy: Enabling On-Face Interaction via Wireless EarbudsXuhai Xu, Haitian Shi, Xin Yi, Wenjia Liu et al.CHI 2020 · 91 citations
- NeuroPose: 3D Hand Pose Tracking using EMG WearablesYilin Liu, Shijia Zhang, Mahanth GowdaWWW 2021 · 84 citations
- DualRing: Enabling Subtle and Expressive Hand Interaction with Dual IMU RingsChen Liang, Chun Yu, Yue Qin, Yuntao Wang et al.UbiComp 2021 · 69 citations
- Enhancing Mobile Voice Assistants with WorldGazeSven Mayer, Gierad Laput, Chris HarrisonCHI 2020 · 65 citations
Related papers
- EarHover: Mid-Air Gesture Recognition for Hearables Using Sound Leakage SignalsShunta Suzuki, Takashi Amesaka, Hiroki Watanabe, Buntarou Shizuki et al.UIST 2024 · 5 citations
- Designing and Evaluating Hand-to-Hand Gestures with Dual Commodity Wrist-Worn DevicesYiqin Lu, Bingjian Huang, Chun Yu, Guahong Liu et al.UbiComp 2020 · 37 citations
- SonicFace: Tracking Facial Expressions Using a Commodity Microphone ArrayYang Gao, Yincheng Jin, Seokmin Choi, Jiyang Li et al.UbiComp 2022 · 14 citations
- Hybrid Zone: Bridging Acoustic and Wi-Fi for Enhanced Gesture RecognitionMengning Li, Wenye WangINFOCOM 2024 · 12 citations
- AO-Finger: Hands-free Fine-grained Finger Gesture Recognition via Acoustic-Optic Sensor FusingChenhan Xu, Bing Zhou, Gurunandan Krishnan, Shree K. NayarCHI 2023 · 25 citations
