MuteIt: Jaw Motion Based Unvoiced Command Recognition Using Earable
Tanmay Srivastava, Prerna Khanna, Shijia Pan, Phuc Nguyen, Shubham Jain
摘要
In this paper, we present MuteIt, an ear-worn system for recognizing unvoiced human commands. MuteIt presents an intuitive alternative to voice-based interactions that can be unreliable in noisy environments, disruptive to those around us, and compromise our privacy. We propose a twin-IMU set up to track the user's jaw motion and cancel motion artifacts caused by head and body movements. MuteIt processes jaw motion during word articulation to break each word signal into its constituent syllables, and further each syllable into phonemes (vowels, visemes, and plosives). Recognizing unvoiced commands by only tracking jaw motion is challenging. As a secondary articulator, jaw motion is not distinctive enough for unvoiced speech recognition. MuteIt combines IMU data with the anatomy of jaw movement as well as principles from linguistics, to model the task of word recognition as an estimation problem. Rather than employing machine learning to train a word classifier, we reconstruct each word as a sequence of phonemes using a bi-directional particle filter, enabling the system to be easily scaled to a large set of words. We validate MuteIt for 20 subjects with diverse speech accents to recognize 100 common command words. MuteIt achieves a mean word recognition accuracy of 94.8% in noise-free conditions. When compared with common voice assistants, MuteIt outperforms them in noisy acoustic environments, achieving higher than 90% recognition accuracy. Even in the presence of motion artifacts, such as head movement, walking, and riding in a moving vehicle, MuteIt achieves mean word recognition accuracy of 91% over all scenarios.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- TRAMBA: A Hybrid Transformer and Mamba Architecture for Practical Audio and Bone Conduction Speech Super Resolution and Enhancement on Mobile and Wearable PlatformsYueyuan Sui, Minghui Zhao, Junxi Xia, Xiaofan Jiang 等UbiComp 2025 · 被引用 18 次
- MELDER: The Design and Evaluation of a Real-time Silent Speech Recognizer for Mobile DevicesLaxmi Pandey, Ahmed Sabbir ArifCHI 2024 · 被引用 14 次
- Exploring Uni-manual Around Ear Off-Device Gestures for EarablesShaikh Shawon Arefin Shimon, Ali Neshati, Junwei Sun, Qiang Xu 等UbiComp 2024 · 被引用 4 次
- A Survey of Earable Technology: Trends, Tools, and the Road AheadChangshuo Hu, Qiang Yang, Yang Liu, Tobias Röddiger 等UbiComp 2026 · 被引用 4 次
- MorsEar: Toward Generalizable Low-Resource Covert Messaging via Earable based Inertial SensingGarvit Chugh, Indrajeet Ghosh, Nirmalya Roy, Sandip Chakraborty 等CHI 2026 · 被引用 1 次
相关 Paper
- EarCommand: "Hearing" Your Silent Speech Commands In EarYincheng Jin, Yang Gao, Xuhai Xu, Seokmin Choi 等UbiComp 2022 · 被引用 37 次
- Lipwatch: Enabling Silent Speech Recognition on Smartwatches using Acoustic SensingQian Zhang, Yubin Lan, Kaiyi Guo, Dong WangUbiComp 2024 · 被引用 20 次
- Baro2Talk: Reconstructing Spectrograms from Ear Canal Pressure for Voice-free CommunicationLuo Zhou, Shan Chang, Han Wang, Xianbo Wang 等INFOCOM 2026
- ReHEarSSE: Recognizing Hidden-in-the-Ear Silently Spelled ExpressionsXuefu Dong, Yifei Chen, Yuuki Nishiyama, Kaoru Sezaki 等CHI 2024 · 被引用 20 次
- Endophasia: Utilizing Acoustic-Based Imaging for Issuing Contact-Free Silent Speech CommandsYongzhao Zhang, Wei-Hsiang Huang, Chih-Yun Yang, Wen-Ping Wang 等UbiComp 2020 · 被引用 43 次
