Endophasia: Utilizing Acoustic-Based Imaging for Issuing Contact-Free Silent Speech Commands
Yongzhao Zhang, Wei-Hsiang Huang, Chih-Yun Yang, Wen-Ping Wang, Yi-Chao Chen, Chuang-Wen You, Da-Yuan Huang, Guangtao Xue, Jiadi Yu
Abstract
Using silent speech to issue commands has received growing attention, as users can utilize existing command sets from voice-based interfaces without attracting other people's attention. Such interaction maintains privacy and social acceptance from others. However, current solutions for recognizing silent speech mainly rely on camera-based data or attaching sensors to the throat. Camera-based solutions require 5.82 times larger power consumption or have potential privacy issues; attaching sensors to the throat is not practical for commercial-off-the-shell (COTS) devices because additional sensors are required. In this paper, we propose a sensing technique that only needs a microphone and a speaker on COTS devices, which not only consumes little power but also has fewer privacy concerns. By deconstructing the received acoustic signals, a 2D motion profile can be generated. We propose a classifier based on convolutional neural networks (CNN) to identify the corresponding silent command from the 2D motion profiles. The proposed classifier can adapt to users and is robust when tested by environmental factors. Our evaluation shows that the system achieves 92.5% accuracy in classifying 20 commands.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 1b17fcb4-5960-42c5-a12e-eacc1f9a2c7eCited by top-tier papers7
- mSilent: Towards General Corpus Silent Speech Recognition Using COTS mmWave RadarShang Zeng, Haoran Wan, Shuyu Shi, Wei WangUbiComp 2023 · 34 citations
- EyeEcho: Continuous and Low-power Facial Expression Tracking on GlassesKe Li, Ruidong Zhang, Siyuan Chen, Boao Chen et al.CHI 2024 · 27 citations
- Underwater messaging using mobile devicesTuochao Chen, Justin Chan, Shyamnath GollakotaSIGCOMM 2022 · 23 citations
- Enabling Voice-Accompanying Hand-to-Face Gesture Recognition with Cross-Device SensingZisu Li, Chen Liang, Yuntao Wang, Yue Qin et al.CHI 2023 · 18 citations
- EyeGesener: Eye Gesture Listener for Smart Glasses Interaction Using Acoustic SensingTao Sun, Yankai Zhao, Wentao Xie, Jiao Li et al.UbiComp 2024 · 15 citations
Related papers
- EarCommand: "Hearing" Your Silent Speech Commands In EarYincheng Jin, Yang Gao, Xuhai Xu, Seokmin Choi et al.UbiComp 2022 · 37 citations
- Watch Your Mouth: Silent Speech Recognition with Depth SensingXue Wang, Zixiong Su, Jun Rekimoto, Yang ZhangCHI 2024 · 19 citations
- SoundLip: Enabling Word and Sentence-level Lip Interaction for Smart DevicesQian Zhang, Dong Wang, Run Zhao, Yinggang YuUbiComp 2021 · 34 citations
- EchoSpeech: Continuous Silent Speech Recognition on Minimally-obtrusive Eyewear Powered by Acoustic SensingRuidong Zhang, Ke Li, Yihong Hao, Yufan Wang et al.CHI 2023 · 53 citations
- ReHEarSSE: Recognizing Hidden-in-the-Ear Silently Spelled ExpressionsXuefu Dong, Yifei Chen, Yuuki Nishiyama, Kaoru Sezaki et al.CHI 2024 · 20 citations
