Baro2Talk: Reconstructing Spectrograms from Ear Canal Pressure for Voice-free Communication
Luo Zhou, Shan Chang, Han Wang, Xianbo Wang, Hongzi Zhu
摘要
The increasing demand for private and noise-resilient speech interaction has motivated the development of Silent Speech Interfaces (SSIs) that infer user intent without vocalization. Existing SSI solutions face limitations such as intrusiveness, privacy leakage, environmental sensitivity, deployment complexity, and motion vulnerability. In this work, we present Baro2Talk, a wearable SSI system that reconstructs speech content from TMJ-dominated Pressure Variation Sequences (TPVSs) captured by miniature barometers embedded in standard earbuds. Baro2Talk is inspired by two key observations. First, silent articulation induces consistent ear canal deformation via temporomandibular joint (TMJ) movements, producing pressure fluctuations that reflect articulatory patterns associated with speech. Second, TPVSs exhibit repeatable temporal and articulatory structures within phrases, offering structured signals to support semantic modeling without acoustic input. We develop a lightweight in-ear pressure sensing prototype and propose a set of modules that first perform articulatory event detection and generalization enhancement, followed by a three-stage reconstruction pipeline: Semantic Encoding, Coarse Mel-spectrogram Construction, and Phonetic Enhancement. The resulting spectrograms are decoded into text using a pre-trained automatic speech recognition (ASR) model (e.g., Whisper). Baro2Talk achieves a 6.5% CER, 9.9% WER, and a 0.081 spectral convergence score, demonstrating robust performance in silent, mobile, and noisy environments.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- EarCommand: "Hearing" Your Silent Speech Commands In EarYincheng Jin, Yang Gao, Xuhai Xu, Seokmin Choi 等UbiComp 2022 · 被引用 37 次
- EchoWhisper: Exploring an Acoustic-based Silent Speech Interface for Smartphone UsersYang Gao, Yincheng Jin, Jiyang Li, Seokmin Choi 等UbiComp 2020 · 被引用 44 次
- EchoSpeech: Continuous Silent Speech Recognition on Minimally-obtrusive Eyewear Powered by Acoustic SensingRuidong Zhang, Ke Li, Yihong Hao, Yufan Wang 等CHI 2023 · 被引用 53 次
- Lipwatch: Enabling Silent Speech Recognition on Smartwatches using Acoustic SensingQian Zhang, Yubin Lan, Kaiyi Guo, Dong WangUbiComp 2024 · 被引用 20 次
- ReHEarSSE: Recognizing Hidden-in-the-Ear Silently Spelled ExpressionsXuefu Dong, Yifei Chen, Yuuki Nishiyama, Kaoru Sezaki 等CHI 2024 · 被引用 20 次
