ProxiMic: Convenient Voice Activation via Close-to-Mic Speech Detected by a Single Microphone
Yue Qin, Chun Yu, Zhaoheng Li, Mingyuan Zhong, Yukang Yan, Yuanchun Shi
Abstract
Wake-up-free techniques (e.g., Raise-to-Speak) are important for improving the voice input experience. We present ProxiMic, a close-to-mic (within 5 cm) speech sensing technique using only one microphone. With ProxiMic, a user keeps a microphone-embedded device close to the mouth and speaks directly to the device without wake-up phrases or button presses. To detect close-to-mic speech, we use the feature from pop noise observed when a user speaks and blows air onto the microphone. Sound input is first passed through a low-pass adaptive threshold filter, then analyzed by a CNN which detects subtle close-to-mic features (mainly pop noise). Our two-stage algorithm can achieve 94.1% activation recall, 12.3 False Accepts per Week per User (FAWU) with 68 KB memory size, which can run at 352 fps on the smartphone. The user study shows that ProxiMic is efficient, user-friendly, and practical.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 438297e8-fa91-44a4-b00e-122d19a8e366Cited by top-tier papers5
- DRG-Keyboard: Enabling Subtle Gesture Typing on the Fingertip with Dual IMU RingsChen Liang, Chi Hsia, Chun Yu, Yukang Yan et al.UbiComp 2023 · 33 citations
- PrISM-Tracker: A Framework for Multimodal Procedure Tracking Using Wearable Sensors and State Transition Information with User-Driven Handling of Errors and UncertaintyRiku Arakawa, Hiromu Yakura, Vimal Mollyn, Suzanne Nie et al.UbiComp 2023 · 20 citations
- Enabling Voice-Accompanying Hand-to-Face Gesture Recognition with Cross-Device SensingZisu Li, Chen Liang, Yuntao Wang, Yue Qin et al.CHI 2023 · 18 citations
- DualVoice: Speech Interaction that Discriminates between Normal and Whispered Voice InputJun RekimotoUIST 2022 · 9 citations
- AuthGlass: Benchmarking Voice Liveness Detection and Authentication on Smart Glasses via Comprehensive Acoustic FeaturesWeiye Xu, Zhang Jiang, Siqi Zheng, Xiyuxing Zhang et al.UbiComp 2026
Builds on4
- EchoWhisper: Exploring an Acoustic-based Silent Speech Interface for Smartphone UsersYang Gao, Yincheng Jin, Jiyang Li, Seokmin Choi et al.UbiComp 2020 · 44 citations
- FaceSight: Enabling Hand-to-Face Gesture Interaction on AR Glasses with a Downward-Facing Camera VisionYueting Weng, Chun Yu, Yingtian Shi, Yuhang Zhao et al.CHI 2021 · 39 citations
- PenSight: Enhanced Interaction with a Pen-Top CameraFabrice Matulic, Riku Arakawa, Brian K. Vogel, Daniel VogelCHI 2020 · 33 citations
- FrownOnError: Interrupting Responses from Smart Speakers by Facial ExpressionsYukang Yan, Chun Yu, Wengrui Zheng, Ruining Tang et al.CHI 2020 · 24 citations
Related papers
- ShadowTouch: Enabling Free-Form Touch-Based Hand-to-Surface Interaction with Wrist-Mounted Illuminant by Shadow ProjectionChen Liang, Xutong Wang, Zisu Li, Chi Hsia et al.UIST 2023 · 14 citations
- Endophasia: Utilizing Acoustic-Based Imaging for Issuing Contact-Free Silent Speech CommandsYongzhao Zhang, Wei-Hsiang Huang, Chih-Yun Yang, Wen-Ping Wang et al.UbiComp 2020 · 43 citations
- MorsEar: Toward Generalizable Low-Resource Covert Messaging via Earable based Inertial SensingGarvit Chugh, Indrajeet Ghosh, Nirmalya Roy, Sandip Chakraborty et al.CHI 2026 · 1 citation
- Watch Your Mouth: Silent Speech Recognition with Depth SensingXue Wang, Zixiong Su, Jun Rekimoto, Yang ZhangCHI 2024 · 19 citations
- EchoSpeech: Continuous Silent Speech Recognition on Minimally-obtrusive Eyewear Powered by Acoustic SensingRuidong Zhang, Ke Li, Yihong Hao, Yufan Wang et al.CHI 2023 · 53 citations
