ProxiMic: Convenient Voice Activation via Close-to-Mic Speech Detected by a Single Microphone
Yue Qin, Chun Yu, Zhaoheng Li, Mingyuan Zhong, Yukang Yan, Yuanchun Shi
摘要
Wake-up-free techniques (e.g., Raise-to-Speak) are important for improving the voice input experience. We present ProxiMic, a close-to-mic (within 5 cm) speech sensing technique using only one microphone. With ProxiMic, a user keeps a microphone-embedded device close to the mouth and speaks directly to the device without wake-up phrases or button presses. To detect close-to-mic speech, we use the feature from pop noise observed when a user speaks and blows air onto the microphone. Sound input is first passed through a low-pass adaptive threshold filter, then analyzed by a CNN which detects subtle close-to-mic features (mainly pop noise). Our two-stage algorithm can achieve 94.1% activation recall, 12.3 False Accepts per Week per User (FAWU) with 68 KB memory size, which can run at 352 fps on the smartphone. The user study shows that ProxiMic is efficient, user-friendly, and practical.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- DRG-Keyboard: Enabling Subtle Gesture Typing on the Fingertip with Dual IMU RingsChen Liang, Chi Hsia, Chun Yu, Yukang Yan 等UbiComp 2023 · 被引用 33 次
- PrISM-Tracker: A Framework for Multimodal Procedure Tracking Using Wearable Sensors and State Transition Information with User-Driven Handling of Errors and UncertaintyRiku Arakawa, Hiromu Yakura, Vimal Mollyn, Suzanne Nie 等UbiComp 2023 · 被引用 20 次
- Enabling Voice-Accompanying Hand-to-Face Gesture Recognition with Cross-Device SensingZisu Li, Chen Liang, Yuntao Wang, Yue Qin 等CHI 2023 · 被引用 18 次
- DualVoice: Speech Interaction that Discriminates between Normal and Whispered Voice InputJun RekimotoUIST 2022 · 被引用 9 次
- AuthGlass: Benchmarking Voice Liveness Detection and Authentication on Smart Glasses via Comprehensive Acoustic FeaturesWeiye Xu, Zhang Jiang, Siqi Zheng, Xiyuxing Zhang 等UbiComp 2026
它引用的顶会 Paper4
- EchoWhisper: Exploring an Acoustic-based Silent Speech Interface for Smartphone UsersYang Gao, Yincheng Jin, Jiyang Li, Seokmin Choi 等UbiComp 2020 · 被引用 44 次
- FaceSight: Enabling Hand-to-Face Gesture Interaction on AR Glasses with a Downward-Facing Camera VisionYueting Weng, Chun Yu, Yingtian Shi, Yuhang Zhao 等CHI 2021 · 被引用 39 次
- PenSight: Enhanced Interaction with a Pen-Top CameraFabrice Matulic, Riku Arakawa, Brian K. Vogel, Daniel VogelCHI 2020 · 被引用 33 次
- FrownOnError: Interrupting Responses from Smart Speakers by Facial ExpressionsYukang Yan, Chun Yu, Wengrui Zheng, Ruining Tang 等CHI 2020 · 被引用 24 次
相关 Paper
- ShadowTouch: Enabling Free-Form Touch-Based Hand-to-Surface Interaction with Wrist-Mounted Illuminant by Shadow ProjectionChen Liang, Xutong Wang, Zisu Li, Chi Hsia 等UIST 2023 · 被引用 14 次
- Endophasia: Utilizing Acoustic-Based Imaging for Issuing Contact-Free Silent Speech CommandsYongzhao Zhang, Wei-Hsiang Huang, Chih-Yun Yang, Wen-Ping Wang 等UbiComp 2020 · 被引用 43 次
- MorsEar: Toward Generalizable Low-Resource Covert Messaging via Earable based Inertial SensingGarvit Chugh, Indrajeet Ghosh, Nirmalya Roy, Sandip Chakraborty 等CHI 2026 · 被引用 1 次
- Watch Your Mouth: Silent Speech Recognition with Depth SensingXue Wang, Zixiong Su, Jun Rekimoto, Yang ZhangCHI 2024 · 被引用 19 次
- EchoSpeech: Continuous Silent Speech Recognition on Minimally-obtrusive Eyewear Powered by Acoustic SensingRuidong Zhang, Ke Li, Yihong Hao, Yufan Wang 等CHI 2023 · 被引用 53 次
