UltraSpeech: Speech Enhancement by Interaction between Ultrasound and Speech
Han Ding, Yizhan Wang, Hao Li, Cui Zhao, Ge Wang, Wei Xi, Jizhong Zhao
摘要
Speech enhancement can benefit lots of practical voice-based interaction applications, where the goal is to generate clean speech from noisy ambient conditions. This paper presents a practical design, namely UltraSpeech, to enhance speech by exploring the correlation between the ultrasound (profiled articulatory gestures) and speech. UltraSpeech uses a commodity smartphone to emit the ultrasound and collect the composed acoustic signal for analysis. We design a complex masking framework to deal with complex-valued spectrograms, incorporating the magnitude and phase rectification of speech simultaneously. We further introduce an interaction module to share information between ultrasound and speech two branches and thus enhance their discrimination capabilities. Extensive experiments demonstrate that UltraSpeech increases the Scale Invariant SDR by 12dB, improves the speech intelligibility and quality effectively, and is capable to generalize to unknown speakers.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- We Can Hear You with mmWave Radar! An End-to-End Eavesdropping SystemDachao Han, Teng Huang, Han Ding, Cui Zhao 等UbiComp 2026 · 被引用 5 次
- AdaStreamLite: Environment-adaptive Streaming Speech Recognition on Mobile DevicesYuheng Wei, Jie Xiong, Hui Liu, Yingtao Yu 等UbiComp 2024 · 被引用 5 次
- FeelWave: Enabling Emotion-Aware Voice Interaction through Noise-Robust mmWave Emotion SensingLingyu Wang, You Zuo, Dequan Wang, Chenming He 等CHI 2026 · 被引用 1 次
- USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal SynthesisLuca Jiang-Tao Yu, Running Zhao, Sijie Ji, Edith C. H. Ngai 等UbiComp 2025 · 被引用 1 次
相关 Paper
- UltraSE: single-channel speech enhancement using ultrasoundKe Sun, Xinyu ZhangMobiCom 2021 · 被引用 69 次
- Sensing to Hear through Memory: Ultrasound Speech Enhancement without Real Ultrasound SignalsQian Zhang, Ke Liu, Dong WangUbiComp 2024 · 被引用 5 次
- EarSE: Bringing Robust Speech Enhancement to COTS HeadphonesDi Duan, Yongliang Chen, Weitao Xu, Tianxing LiUbiComp 2024 · 被引用 14 次
- ClearSpeech: Improving Voice Quality of Earbuds Using Both In-Ear and Out-Ear MicrophonesDong Ma, Ting Dang, Ming Ding, Rajesh BalanUbiComp 2024 · 被引用 5 次
- ExpresSense: Exploring a Standalone Smartphone to Sense Engagement of Users from Facial Expressions Using Acoustic SensingPragma Kar, Shyamvanshikumar Singh, Avijit Mandal, Samiran Chattopadhyay 等CHI 2023 · 被引用 7 次
