We Can Hear You with mmWave Radar! An End-to-End Eavesdropping System
Dachao Han, Teng Huang, Han Ding, Cui Zhao, Fei Wang, Ge Wang, Wei Xi
Abstract
With the rise of voice-enabled technologies, loudspeaker playback has become widespread, posing increasing risks to speech privacy. Traditional eavesdropping methods often require invasive access or line-of-sight, limiting their practicality. In this paper, we present mmSpeech, an end-to-end mmWave-based eavesdropping system that reconstructs intelligible speech solely from vibration signals induced by loudspeaker playback, even through walls and without prior knowledge of the speaker. To achieve this, we reveal an optimal combination of vibrating material and radar sampling rate for capturing highquality vibrations using narrowband mmWave signals. We then design a deep neural network that reconstructs intelligible speech from the estimated noisy spectrograms. To further support downstream speech understanding, we introduce a synthetic training pipeline and selectively fine-tune the encoder of a pre-trained ASR model. We implement mmSpeech with a commercial mmWave radar and validate its performance through extensive experiments. Results show that mmSpeech achieves state-of-the-art speech quality and generalizes well across unseen speakers and various conditions.
CCS Concepts: • Human-centered computing → Ubiquitous and mobile computing systems and tools.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7024c2ec-cd87-4284-823d-289defff3e0cBuilds on20
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
- HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech SynthesisJungil Kong, Jaehyeon Kim, Jaekyoung BaeNeurIPS 2020 · 2,890 citations
- An Exponential Learning Rate Schedule for Deep LearningZhiyuan Li, Sanjeev AroraICLR 2020 · 267 citations
- TouchPass: towards behavior-irrelevant on-touch user authentication on smartphones leveraging vibrationsXiangyu Xu, Jiadi Yu, Yingying Chen, Qin Hua et al.MobiCom 2020 · 101 citations
- Towards Generalized mmWave-based Human Pose Estimation through Signal AugmentationHongfei Xue, Qiming Cao, Chenglin Miao, Yan Ju et al.MobiCom 2023 · 69 citations
Related papers
- Privacy Leakage via Speech-induced Vibrations on Room Objects through Remote Sensing based on Phased-MIMOCong Shi, Tianfang Zhang, Zhaoyi Xu, Shuping Li et al.CCS 2023 · 10 citations
- MILLIEAR: Millimeter-wave Acoustic Eavesdropping with Unconstrained VocabularyPengfei Hu, Yifan Ma, Panneer Selvam Santhalingam, Parth H. Pathak et al.INFOCOM 2022 · 60 citations
- mmSpy: Spying Phone Calls using mmWave RadarsSuryoday Basak, Mahanth GowdaS&P 2022 · 57 citations
- mmEar: Push the Limit of COTS mmWave Eavesdropping on HeadphonesXiangyu Xu, Yu Chen, Zhen Ling, Li Lu et al.INFOCOM 2024 · 9 citations
- LAM-assisted Acoustic Eavesdropping in Multi-speaker Scenarios via Commercial mmWave RadarGuodong Liu, Lei Wang, Minjun Jiang, Qianran Qiao et al.UbiComp 2026
