USENIX Security2025Top-tier venue
EchoLLM: LLM-Augmented Acoustic Eavesdropping Attack on Bone Conduction Headphones with mmWave Radar
Xin Yao, Kecheng Huang, Yimin Chen, Jiawei Guo, Jie Tang, Ming Zhao
Abstract
Bone conduction headphones have gained popularity due to their comfort and versatility, as they transmit sound via vibrations through the user's skull bones rather than the eardrums. However, the transmitted audio may contain sensitive information, posing serious privacy risks if intercepted by unauthorized parties. While prior research has explored acoustic eavesdropping attacks via side channels such as millimeter wave (mmWave) radar, existing methods remain constrained by limited generalizability and degraded reconstruction quality. In this paper, we propose EchoLLM, the first mmWave-based eavesdropping attack specifically targeting the semantic content of audio transmitted through a victim's bone conduction headphones. EchoLLM exploits the fact that audio signals in bone conduction headphones are transmitted as mechanical vibrations, which can be captured as a side channel. The attack is designed to be context-aware, target-aware, low-cost, and robust. To improve automatic speech recognition (ASR) under weak mmWave signals, EchoLLM introduces a novel multi-modal speech recognition model that leverages the victim's own speech as contextual input for better ASR accuracy. To provide higher-quality input to the multi-modal model, EchoLLM incorporates signal enhancement techniques such as target identification and background reflection reduction. Our experiments show that EchoLLM achieves a word error rate (WER) as low as 5.23%, and confirm that it can effectively reconstruct the content from the audio transmitted through a targeted bone conduction headphone under realistic scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on16
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 9,451 citations
- Robust Speech Recognition via Large-Scale Weak SupervisionAlec Radford, Jong Wook Kim, Tao Xu, Greg Brockman et al.ICML 2023 · 6,966 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- mmVib: micrometer-level vibration measurement with mmwave radarChengkun Jiang, Junchen Guo, Yuan He, Meng Jin et al.MobiCom 2020 · 154 citations
- Speechless: Analyzing the Threat to Speech Privacy from Smartphone Motion SensorsS. Abhishek Anand, Nitesh SaxenaS&P 2018 · 110 citations
Related papers
- mmEar: Push the Limit of COTS mmWave Eavesdropping on HeadphonesXiangyu Xu, Yu Chen, Zhen Ling, Li Lu et al.INFOCOM 2024 · 9 citations
- mmEcho: A mmWave-based Acoustic Eavesdropping MethodPengfei Hu, Wenhao Li, Riccardo Spolaor, Xiuzhen ChengS&P 2023
- MILLIEAR: Millimeter-wave Acoustic Eavesdropping with Unconstrained VocabularyPengfei Hu, Yifan Ma, Panneer Selvam Santhalingam, Parth H. Pathak et al.INFOCOM 2022 · 60 citations
- LAM-assisted Acoustic Eavesdropping in Multi-speaker Scenarios via Commercial mmWave RadarGuodong Liu, Lei Wang, Minjun Jiang, Qianran Qiao et al.UbiComp 2026
- We Can Hear You with mmWave Radar! An End-to-End Eavesdropping SystemDachao Han, Teng Huang, Han Ding, Cui Zhao et al.UbiComp 2026 · 5 citations
