Towards Voice Reconstruction from EEG during Imagined Speech
Young-Eun Lee, Seo-Hyun Lee, Sang-Ho Kim, Seong-Whan Lee
摘要
Translating imagined speech from human brain activity into voice is a challenging and absorbing research issue that can provide new means of human communication via brain signals. Endeavors toward reconstructing speech from brain activity have shown their potential using invasive measures of spoken speech data, however, have faced challenges in reconstructing imagined speech. In this paper, we propose Neu-roTalk, which converts non-invasive brain signals of imagined speech into the user's own voice. Our model was trained with spoken speech EEG which was generalized to adapt to the domain of imagined speech, thus allowing natural correspondence between the imagined speech and the voice as a ground truth. In our framework, automatic speech recognition decoder contributed to decomposing the phonemes of generated speech, thereby displaying the potential of voice reconstruction from unseen words. Our results imply the potential of speech synthesis from human EEG signals, not only from spoken speech but also from the brain signals of imagined speech.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- DMF2Mel: A Dynamic Multiscale Fusion Network for EEG-Driven Mel Spectrogram ReconstructionCunhang Fan, Sheng Zhang, Jingjing Zhang, Enrui Liu 等ACM MM 2025 · 被引用 4 次
- SM-Former: Spiking Symmetric Mixing Branchformer for Brain Auditory Attention DetectionJiaqi Wang, Zhengyu Ma, Xiongri Shen, Chenlin Zhou 等NeurIPS 2025 · 被引用 3 次
- Let EEG Models Learn EEGYifan Wang, Yijia Ma, Wen Li, Chenyu YouICML 2026
它引用的顶会 Paper4
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech SynthesisJungil Kong, Jaehyeon Kim, Jaekyoung BaeNeurIPS 2020 · 被引用 2,890 次
- Open Vocabulary Electroencephalography-to-Text Decoding and Zero-Shot Sentiment ClassificationZhenhailong Wang, Heng JiAAAI 2022 · 被引用 122 次
- Digital Voicing of Silent SpeechDavid Gaddy, Dan KleinEMNLP 2020 · 被引用 55 次
相关 Paper
- EEG2Video: Towards Decoding Dynamic Visual Perception from EEG SignalsXuan-Hao Liu, Yan-Kai Liu, Yansen Wang, Kan Ren 等NeurIPS 2024 · 被引用 59 次
- Mind the State: Towards Unified, Context-Aware EEG-to-fMRI SynthesisYamin Li, Shiyu Wang, Chang Li, Ange Lou 等ICML 2026
- MINDEV: Multi-modal Integrated Diffusion Framework for Video Reconstruction from EEG SignalsShuai Huang, Yongxiong Wang, Huan Luo, Haodong Jing 等ACM MM 2025 · 被引用 1 次
- NeuroBOLT: Resting-state EEG-to-fMRI Synthesis with Multi-dimensional Feature MappingYamin Li, Ange Lou, Ziyuan Xu, Shengchao Zhang 等NeurIPS 2024 · 被引用 22 次
- EVOKE: Efficient and High-Fidelity EEG-to-Video Reconstruction via Decoupling Implicit Neural RepresentationHaodong Jing, Panqi Yang, Dongyao Jiang, Zhipeng Liu 等AAAI 2026 · 被引用 1 次
