Unlocking Cross-Modal Biosignal Synthesis: A Temporally-Aware VAE-Diffusion Model
Chenyang Xu, Dezhen Wang, Hao Wang
摘要
Synthesizing authentic phonocardiograms (PCG) from ubiquitous electrocardiograms (ECG) is a critical task for accessible cardiac monitoring. Existing generative models, however, struggle to capture the heart's complex electromechanical coupling, failing to meet the dual requirements of temporal precision and physiological fidelity needed for clinically relevant waveform analysis. We introduce the Temporally-Aware VAE-Diffusion model, a synergistic hybrid architecture that resolves this trade-off. Our architecture enforces tight physiological coupling through an Enhanced Condition Fusion mechanism and explicitly models long-range cardiac dynamics via Temporal Attention Blocks. On the EPHNOGRAM benchmark, our model sets a new state-of-the-art, achieving a Pearson correlation of 0.9100.008, 95.95% S1 detection accuracy, and a precise 12.0 ms timing error, significantly outperforming leading diffusion and Transformer baselines. Crucially, our work provides a reproducible zero-shot transfer evaluation for ECG-to-PCG synthesis. Evaluated on the synchronized PhysioNet/CinC 2016 training-a/MITHSDB subset without target-domain training, our model preserves high waveform fidelity and clinically relevant timing structure under domain shift, including on pathological recordings. These results support cross-dataset robustness of the proposed synthesis framework, while downstream diagnostic validation remains an important direction for future work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 被引用 11,724 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
相关 Paper
- SE-Diff: Simulator and Experience Enhanced Diffusion Model for Comprehensive ECG GenerationXiaoda Wang, Kaiqiao Han, Yuhao Xu, Xiao Luo 等ICLR 2026 · 被引用 4 次
- Region-Disentangled Diffusion Model for High-Fidelity PPG-to-ECG TranslationDebaditya Shome, Pritam Sarkar, Ali EtemadAAAI 2024 · 被引用 42 次
- CardioGAN: Attentive Generative Adversarial Network with Dual Discriminators for Synthesis of ECG from PPGPritam Sarkar, Ali EtemadAAAI 2021 · 被引用 99 次
- EchoVDiff: Cardiac-Cycle Echocardiography Video Generation from Arbitrary FrameJiansong Zhang, Xiaying Yang, Xiaoling Luo, Linlin ShenCVPR 2026
- ME-GAN: Learning Panoptic Electrocardio Representations for Multi-view ECG Synthesis Conditioned on Heart DiseasesJintai Chen, Kuanlun Liao, Kun Wei, Haochao Ying 等ICML 2022 · 被引用 29 次
