Signal Enhancement via Multi-view Dynamic Representation and Alignment-aware Fusion
Zikun Jin, Yuhua Qian, Xinyan Liang, Jiaqian Zhang, Jinpeng Yuan, Shen Hu, Haijun Geng, Honghong Cheng
Abstract
Robust signal enhancement under non-stationary and low SNR conditions remains challenging, as methods based on the short-time Fourier transform (STFT) with fixed resolution struggle to represent complex and time–frequency structures. While leveraging the fractional domain as an auxiliary view offers flexibility in modeling time-frequency structures, existing methods typically adopt fixed transform orders and overlook alignment between views, hindering effective integration of complementary representations and leaving frequency domain misalignment unresolved. Therefore, we propose FracFusion, a novel framework that integrates a learnable short-time fractional Fourier Transform (STFrFT) module to generate dynamic auxiliary views, combined with two stage alignment-aware fusion modules: Pearson Channel Fusion for correlation-guided consistency and Efficient Align Fusion for fine-grained, frequency aligned interaction. Experiments on speech and electromagnetic (EM) datasets show that FracFusion consistently outperforms state-of-the-art baselines across diverse noise levels and signal types, demonstrating robust adaptability across domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- KITE: Knowledge-Guided Probabilistic Modeling for Time Series Forecasting with Exogenous VariablesHanyin Cheng, Jingrong Zhou, Yang Shu, Chenjuan GuoICML 2026
- TeamWork: Multivariate Time Series Anomaly Detection via Asymmetric Role-aware Channel ModelingShiyan Hu, Tengxue Zhang, Jianxin Jin, Xiangfei Qiu et al.ICML 2026
- Robust Signal Enhancement via Fractional Detail Views and Knowledge Guided Multi-view FusionZikun Jin, Yuhua Qian, Xinyan Liang, Jiaqian Zhang et al.ICML 2026
Builds on4
- MetricGAN-OKD: Multi-Metric Optimization of MetricGAN via Online Knowledge Distillation for Speech EnhancementWooseok Shin, Byung Hoon Lee, Jin Sob Kim, Hyun Joon Park et al.ICML 2023 · 18 citations
- Time-Frequency Domain Fusion Enhancement for Audio Super-ResolutionYe Tian, Zhe Wang, Jianguo Sun, Liguo ZhangACM MM 2024 · 1 citation
- Trusted Multi-View Classification with Expert Knowledge ConstraintsXinyan Liang, Shijie Wang, Yuhua Qian, Qian Guo et al.ICML 2025
- BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech EnhancementCunhang Fan, Enrui Liu, Andong Li, Jianhua Tao et al.AAAI 2025
Related papers
- Learnable Fractional Superlets with a Spectro-Temporal Emotion Encoder for Speech Emotion RecognitionAlaa Nfissi, Wassim Bouachir, Nizar Bouguila, Brian L MisharaICLR 2026
- Towards Non-Stationary Time Series Forecasting with Temporal Stabilization and Frequency DifferencingJunkai Lu, Peng Chen, Chenjuan Guo, Yang Shu et al.AAAI 2026 · 1 citation
- RTFS-Net: Recurrent Time-Frequency Modelling for Efficient Audio-Visual Speech SeparationSamuel Pegg, Kai Li, Xiaolin HuICLR 2024 · 13 citations
- Semantic-Adaptive Diffusion for Dynamic Spatiotemporal FusionJinsong Zhang, Ying Qu, Yuan Liao, Hairong Qi et al.CVPR 2026
- Reference-Based Speech Enhancement via Feature Alignment and Fusion NetworkHuanjing Yue, Wenxin Duo, Xiulian Peng, Jingyu YangAAAI 2022 · 19 citations
