EmotionKD: A Cross-Modal Knowledge Distillation Framework for Emotion Recognition Based on Physiological Signals
Yucheng Liu, Ziyu Jia, Haichao Wang
Abstract
Emotion recognition using multi-modal physiological signals is an emerging field in affective computing that significantly improves performance compared to unimodal approaches. The combination of Electroencephalogram(EEG) and Galvanic Skin Response(GSR) signals are particularly effective for objective and complementary emotion recognition. However, the high cost and inconvenience of EEG signal acquisition severely hinder the popularity of multi-modal emotion recognition in real-world scenarios, while GSR signals are easier to obtain. To address this challenge, we propose EmotionKD, a framework for cross-modal knowledge distillation that simultaneously models the heterogeneity and interactivity of GSR and EEG signals under a unified framework. By using knowledge distillation, fully fused multi-modal features can be transferred to an unimodal GSR model to improve performance. Additionally, an adaptive feedback mechanism is proposed to enable the multi-modal model to dynamically adjust according to the performance of the unimodal model during knowledge distillation, which guides the unimodal model to enhance its performance in emotion recognition. Our experiment results demonstrate that the proposed model achieves state-of-the-art performance on two public datasets. Furthermore, our approach has the potential to reduce reliance on multi-modal data with lower sacrificed performance, making emotion recognition more applicable and feasible. The source code is available at https://github.com/YuchengLiu-Alex/EmotionKD
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get c6c042e0-b0a6-4155-9fc7-3d09633fbf84Cited by top-tier papers4
- Revisiting Multimodal Emotion Recognition in Conversation from the Perspective of Graph SpectrumWei Ai, Fuchen Zhang, Yuntao Shou, Tao Meng et al.AAAI 2025 · 64 citations
- Distilling Cross-Modal Knowledge via Feature DisentanglementJunhong Liu, Yuan Zhang, Tao Huang, Wenchao Xu et al.AAAI 2026 · 2 citations
- SleepSMC: Ubiquitous Sleep Staging via Supervised Multimodal CoordinationShuo Ma, Yingwei Zhang, Yiqiang Chen, Hualei Wang et al.ICLR 2025
- Multi-modal Medical Diagnosis via Large-small Model CollaborationWanyi Chen, Zihua Zhao, Jiangchao Yao, Ya Zhang et al.CVPR 2025
Related papers
- Decoupled Multimodal Distilling for Emotion RecognitionYong Li, Yuanzhi Wang, Zhen CuiCVPR 2023
- Correlation-Driven Multi-Modality Graph Decomposition for Cross-Subject Emotion RecognitionWuliang Huang, Yiqiang Chen, Xinlong Jiang, Chenlong Gao et al.ACM MM 2024 · 2 citations
- HeLo: Heterogeneous Multi-Modal Fusion with Label Correlation for Emotion Distribution LearningChuhang Zheng, Chunwei Tian, Jie Wen, Daoqiang Zhang et al.ACM MM 2025 · 13 citations
- VBH-GNN: Variational Bayesian Heterogeneous Graph Neural Networks for Cross-subject Emotion RecognitionChenyu Liu, Xinliang Zhou, Zhengri Zhu, Liming Zhai et al.ICLR 2024 · 25 citations
- EMOE: Modality-Specific Enhanced Dynamic Emotion ExpertsYiyang Fang, Wenke Huang, Guancheng Wan, Kehua Su et al.CVPR 2025
