Simplifying Multimodal Emotion Recognition with Single Eye Movement Modality
Xu Yan, Li-Ming Zhao, Bao-Liang Lu
Abstract
Multimodal emotion recognition has long been a popular topic in affective computing since it significantly enhances the performance compared with that of a single modality. Among all, the combination of electroencephalography (EEG) and eye movement signals is one of the most attractive practices due to their complementarity and objectivity. However, the high cost and inconvenience of EEG signal acquisition severely hamper the popularization of multimodal emotion recognition in practical scenarios, while eye movement signals are much easier to acquire. To increase the feasibility and the generalization ability of emotion decoding without compromising the performance, we propose a generative adversarial network-based framework. In our model, a single modality of eye movements is used as input and it is capable of mapping the information onto multimodal features. Experimental results on SEED series datasets with different emotion categories demonstrate that our model with multimodal features generated by the single eye movement modality maintains competitive accuracies compared to those with multimodality input and drastically outperforms those single-modal emotion classifiers. This illustrates that the model has the potential to reduce the dependence on multimodalities without sacrificing performance which makes emotion recognition more applicable and practicable.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1f99ce6-0d68-498d-b972-c81900b417fbCited by top-tier papers2
- Counterfactual Reasoning for Out-of-distribution Multimodal Sentiment AnalysisTeng Sun, Wenjie Wang, Liqiang Jing, Yiran Cui et al.ACM MM 2022 · 65 citations
- Multi-to-Single: Reducing Multimodal Dependency in Emotion Recognition Through Contrastive LearningYan-Kai Liu, Jinyu Cai, Bao-Liang Lu, Wei-Long ZhengAAAI 2025 · 6 citations
Builds on3
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Context-Aware Emotion Recognition NetworksJiyoung Lee, Seungryong Kim, Sunok Kim, Jungin Park et al.ICCV 2019 · 285 citations
- L-Verse: Bidirectional Generation Between Image and TextTaehoon Kim, Gwangmo Song, Sihaeng Lee, Sangyun Kim et al.CVPR 2022 · 23 citations
Related papers
- Multimodal Adaptive Emotion Transformer with Flexible Modality Inputs on A Novel Dataset with Continuous LabelsWei-Bang Jiang, Xuan-Hao Liu, Wei-Long Zheng, Bao-Liang LuACM MM 2023 · 44 citations
- EmotionKD: A Cross-Modal Knowledge Distillation Framework for Emotion Recognition Based on Physiological SignalsYucheng Liu, Ziyu Jia, Haichao WangACM MM 2023 · 55 citations
- A Multimodal EEG-Eye Movement Model for Automatic Depression DetectionHao-Long Yin, Jian-Ming Zhang, Ren-Jie Dai, Wei-Long Zheng et al.AAAI 2026
- A Multi-Domain Adaptive Graph Convolutional Network for EEG-based Emotion RecognitionRui Li, Yiting Wang, Bao-Liang LuACM MM 2021 · 70 citations
- Graph to Grid: Learning Deep Representations for Multimodal Emotion RecognitionMing Jin, Jinpeng LiACM MM 2023 · 18 citations
