Multi-dataset Joint Pre-training of Emotional EEG Enables Generalizable Affective Computing
Qingzhu Zhang, Jiani Zhong, Zongsheng Li, Xinke Shen, Quanying Liu
Abstract
Task-specific pre-training is essential when task representations diverge from generic pre-training features. Existing task-general pre-training EEG models struggle with complex tasks like emotion recognition due to mismatches between task-specific features and broad pre-training approaches. This work aims to develop a task-specific multi-dataset joint pre-training framework for cross-dataset emotion recognition, tackling problems of large inter-dataset distribution shifts, inconsistent emotion category definitions, and substantial inter-subject variability. We introduce a cross-dataset covariance alignment loss to align second-order statistical properties across datasets, enabling robust generalization without the need for extensive labels or per-subject calibration. To capture the long-term dependency and complex dynamics of EEG, we propose a hybrid encoder combining a Mamba-like linear attention channel encoder and a spatiotemporal dynamics model. Our method outperforms state-of-the-art large-scale EEG models by an average of 4.57% in AUROC for few-shot emotion recognition and 11.92% in accuracy for zero-shot generalization to a new dataset. Performance scales with the increase of datasets used in pre-training. Multi-dataset joint pre-training achieves a performance gain of 8.55% over single-dataset training. This work provides a scalable framework for task-specific pre-training and highlights its benefit in generalizable affective computing. Our code is available at https://github.com/ncclab-sustech/ mdJPT_nips2025 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b9aecb9f-a4d3-4374-87c6-41420d902c3aCited by top-tier papers2
- Harnessing Spectrum Video for Subject-Level Few-Shot and Cross-Montage EEG GeneralizationWei Wang, Fang He, Yifan Li, Wanying Qu et al.ICML 2026
- RECTOR: Masked Region-Channel-Temporal Modeling for Affective and Cognitive Representation LearningJinhan Liu, Mahsa ShoaranICML 2026
Builds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- BIOT: Biosignal Transformer for Cross-data Learning in the WildChaoqi Yang, M. Brandon Westover, Jimeng SunNeurIPS 2023 · 345 citations
- Large Brain Model for Learning Generic Representations with Tremendous EEG Data in BCIWei-Bang Jiang, Li-Ming Zhao, Bao-Liang LuICLR 2024 · 298 citations
- Demystify Mamba in Vision: A Linear Attention PerspectiveDongchen Han, Ziyi Wang, Zhuofan Xia, Yizeng Han et al.NeurIPS 2024 · 287 citations
Related papers
- State Mamba: Spatiotemporal EEG State-Space Model with Dynamic Brain Alignment for Cross-Subject RepresentationWeining Weng, Yang Gu, Yuan Ma, Yuchen Liu et al.AAAI 2026
- DHCM-CACL: Dynamic Hierarchical Cross-modal Mamba with Confidence-Adaptive Contrastive Learning for Multimodal Emotion RecognitionBaiqiang Wu, Yang LiAAAI 2026 · 1 citation
- Sera: Separated Coarse-to-fine Representation Alignment for Cross-subject EEG-based Emotion RecognitionZhihao Jia, Meiyan Xu, Jingyuan Wang, Ziyu Jia et al.ACM MM 2025 · 2 citations
- VBH-GNN: Variational Bayesian Heterogeneous Graph Neural Networks for Cross-subject Emotion RecognitionChenyu Liu, Xinliang Zhou, Zhengri Zhu, Liming Zhai et al.ICLR 2024 · 25 citations
- Learning Topology-Agnostic EEG Representations with Geometry-Aware ModelingKe Yi, Yansen Wang, Kan Ren, Dongsheng LiNeurIPS 2023 · 99 citations
