Interpretable Self-Supervised Facial Micro-Expression Learning to Predict Cognitive State and Neurological Disorders
Arun Das, Jeffrey Mock, Yufei Huang, Edward J. Golob, Peyman Najafirad
Abstract
Human behavior is the confluence of output from voluntary and involuntary motor systems. The neural activities that mediate behavior, from individual cells to distributed networks, are in a state of constant flux. Artificial intelligence (AI) research over the past decade shows that behavior, in the form of facial muscle activity, can reveal information about fleeting voluntary and involuntary motor system activity related to emotion, pain, and deception. However, the AI algorithms often lack an explanation for their decisions, and learning meaningful representations requires large datasets labeled by a subject-matter expert. Motivated by the success of using facial muscle movements to classify brain states and the importance of learning from small amounts of data, we propose an explainable self-supervised representation-learning paradigm that learns meaningful temporal facial muscle movement patterns from limited samples. We validate our methodology by carrying out comprehensive empirical study to predict future speech behavior in a real-world dataset of adults who stutter (AWS). Our explainability study found facial muscle movements around the eyes (p <0.×001) and lips (p <0.001) differ significantly before producing fluent vs. disfluent speech. Evaluations using the AWS dataset demonstrates that the proposed self-supervised approach achieves a minimum of 2.51% accuracy improvement over fully-supervised approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on4
- Auto-GAN: Self-Supervised Collaborative Learning for Medical Image SynthesisBing Cao, Han Zhang, Nannan Wang, Xinbo Gao et al.AAAI 2020 · 94 citations
- Multi-Task Self-Supervised Learning for Disfluency DetectionShaolei Wang, Wanxiang Che, Qi Liu, Pengda Qin et al.AAAI 2020 · 56 citations
- Revisiting Image Aesthetic Assessment via Self-Supervised Feature LearningKekai Sheng, Weiming Dong, Menglei Chai, Guohui Wang et al.AAAI 2020 · 35 citations
- Self-Supervised Learning of Video-Induced Visual InvariancesMichael Tschannen, Josip Djolonga, Marvin Ritter, Aravindh Mahendran et al.CVPR 2020
Related papers
- Towards End-to-End Explainable Facial Action Unit Recognition via Vision-Language Joint LearningXuri Ge, Junchen Fu, Fuhai Chen, Shan An et al.ACM MM 2024 · 12 citations
- Disability-First AI Dataset Annotation: Co-designing Stuttered Speech Annotation Guidelines with People Who StutterXinru Tang, Jingjin Li, Shaomei WuCHI 2026 · 3 citations
- Knowledge-Driven Self-Supervised Representation Learning for Facial Action Unit RecognitionYanan Chang, Shangfei WangCVPR 2022 · 38 citations
- emg2speech: synthesizing speech from electromyography using self-supervised speech modelsHarshavardhana T. Gowda, Daniel C. Comstock, Lee M. MillerACL 2026 · 2 citations
- Facial-R1: Aligning Reasoning and Recognition for Facial Emotion AnalysisJiulong Wu, Yucheng Shen, Lingyong Yan, Haixin Sun et al.AAAI 2026 · 3 citations
