Learning Unseen Emotions from Gestures via Semantically-Conditioned Zero-Shot Perception with Adversarial Autoencoders
Abhishek Banerjee, Uttaran Bhattacharya, Aniket Bera
Abstract
We present a novel generalized zero-shot algorithm to recognize perceived emotions from gestures. Our task is to map gestures to novel emotion categories not encountered in training. We introduce an adversarial autoencoder-based representation learning that correlates 3D motion-captured gesture sequences with the vectorized representation of the naturallanguage perceived emotion terms using word2vec embeddings. The language-semantic embedding provides a representation of the emotion label space, and we leverage this underlying distribution to map the gesture sequences to the appropriate categorical emotion labels. We train our method using a combination of gestures annotated with known emotion terms and gestures not annotated with any emotions. We evaluate our method on the MPI Emotional Body Expressions Database (EBEDB) and obtain an accuracy of 58.43% . We see an improvement in performance compared to current state-of-the-art algorithms for generalized zero-shot learning by 25-27% on the absolute. We also demonstrate our approach on publicly available videos from the internet and movie scenes, where the actors' pose has been extracted and map to their respective emotive states.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b62b0fd2-52ce-485a-8682-945a1478d141Cited by top-tier papers4
- Text2Gestures: A Transformer-Based Network for Generating Emotive Body Gestures for Virtual Agents**This work has been supported in part by ARO Grants W911NF1910069 and W911NF1910315, and Intel. Code and additional materials available at: https: //gamma.umd.edu/t2gUttaran Bhattacharya, Nicholas Rewkowski, Abhishek Banerjee, Pooja Guhan et al.IEEE VR 2021 · 147 citations
- Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression LearningUttaran Bhattacharya, Elizabeth Childs, Nicholas Rewkowski, Dinesh ManochaACM MM 2021 · 94 citations
- Which Demographics do LLMs Default to During Annotation?Johannes Schäfer, Aidan Combs, Christopher Bagdon, Jiahui Li et al.ACL 2025 · 11 citations
- Enhance Stealthiness and Transferability of Adversarial Attacks with Class Activation Mapping Ensemble AttackHui Xia, Rui Zhang, Zi Kang, Shuliang Jiang et al.NDSS 2024
Builds on1
Related papers
- iMiGUE: An Identity-Free Video Dataset for Micro-Gesture Understanding and Emotion AnalysisXin Liu, Henglin Shi, Haoyu Chen, Zitong Yu et al.CVPR 2021
- Zero-Shot Learning for IMU-Based Activity Recognition Using Video EmbeddingsCatherine Tong, Jinchen Ge, Nicholas D. LaneUbiComp 2022 · 39 citations
- A Variational Autoencoder with Deep Embedding Model for Generalized Zero-Shot LearningPeirong Ma, Xiao HuAAAI 2020 · 43 citations
- Reading Your Actions: Learning Generalizable Action Representations via Pre-training AEMGZhenghao Huang, Huilin Yao, Kaikai Wang, Lin ShuCVPR 2026
- Discovering Human Interactions With Novel Objects via Zero-Shot LearningSuchen Wang, Kim-Hui Yap, Junsong Yuan, Yap-Peng TanCVPR 2020
