Pose as a Modality: A Psychology-Inspired Network for Personality Recognition with a New Multimodal Dataset
Bin Tang, Keqi Pan, Miao Zheng, Ning Zhou, Jialu Sui, Dandan Zhu, Cheng-Long Deng, Shu-Guang Kuai
Abstract
In recent years, predicting Big Five personality traits from multimodal data has received significant attention in artificial intelligence (AI). However, existing computational models often fail to achieve satisfactory performance. Psychological research has shown a strong correlation between pose and personality traits, yet previous research has largely ignored pose data in computational models. To address this gap, we develop a novel multimodal dataset that incorporates full-body pose data. The dataset includes video recordings of 287 participants completing a virtual interview with 36 questions, along with self-reported Big Five personality scores as labels. To effectively utilize this multimodal data, we introduce the Psychology-Inspired Network (PINet), which consists of three key modules: Multimodal Feature Awareness (MFA), Multimodal Feature Interaction (MFI), and Psychology-Informed Modality Correlation Loss (PIMC Loss). The MFA module leverages the Vision Mamba Block to capture comprehensive visual features related to personality, while the MFI module efficiently fuses the multimodal features. The PIMC Loss, grounded in psychological theory, guides the model to emphasize different modalities for different personality dimensions. Experimental results show that the PINet outperforms several state-of-the-art baseline models. Furthermore, the three modules of PINet contribute almost equally to the model’s overall performance. Incorporating pose data significantly enhances the model’s performance, with the pose modality ranking mid-level in importance among the five modalities. These findings address the existing gap in personality-related datasets that lack full-body pose data and provide a new approach for improving the accuracy of personality prediction models, highlighting the importance of integrating psychological insights into AI frameworks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8c9bf6b2-e96b-4ccf-ac58-7d19c3e62831Cited by top-tier papers1
Ask how each one uses itBuilds on4
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and TextHassan Akbari, Liangzhe Yuan, Rui Qian, Wei-Hong Chuang et al.NeurIPS 2021 · 782 citations
- Low-Rank Bottleneck in Multi-head Attention ModelsSrinadh Bhojanapalli, Chulhee Yun, Ankit Singh Rawat, Sashank J. Reddi et al.ICML 2020 · 130 citations
- Investigating how speech and animation realism influence the perceived personality of virtual characters and agentsSean Thomas, Ylva Ferstl, Rachel McDonnell, Cathy EnnisIEEE VR 2022 · 34 citations
Related papers
- Multimodal Fine-Grained Apparent Personality Trait Recognition: Joint Modeling of Big Five and Questionnaire Item-level ScoresRyo Masumura, Shota Orihashi, Mana Ihori, Tomohiro Tanaka et al.AAAI 2025 · 3 citations
- Cross-Modal Interactive Perception Network with Mamba for Lung Tumor Segmentation in PET-CT ImagesJie Mei, Chenyu Lin, Yu Qiu, Yaonan Wang et al.CVPR 2025
- PersonalitySensing: A Multi-View Multi-Task Learning Approach for Personality Detection based on Smartphone UsageSongcheng Gao, Wenzhong Li, Lynda J. Song, Xiao Zhang et al.ACM MM 2020 · 14 citations
- PSA-MF: Personality-Sentiment Aligned Multi-Level Fusion for Multimodal Sentiment AnalysisHeng Xie, Kang Zhu, Zhengqi Wen, Jianhua Tao et al.AAAI 2026 · 1 citation
- Robust Pose Estimation in Crowded Scenes with Direct Pose-Level InferenceDongkai Wang, Shiliang Zhang, Gang HuaNeurIPS 2021 · 36 citations
