Debiased Multimodal Personality Understanding through Dual Causal Intervention
Yangfu Zhu, Zitong Han, Nianwen Ning, Yuting Wei, Yuandong Wang, Hang Feng, Zhenzhou Shao
摘要
Multimodal personality understanding plays a critical role in human-centered artificial intelligence. Previous work mainly focus on learning rich multimodal representations for video personality understanding. However, they often suffer from potential harm caused by subject bias (e.g., observable age and unobservable mental states), as subjects originate from diverse demographic backgrounds. Learning such spurious associations between multimodal features and traits may lead to unfair personality understanding. In this work, we construct a Structural Causal Model (SCM) to analyze the impact of these biases from a causal perspective, and propose a novel Dual Causal Adjustment Network (DCAN) to mitigate the interference of subject attributes on personality understanding. Specifically, we design a Back-door Adjustment Causal Learning (BACL) module to block spurious correlations from observable demographic factors via a prototype-based confounder dictionary, and subsequently apply a Front-door Adjustment Causal Learning (FACL) module to address latent and unobservable biases through a learned mediator dictionary intervention, thereby achieving causal disentanglement of representations for deconfounded reasoning. Importantly, we construct a Demographic-annotated Multimodal Student Personality (DMSP) dataset to support the analysis and discussion of fairness-related factors. Extensive experiments on the benchmark dataset CFI-V2 and our DMSP dataset demonstrate that DCAN consistently improves prediction accuracy, reaching 92.11% and 92.90%, respectively. Meanwhile, the improvements in the fairness metrics of equal opportunity and demographic parity are 6.57% and 7.97% on CFI-V2, and 15.38% and 20.06% on the DMSP dataset. Our code and DMSP dataset are available at https://github.com/Sabrina-han/DCAN
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient InferenceHan Zhao, Min Zhang, Wei Zhao, Pengxiang Ding 等AAAI 2025 · 被引用 125 次
- MDD-5k: A New Diagnostic Conversation Dataset for Mental Disorders Synthesized via Neuro-Symbolic LLM AgentsCongchi Yin, Feng Li, Shu Zhang, Zike Wang 等AAAI 2025 · 被引用 18 次
- Debiased Multimodal Understanding for Human Language SequencesZhi Xu, Dingkang Yang, Mingcheng Li, Yuzheng Wang 等AAAI 2025 · 被引用 16 次
- Fairness without Demographics through Shared Latent Space-Based DebiasingRashidul Islam, Huiyuan Chen, Yiwei CaiAAAI 2024 · 被引用 8 次
相关 Paper
- Towards Deconfounded Image-Text Matching with Causal InferenceWenhui Li, Xinqi Su, Dan Song, Lanjun Wang 等ACM MM 2023 · 被引用 12 次
- Towards Unbiased Visual Emotion Recognition via Causal InterventionYuedong Chen, Xu Yang, Tat-Jen Cham, Jianfei CaiACM MM 2022 · 被引用 27 次
- Towards Ultrasound-based Reliable Disease Diagnosis Using Causal InferenceBolei Chen, Jiaxu Kang, Haonan Yang, Ping Zhong 等AAAI 2026
- Context De-Confounded Emotion RecognitionDingkang Yang, Zhaoyu Chen, Yuzheng Wang, Shunli Wang 等CVPR 2023
- Fair Deepfake Detectors Can GeneralizeHarry Cheng, Ming-Hui Liu, Yangyang Guo, Tianyi Wang 等NeurIPS 2025 · 被引用 11 次
