Personality-guided Public-Private Domain Disentangled Hypergraph-Former Network for Multimodal Depression Detection
Changzeng Fu, Shiwen Zhao, Yunze Zhang, Zhongquan Jian, Shiqi Zhao, Chaoran Liu
摘要
Depression represents a global mental health challenge requiring efficient and reliable automated detection methods. Current Transformer-or Graph Neural Networks (GNNs)based multimodal depression detection methods face significant challenges in modeling individual differences and cross-modal temporal dependencies across diverse behavioral contexts. Therefore, we propose P 3 HF (Personalityguided Public-Private Domain Disentangled Hypergraph-Former Network) with three key innovations: (1) personalityguided representation learning using LLMs to transform discrete individual features into contextual descriptions for personalized encoding; (2) Hypergraph-Former architecture modeling high-order cross-modal temporal relationships; (3) event-level domain disentanglement with contrastive learning for improved generalization across behavioral contexts. Experiments on MPDD-Young dataset show P 3 HF achieves around 10% improvement on accuracy and weighted F1 for binary and ternary depression classification task over existing methods. Extensive ablation studies validate the independent contribution of each architectural component, confirming that personality-guided representation learning and highorder hypergraph reasoning are both essential for generating robust, individual-aware depression-related representations. The code is released at https://github.com/hacilab/P3HF .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- MISA: Modality-Invariant and -Specific Representations for Multimodal Sentiment AnalysisDevamanyu Hazarika, Roger Zimmermann, Soujanya PoriaACM MM 2020 · 被引用 1,037 次
- AM-GCN: Adaptive Multi-channel Graph Convolutional NetworksXiao Wang, Meiqi Zhu, Deyu Bo, Peng Cui 等KDD 2020 · 被引用 464 次
- EPIC-Fusion: Audio-Visual Temporal Binding for Egocentric Action RecognitionEvangelos Kazakos, Arsha Nagrani, Andrew Zisserman, Dima DamenICCV 2019 · 被引用 395 次
- Multimodal Fusion via Hypergraph Autoencoder and Contrastive Learning for Emotion Recognition in ConversationZijian Yi, Ziming Zhao, Zhishu Shen, Tiehua ZhangACM MM 2024 · 被引用 32 次
- Towards Emotion-enriched Text-to-Motion Generation via LLM-guided Limb-level Emotion ManipulatingTan Yu, Jingjing Wang, Jiawen Wang, Jiamin Luo 等ACM MM 2024 · 被引用 9 次
相关 Paper
- Targeting Borderline Fraudsters: Multi-View Hypergraph Fraud Detection with LLM-Guided Contrastive LearningRui Ou, Kun Zhu, Nana Zhang, Jiangtong Li 等AAAI 2026
- Disentangled-Multimodal Privileged Knowledge Distillation for Depression Recognition with Incomplete Multimodal DataYuchen Pan, Junjun Jiang, Kui Jiang, Xianming LiuACM MM 2024 · 被引用 18 次
- Depression Detection from Social Media: A Mutual Guidance Multi-modal Network with Complementary Graph LearningGuocheng Hu, Chaoqun Zheng, Ruifan Zuo, Fengling Li 等SIGIR 2026
- Disentangled Contrastive Hypergraph Learning for Next POI RecommendationYantong Lai, Yijun Su, Lingwei Wei, Tianqi He 等SIGIR 2024 · 被引用 56 次
- Using Graph Representation Learning with Schema Encoders to Measure the Severity of Depressive SymptomsSimin Hong, Anthony G. Cohn, David Crossland HoggICLR 2022 · 被引用 16 次
