FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data
Nuredin Ali Abdelkadir, Anjali Ratnam, Zeerak Talat, Stevie Chancellor
摘要
Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, sharing this sensitive data poses privacy risks and limits the growth of benchmark datasets. We comprehensively evaluate whether privacy-preserving ML techniques can enable safer data sharing while preserving performance. Specifically, we apply federated learning (FL) and Differentially Private FL for two widely-studied mental health prediction tasks: depression detection on X (Twitter) and suicide crisis detection on Reddit. We simulate realistic data-sharing scenarios by treating each user as a client in a non-IID setting, evaluating across different client fractions, aggregation strategies, and privacy budgets. While FL achieves comparable performance to centralized training (centralized F1 = 85.63; best FL model F1 = 83.16) on depression identification, we find that Differentially Private FL has a large performance-privacy trade-off (up to F1 = 27.01 drop) even with low levels of noise (epsilon = 50). This is due to the distortion of highly informative yet sparse mental health linguistic markers related to mental health, like health topics and emotion words. This research empirically demonstrates the potential and limitations of current privacy preservation techniques for mental health inference tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett 等ICLR 2021 · 被引用 1,917 次
- Differentially Private Learning Needs Better Features (or Much More Data)Florian Tramèr, Dan BonehICLR 2021 · 被引用 325 次
- Towards Interpretable Mental Health Analysis with Large Language ModelsKailai Yang, Shaoxiong Ji, Tianlin Zhang, Qianqian Xie 等EMNLP 2023 · 被引用 114 次
- A Prioritization Model for Suicidality Risk AssessmentHan-Chin Shing, Philip Resnik, Douglas W. OardACL 2020 · 被引用 41 次
相关 Paper
- ReDepress: A Cognitive Framework for Detecting Depression Relapse from Social MediaAakash Kumar Agarwal, Saprativa Bhattacharjee, Mauli Rastogi, Jemima Jacob 等EMNLP 2025
- Contextual Gaps in Machine Learning for Mental Illness Prediction: The Case of Diagnostic DisclosuresStevie Chancellor, Jessica L. Feuston, Jayhyun ChangCSCW 2023 · 被引用 6 次
- DepressionNet: Learning Multi-modalities with User Post Summarization for Depression Detection on Social MediaHamad Zogan, Imran Razzak, Shoaib Jameel, Guandong XuSIGIR 2021 · 被引用 99 次
- A Framework for Understanding the Relationship between Social Media Discourse and Mental HealthSanjana Mendu, Anna N. Baglione, Sonia Baee, Congyu Wu 等CSCW 2020 · 被引用 21 次
- Learning Language and Multimodal Privacy-Preserving Markers of Mood from Mobile DataPaul Pu Liang, Terrance Liu, Anna Cai, Michal Muszynski 等ACL 2021
