FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data
Nuredin Ali Abdelkadir, Anjali Ratnam, Zeerak Talat, Stevie Chancellor
Abstract
Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, sharing this sensitive data poses privacy risks and limits the growth of benchmark datasets. We comprehensively evaluate whether privacy-preserving ML techniques can enable safer data sharing while preserving performance. Specifically, we apply federated learning (FL) and Differentially Private FL for two widely-studied mental health prediction tasks: depression detection on X (Twitter) and suicide crisis detection on Reddit. We simulate realistic data-sharing scenarios by treating each user as a client in a non-IID setting, evaluating across different client fractions, aggregation strategies, and privacy budgets. While FL achieves comparable performance to centralized training (centralized F1 = 85.63; best FL model F1 = 83.16) on depression identification, we find that Differentially Private FL has a large performance-privacy trade-off (up to F1 = 27.01 drop) even with low levels of noise (epsilon = 50). This is due to the distortion of highly informative yet sparse mental health linguistic markers related to mental health, like health topics and emotion words. This research empirically demonstrates the potential and limitations of current privacy preservation techniques for mental health inference tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 09d738e2-9b72-4ddd-913a-fdff9c942021Builds on6
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- Differentially Private Learning Needs Better Features (or Much More Data)Florian Tramèr, Dan BonehICLR 2021 · 325 citations
- Towards Interpretable Mental Health Analysis with Large Language ModelsKailai Yang, Shaoxiong Ji, Tianlin Zhang, Qianqian Xie et al.EMNLP 2023 · 114 citations
- A Prioritization Model for Suicidality Risk AssessmentHan-Chin Shing, Philip Resnik, Douglas W. OardACL 2020 · 41 citations
Related papers
- ReDepress: A Cognitive Framework for Detecting Depression Relapse from Social MediaAakash Kumar Agarwal, Saprativa Bhattacharjee, Mauli Rastogi, Jemima Jacob et al.EMNLP 2025
- Contextual Gaps in Machine Learning for Mental Illness Prediction: The Case of Diagnostic DisclosuresStevie Chancellor, Jessica L. Feuston, Jayhyun ChangCSCW 2023 · 6 citations
- DepressionNet: Learning Multi-modalities with User Post Summarization for Depression Detection on Social MediaHamad Zogan, Imran Razzak, Shoaib Jameel, Guandong XuSIGIR 2021 · 99 citations
- A Framework for Understanding the Relationship between Social Media Discourse and Mental HealthSanjana Mendu, Anna N. Baglione, Sonia Baee, Congyu Wu et al.CSCW 2020 · 21 citations
- Learning Language and Multimodal Privacy-Preserving Markers of Mood from Mobile DataPaul Pu Liang, Terrance Liu, Anna Cai, Michal Muszynski et al.ACL 2021
