Early Identification of Depression Severity Levels on Reddit Using Ordinal Classification
Usman Naseem, Adam G. Dunn, Jinman Kim, Matloob Khushi
Abstract
User-generated text on social media is a promising avenue for public health surveillance and has been actively explored for its feasibility in the early identification of depression. Existing methods in the identification of depression have shown promising results; however, these methods were all focused on treating the identification as a binary classification problem. To date, there has been little effort towards identifying users’ depression severity level and disregard the inherent ordinal nature across these fine-grain levels. This paper aims to make early identification of depression severity levels on social media data. To accomplish this, we built a new dataset based on the inherent ordinal nature over depression severity levels using clinical depression standards on Reddit posts. The posts were classified into 4 depression severity levels covering the clinical depression standards on social media. Accordingly, we reformulate the early identification of depression as an ordinal classification task over clinical depression standards such as Beck’s Depression Inventory and the Depressive Disorder Annotation scheme to identify depression severity levels. With these, we propose a hierarchical attention method optimized to factor in the increasing depression severity levels through a soft probability distribution. We experimented using two datasets (a public dataset having more than one post from each user and our built dataset with a single user post) using real-world Reddit posts that have been classified according to questionnaires built by clinical experts and demonstrated that our method outperforms state-of-the-art models. Finally, we conclude by analyzing the minimum number of posts required to identify depression severity level followed by a discussion of empirical and practical considerations of our study.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers8
- Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text DataXuhai Xu, Bingsheng Yao, Yuanzhe Dong, Saadia Gabriel et al.UbiComp 2024 · 281 citations
- Cognitive Reframing of Negative Thoughts through Human-Language Model InteractionAshish Sharma, Kevin Rushton, Inna E. Lin, David Wadden et al.ACL 2023 · 32 citations
- Predicting Information Pathways Across Online CommunitiesYiqiao Jin, Yeon-Chang Lee, Kartik Sharma, Meng Ye et al.KDD 2023 · 18 citations
- Semantic Similarity Models for Depression Severity EstimationAnxo Pérez, Neha Warikoo, Kexin Wang, Javier Parapar et al.EMNLP 2023 · 8 citations
- InterMind: Doctor-Patient-Family Interactive Depression Assessment Empowered by Large Language ModelsZhiyuan Zhou, Jilong Liu, Sanwang Wang, Shijie Hao et al.ACM MM 2025 · 2 citations
Related papers
- ReDepress: A Cognitive Framework for Detecting Depression Relapse from Social MediaAakash Kumar Agarwal, Saprativa Bhattacharjee, Mauli Rastogi, Jemima Jacob et al.EMNLP 2025
- Towards Identifying Fine-Grained Depression Symptoms from MemesShweta Yadav, Cornelia Caragea, Chenye Zhao, Naincy Kumari et al.ACL 2023 · 5 citations
- HOPE: Hybrid Optimized Parallel Encoding with Supervised and Unsupervised Semantic Fusion for Depression Symptom DetectionTu-Phuong Mai, Minh-Ha H. Le, Duc-Luong Tran, Phuong-Anh Chu et al.ACL 2026
- Improving the Generalizability of Depression Detection by Leveraging Clinical QuestionnairesThong Nguyen, Andrew Yates, Ayah Zirikly, Bart Desmet et al.ACL 2022 · 66 citations
- DepressionNet: Learning Multi-modalities with User Post Summarization for Depression Detection on Social MediaHamad Zogan, Imran Razzak, Shoaib Jameel, Guandong XuSIGIR 2021 · 99 citations
