Figurative-cum-Commonsense Knowledge Infusion for Multimodal Mental Health Meme Classification
Abdullah Mazhar, Zuhair Hasan Shaik, Aseem Srivastava, Polly Ruhnke, Lavanya Vaddavalli, Sri Keshav Katragadda, Shweta Yadav, Md. Shad Akhtar
Abstract
The expression of mental health symptoms through non-traditional means, such as memes, has gained remarkable attention over the past few years, with users often highlighting their mental health struggles through figurative intricacies within memes. While humans rely on commonsense knowledge to interpret these complex expressions, current Multimodal Language Models (MLMs) struggle to capture these figurative aspects inherent in memes. To address this gap, we introduce a novel dataset, AxiOM, derived from the GAD anxiety questionnaire, which categorizes memes into six fine-grained anxiety symptoms. Next, we propose a commonsense and domainenriched framework, M3H, to enhance MLMs' ability to interpret figurative language and commonsense knowledge. The overarching goal remains to first understand and then classify the mental health symptoms expressed in memes. We benchmark M3H against 6 competitive baselines (with 20 variations), demonstrating improvements in both quantitative and qualitative metrics, including a detailed human evaluation. We observe a clear improvement of 4.20% and 4.66% on weighted-F1 metric. To assess the generalizability, we perform extensive experiments on a public dataset, RESTORE, for depressive symptom identification, presenting an extensive ablation study that highlights the contribution of each module in both datasets. Our findings reveal limitations in existing models and the advantage of employing commonsense to enhance figurative understanding. CCS Concepts • Computing methodologies → Discourse, dialogue and pragmatics; Natural language generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55d5fe37-55cc-4851-a285-8bfe9f1e3064Cited by top-tier papers3
- All Changes May Have Invariant Principles: Improving Ever-Shifting Harmful Meme Detection via Design Concept ReproductionZiyou Jiang, Mingyang Li, Junjie Wang, Yuekai Huang et al.ACL 2026
- MAMA-Memeia! Multi-Aspect Multi-Agent Collaboration for Depressive Symptoms Identification in MemesSiddhant Agarwal, Adya Dhuler, Polly Ruhnke, Melvin Speisman et al.AAAI 2026
- Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health ConversationAbdullah Mazhar, Het Riteshkumar Shah, Aseem Srivastava, Smriti Joshi et al.ACL 2026
Builds on5
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 3,729 citations
- Towards Identifying Fine-Grained Depression Symptoms from MemesShweta Yadav, Cornelia Caragea, Chenye Zhao, Naincy Kumari et al.ACL 2023 · 5 citations
- Knowledge Planning in Large Language Models for Domain-Aligned Counseling SummarizationAseem Srivastava, Smriti Joshi, Tanmoy Chakraborty, Md. Shad AkhtarEMNLP 2024 · 3 citations
- WorryWords: Norms of Anxiety Association for over 44k English WordsSaif MohammadEMNLP 2024 · 2 citations
Related papers
- Still Not Quite There! Evaluating Large Language Models for Comorbid Mental Health DiagnosisAmey Hengle, Atharva Kulkarni, Shantanu Patankar, Madhumitha Chandrasekaran et al.EMNLP 2024 · 1 citation
- MetaGPT: A Large Vision-Language Model for Meme Metaphor UnderstandingBo Xu, Chenyuan Wang, Xinyu Chen, Hongfei Lin et al.AAAI 2026
- MemeCap: A Dataset for Captioning and Interpreting MemesEunjeong Hwang, Vered ShwartzEMNLP 2023 · 14 citations
- DRMD: Explainable Depression Detection Based on Metaphorical Conceptual MappingDongyu Zhang, Wanqiu Liao, Weichen Hu, Hongfei LinWWW 2026
- MemeQA: Holistic Evaluation for Meme UnderstandingKhoi P. N. Nguyen, Terrence Li, Derek Lou Zhou, Gabriel Xiong et al.ACL 2025 · 3 citations
