Bridging Domains in Mental Stress Assessment via Retrieval-Augmented Reasoning
Yi Dai, Yang Ding, Kaisheng Zeng
Abstract
Mental stress assessment is crucial for mental and physical well-being. However, it faces limitations due to domain fragmentation, in which contextual variations in stress triggers and demographics hinder the generalization of assessment models across real-world scenarios. Additionally, mental stress assessment is sensitive and human-centric due to its implications for mental health interventions, emphasizing the need for model transparency and trustworthiness. To address this gap, we propose Retrieval-Augmented Reasoning, a novel framework that bridges domain gaps in mental stress assessment through transparent step-by-step reasoning and dynamic in-context example retrieval. Our framework introduces two key components: (1) a ''detect-then-assess'' reasoning chain decouples stress-relevant facial action units (AUs) from domain-specific noise by first generating textual descriptions as intermediate reasoning step (e.g., ''eyebrow: inner portions raised''). The model then reflects on and learns to refine these descriptions via Direct Preference Optimization (DPO), ensuring faithfulness and helpfulness; (2) a dual-encoder multimodal retriever dynamically selects proper in-context examples from source domain to enhance target-domain assessments, leveraging feedback from the assessment model to optimize retrieval. Experimental results demonstrate that our framework consistently outperforms large multimodal foundation models, stress assessment baselines, and domain generalization methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 8db4f93d-a1e1-4373-82f2-8dd67b9d7901Related papers
- Interpretable Video based Stress Detection with Self-Refine Chain ReasoningYi Dai, Yang Ding, Lei Cao, Kaisheng Zeng et al.ICDE 2025 · 1 citation
- Integrating Content-Semantics-World Knowledge to Detect Stress from VideosYang Ding, Yi Dai, Xin Wang, Ling Feng et al.ACM MM 2024 · 3 citations
- From Subtle Hints to Grand Expressions - Mastering Fine-grained Emotions with Dynamic Multimodal AnalysisQinfu Xu, Liyuan Pan, Shaozu Yuan, Yiwei Wei et al.ACM MM 2025
- Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video ReasoningSongyuan Yang, Weijiang Yu, Ziyu Liu, Guijian Tang et al.CVPR 2026 · 2 citations
- Uncertainty Quantification for Retrieval-Augmented ReasoningHeydar Soudani, Hamed Zamani, Faegheh HasibiSIGIR 2026 · 1 citation
