Bridging Domains in Mental Stress Assessment via Retrieval-Augmented Reasoning
Yi Dai, Yang Ding, Kaisheng Zeng
摘要
Mental stress assessment is crucial for mental and physical well-being. However, it faces limitations due to domain fragmentation, in which contextual variations in stress triggers and demographics hinder the generalization of assessment models across real-world scenarios. Additionally, mental stress assessment is sensitive and human-centric due to its implications for mental health interventions, emphasizing the need for model transparency and trustworthiness. To address this gap, we propose Retrieval-Augmented Reasoning, a novel framework that bridges domain gaps in mental stress assessment through transparent step-by-step reasoning and dynamic in-context example retrieval. Our framework introduces two key components: (1) a ''detect-then-assess'' reasoning chain decouples stress-relevant facial action units (AUs) from domain-specific noise by first generating textual descriptions as intermediate reasoning step (e.g., ''eyebrow: inner portions raised''). The model then reflects on and learns to refine these descriptions via Direct Preference Optimization (DPO), ensuring faithfulness and helpfulness; (2) a dual-encoder multimodal retriever dynamically selects proper in-context examples from source domain to enhance target-domain assessments, leveraging feedback from the assessment model to optimize retrieval. Experimental results demonstrate that our framework consistently outperforms large multimodal foundation models, stress assessment baselines, and domain generalization methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Interpretable Video based Stress Detection with Self-Refine Chain ReasoningYi Dai, Yang Ding, Lei Cao, Kaisheng Zeng 等ICDE 2025 · 被引用 1 次
- Integrating Content-Semantics-World Knowledge to Detect Stress from VideosYang Ding, Yi Dai, Xin Wang, Ling Feng 等ACM MM 2024 · 被引用 3 次
- From Subtle Hints to Grand Expressions - Mastering Fine-grained Emotions with Dynamic Multimodal AnalysisQinfu Xu, Liyuan Pan, Shaozu Yuan, Yiwei Wei 等ACM MM 2025
- Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video ReasoningSongyuan Yang, Weijiang Yu, Ziyu Liu, Guijian Tang 等CVPR 2026 · 被引用 2 次
- Uncertainty Quantification for Retrieval-Augmented ReasoningHeydar Soudani, Hamed Zamani, Faegheh HasibiSIGIR 2026 · 被引用 1 次
