Dialogue Based Disease Screening Through Domain Customized Reinforcement Learning
Zhuo Liu, Yanxuan Li, Xingzhi Sun, Fei Wang, Gang Hu, Guotong Xie
摘要
In this paper, we study the problem of leveraging dialogue agents learned from reinforcement learning (RL) that can interact with patients for automatic disease screening. This application requires efficient and effective inquiry of appropriate symptoms to make accurate diagnosis recommendations. Existing studies have tried to use RL to perform both symptom inquiry and diagnosis simultaneously, which needs to deal with a large, heterogeneous action space that affects the learning efficiency and effectiveness. To address the challenge, we propose to leverage the models learned from the dialogue data to customize the settings of the reinforcement learning for more efficient action space exploration. In particular, a supervised diagnosis model is built and involved in the definition of state and reward. We also develop the clustering method to form a hierarchy in the action space. These customizations can make the learning task focus on checking the most relevant symptoms, which effectively boost the confidence of diagnosis. Besides, a novel hierarchical reinforcement learning framework with the pretraining strategy is used to reduce the dimension of action space and help the model to converge. For empirical evaluations, we conduct extensive experiments on both synthetic and real-world datasets. The results have demonstrated the superiority of our approach in diagnostic accuracy and interaction efficiency compared with other baseline methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Diaformer: Automatic Diagnosis via Symptoms Sequence GenerationJunying Chen, Dongfang Li, Qingcai Chen, Wenxiu Zhou 等AAAI 2022 · 被引用 35 次
- PatientVLM Meets DocVLM: Pre-Consultation Dialogue Between Vision-Language Models for Efficient DiagnosisK. Lokesh, Abhirama Subramanyam Penamakuri, Uday Agarwal, Apoorva Challa 等AAAI 2026
- Generative Adversarial Regularized Mutual Information Policy Gradient Framework for Automatic DiagnosisYuan Xia, Jingbo Zhou, Zhenhui Shi, Chao Lu 等AAAI 2020 · 被引用 85 次
- Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement LearningDavid Bani-Harouni, Chantal Pellegrini, Ege Özsoy, Nassir Navab 等ICLR 2026 · 被引用 16 次
- Towards Trustworthy Automatic Diagnosis Systems by Emulating Doctors' Reasoning with Deep Reinforcement LearningArsène Fansi Tchango, Rishab Goel, Julien Martel, Zhi Wen 等NeurIPS 2022 · 被引用 14 次
