Deep Reinforcement Learning for Cost-Effective Medical Diagnosis
Zheng Yu, Yikuan Li, Joseph C. Kim, Kaixuan Huang, Yuan Luo, Mengdi Wang
摘要
Dynamic diagnosis is desirable when medical tests are costly or time-consuming. In this work, we use reinforcement learning (RL) to find a dynamic policy that selects lab test panels sequentially based on previous observations, ensuring accurate testing at a low cost. Clinical diagnostic data are often highly imbalanced; therefore, we aim to maximize the score instead of the error rate. However, optimizing the non-concave score is not a classic RL problem, thus invalidates standard RL methods. To remedy this issue, we develop a reward shaping approach, leveraging properties of the score and duality of policy optimization, to provably find the set of all Pareto-optimal policies for budget-constrained score maximization. To handle the combinatorially complex state space, we propose a Semi-Model-based Deep Diagnosis Policy Optimization (SM-DDPO) framework that is compatible with end-to-end training and online learning. SM-DDPO is tested on diverse clinical tasks: ferritin abnormality detection, sepsis mortality prediction, and acute kidney injury diagnosis. Experiments with real-world data validate that SM-DDPO trains efficiently and identifies all Pareto-front solutions. Across all tasks, SM-DDPO is able to achieve state-of-the-art diagnosis accuracy (in some cases higher than conventional methods) with up to reduction in testing cost. The code is available at [https://github.com/Zheng321/Deep-Reinforcement-Learning-for-Cost-Effective-Medical-Diagnosis].
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement LearningDavid Bani-Harouni, Chantal Pellegrini, Ege Özsoy, Nassir Navab 等ICLR 2026 · 被引用 16 次
- ED-Copilot: Reduce Emergency Department Wait Time with Language Model Diagnostic AssistanceLiwen Sun, Abhineet Agarwal, Aaron Kornblith, Bin Yu 等ICML 2024 · 被引用 7 次
- Limited Query Graph Connectivity TestMingyu Guo, Jialiang Li, Aneta Neumann, Frank Neumann 等AAAI 2024 · 被引用 7 次
- Timely Clinical Diagnosis through Active Test SelectionSilas Ruhrberg Estévez, Nicolás Astorga, Mihaela van der SchaarNeurIPS 2025 · 被引用 4 次
它引用的顶会 Paper2
相关 Paper
- Acquisition Conditioned Oracle for Nongreedy Active Feature AcquisitionMichael Valancius, Max Lennon, Junier OlivaICML 2024 · 被引用 7 次
- Salus: Strategic Diagnostic Testing for Complex Diagnosis via Multi-Agent Reinforcement LearningShuohao Gao, Xuanzhong Chen, Lingxiao Luo, Zilin Ding 等ICML 2026
- Adversarial Cooperative Imitation Learning for Dynamic Treatment Regimes✱Lu Wang, Wenchao Yu, Xiaofeng He, Wei Cheng 等WWW 2020 · 被引用 33 次
- Dialogue Based Disease Screening Through Domain Customized Reinforcement LearningZhuo Liu, Yanxuan Li, Xingzhi Sun, Fei Wang 等KDD 2021 · 被引用 8 次
- QoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO TrainingDavid Dai, Peilin Chen, Chanakya Ekbote, Paul Pu LiangNeurIPS 2025 · 被引用 48 次
