Delphi: A Neuro-Symbolic Framework for Individualized, Safe and Interpretable Treatment Recommendation
Muchan Tao, Haonan Qin, Yuqi Fang, Caifeng Shan, Tieniu Tan
摘要
Clinical reinforcement learning (RL) holds promise for treatment recommendation. However, its adoption is hindered by black box decision processes, limited safety guarantees, and a lack of individualized treatment. To address these issues, we introduce Delphi, the first trainable neuro symbolic causal RL framework for dynamic treatment planning, designed to answer three core clinical questions: Why for this patient? Why is it safe? Why this action? Specifically, Delphi constructs: 1) causality aware state modeling, using discretized physiological variables and subgroup specific causal graphs; 2) adaptive symbolic rule constraints, combining clinical guidelines and behavior-derived rules into the RL system; and 3) interpretable decision fusion, where actions are selected based on joint neural symbolic Q values and explained via structured LLM-based justifications. We evaluate Delphi on MIMIC-III sepsis cohort with more than 20,000 trajectories, and experiments show that our Delphi achieves leading performance among existing methods. Moreover, Delphi introduces the first blinded physician evaluation of an explainable RL system in healthcare. Results demonstrate that Delphi consistently outperforms historical physicians' treatments in six dimensions, including adoption rate (+5.75%), understandability (+8.9%), safety (+10.4%), satisfaction (+9.35%), trust (+8.78%), effectiveness (+8.87%). These results highlight Delphi's potential as an interpretable, safe, and patientspecific AI assistant for critical care medicine.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
- Medical Dead-ends and Learning to Identify High-Risk States and TreatmentsMehdi Fatemi, Taylor W. Killian, Jayakumar Subramanian, Marzyeh GhassemiNeurIPS 2021 · 被引用 51 次
- Deconfounding Actor-Critic Network with Policy Adaptation for Dynamic Treatment RegimesChangchang Yin, Ruoqi Liu, Jeffrey M. Caterino, Ping ZhangKDD 2022 · 被引用 5 次
相关 Paper
- Benchmarking Reinforcement Learning Algorithms for ICU Ventilator Settings: An Interpretable and Probabilistic Patient Environment for Doctor AgentsYa-Hsi Chang, Po-Chih KuoAAAI 2026
- Ignore, Trust, or Negotiate: Understanding Clinician Acceptance of AI-Based Treatment Recommendations in Health CareVenkatesh Sivaraman, Leigh A. Bukowski, Joel Levin, Jeremy M. Kahn 等CHI 2023 · 被引用 126 次
- Delphic Offline Reinforcement Learning under Nonidentifiable Hidden ConfoundingAlizée Pace, Hugo Yèche, Bernhard Schölkopf, Gunnar Rätsch 等ICLR 2024 · 被引用 9 次
- CausalXRL: Explainable Reinforcement Learning through Causal Graph ReasoningYanming Zhang, Eric Papenhausen, Klaus MuellerICML 2026
- RES-MR: Risk-Aware Reasoning for Explainable and Safe Medication RecommendationCong Wang, Jin Li, Shoujin Wang, Yishuo Li 等SIGIR 2026 · 被引用 1 次
