EviCare: Enhancing Diagnosis Prediction with Deep Model-Guided Evidence for In-Context Reasoning
Hengyu Zhang, Xuyun Zhang, Pengxiang Zhan, Linhao Luo, Hang Lv, Yanchao Tan, Shirui Pan, Carl Yang
Abstract
Recent advances in large language models (LLMs) have enabled promising progress in diagnosis prediction from electronic health records (EHRs). However, existing LLM-based approaches tend to overfit to historically observed diagnoses, often overlooking novel yet clinically important conditions that are critical for early intervention. To address this, we propose EviCare, an in-context reasoning framework that integrates deep model guidance into LLM-based diagnosis prediction. Rather than prompting LLMs directly with raw EHR inputs, EviCare performs (1) deep model inference for candidate selection, (2) evidential prioritization for set-based EHRs, and (3) relational evidence construction for novel diagnosis prediction. These signals are then composed into an adaptive in-context prompt to guide LLM reasoning in an accurate and interpretable manner. Extensive experiments on two real-world EHR benchmarks (MIMIC-III and MIMIC-IV) demonstrate that EviCare 1 achieves significant performance gains, which consistently outperforms both LLM-only and deep model-only baselines by an average of 20.65% across precision and accuracy metrics. The improvements are particularly notable in challenging novel diagnosis prediction, yielding average improvements of 30.97%.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dffc16fc-8b48-4c83-8aeb-1438b75f5a5cBuilds on11
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le et al.ICLR 2023 · 681 citations
- HiTANet: Hierarchical Time-Aware Attention Networks for Risk Prediction on Electronic Health RecordsJunyu Luo, Muchao Ye, Cao Xiao, Fenglong MaKDD 2020 · 187 citations
- StageNet: Stage-Aware Neural Networks for Health Risk PredictionJunyi Gao, Cao Xiao, Yasha Wang, Wen Tang et al.WWW 2020 · 131 citations
- GraphCare: Enhancing Healthcare Predictions with Personalized Knowledge GraphsPengcheng Jiang, Cao Xiao, Adam Cross, Jimeng SunICLR 2024 · 77 citations
Related papers
- Toward Better EHR Reasoning in LLMs: Reinforcement Learning with Expert Attention GuidanceYue Fang, Yuxin Guo, Jiaran Gao, Hongxin Ding et al.AAAI 2026 · 4 citations
- CARER - ClinicAl Reasoning-Enhanced Representation for Temporal Health Risk PredictionTuan Nguyen, Thanh Trung Huynh, Minh Hieu Phan, Quoc Viet Hung Nguyen et al.EMNLP 2024 · 1 citation
- Reasoning-Enhanced Healthcare Predictions with Knowledge Graph Community RetrievalPengcheng Jiang, Cao Xiao, Minhao Jiang, Parminder Bhatia et al.ICLR 2025
- BoxLM: Unifying Structures and Semantics of Medical Concepts for Diagnosis Prediction in HealthcareYanchao Tan, Hang Lv, Yunfei Zhan, Guofang Ma et al.ICML 2025
- Large Language Models Are Clinical Reasoners: Reasoning-Aware Diagnosis Framework with Prompt-Generated RationalesTaeyoon Kwon, Kai Tzu-iunn Ong, Dongjin Kang, Seungjun Moon et al.AAAI 2024 · 71 citations
