Deep Contextual Clinical Prediction with Reverse Distillation
Rohan S. Kodialam, Rebecca Boiarsky, Justin Lim, Aditya Sai, Neil Dixit, David A. Sontag
摘要
Healthcare providers are increasingly using machine learning to predict patient outcomes to make meaningful interventions. However, despite innovations in this area, deep learning models often struggle to match performance of shallow linear models in predicting these outcomes, making it difficult to leverage such techniques in practice. In this work, motivated by the task of clinical prediction from insurance claims, we present a new technique called reverse distillation which pretrains deep models by using high-performing linear models for initialization. We make use of the longitudinal structure of insurance claims datasets to develop Self Attention with Reverse Distillation, or SARD, an architecture that utilizes a combination of contextual embedding, temporal embedding and self-attention mechanisms and most critically is trained via reverse distillation. SARD outperforms state-of-the-art methods on multiple clinical prediction outcomes, with ablation studies revealing that reverse distillation is a primary driver of these improvements. Code is available at https://github.com/clinicalml/omop-learn .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- Fast, Accurate, and Simple Models for Tabular Data via Augmented DistillationRasool Fakoor, Jonas Mueller, Nick Erickson, Pratik Chaudhari 等NeurIPS 2020 · 被引用 65 次
- Joint Fine-tuning and Conversion of Pretrained Speech and Language Models towards Linear ComplexityMutian He, Philip N. GarnerICLR 2025
- SeqCare: Sequential Training with External Medical Knowledge Graph for Diagnosis Prediction in Healthcare DataYongxin Xu, Xu Chu, Kai Yang, Zhiyuan Wang 等WWW 2023 · 被引用 41 次
- Distilling Knowledge from Publicly Available Online EMR Data to Emerging Epidemic for PrognosisLiantao Ma, Xinyu Ma, Junyi Gao, Xianfeng Jiao 等WWW 2021 · 被引用 32 次
- Reverse Distillation: Consistently Scaling Protein Language Model RepresentationsDarius Catrina, Christian Bepler, Samuel Sledzieski, Rohit SinghICLR 2026 · 被引用 2 次
