Contextualized Policy Recovery: Modeling and Interpreting Medical Decisions with Adaptive Imitation Learning
Jannik Deuschel, Caleb Ellington, Yingtao Luo, Benjamin J. Lengerich, Pascal Friederich, Eric P. Xing
Abstract
Interpretable policy learning seeks to estimate intelligible decision policies from observed actions; however, existing models force a tradeoff between accuracy and interpretability, limiting data-driven interpretations of human decision-making processes. Fundamentally, existing approaches are burdened by this tradeoff because they represent the underlying decision process as a universal policy, when in fact human decisions are dynamic and can change drastically under different contexts. Thus, we develop Contextualized Policy Recovery (CPR), which re-frames the problem of modeling complex decision processes as a multi-task learning problem, where each context poses a unique task and complex decision policies can be constructed piece-wise from many simple context-specific policies. CPR models each context-specific policy as a linear map, and generates new policy models as contexts are updated with new observations. We provide two flavors of the CPR framework: one focusing on exact local interpretability, and one retaining full global interpretability. We assess CPR through studies on simulated and real data, achieving state-of-the-art performance on predicting antibiotic prescription in intensive care units ( AUROC vs. previous SOTA) and predicting MRI prescription for Alzheimer's patients ( AUROC vs. previous SOTA). With this improvement, CPR closes the accuracy gap between interpretable and black-box methods, allowing high-resolution exploration and analysis of context-specific decision models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4d4b0b21-ebaf-4ffa-8e85-a10a787c2bf6Builds on5
- Imitation Learning via Off-Policy Distribution MatchingIlya Kostrikov, Ofir Nachum, Jonathan TompsonICLR 2020 · 239 citations
- What Did You Think Would Happen? Explaining Agent Behaviour through Intended OutcomesHerman Yau, Chris Russell, Simon HadfieldNeurIPS 2020 · 44 citations
- Clairvoyance: A Pipeline Toolkit for Medical Time SeriesDaniel Jarrett, Jinsung Yoon, Ioana Bica, Zhaozhi Qian et al.ICLR 2021 · 43 citations
- Explaining by Imitating: Understanding Decisions by Interpretable Policy LearningAlihan Hüyük, Daniel Jarrett, Cem Tekin, Mihaela van der SchaarICLR 2021 · 22 citations
- POETREE: Interpretable Policy Learning with Adaptive Decision TreesAlizée Pace, Alex J. Chan, Mihaela van der SchaarICLR 2022 · 18 citations
Related papers
- Interpretable Off-Policy Learning via Hyperbox SearchDaniel Tschernutter, Tobias Hatt, Stefan FeuerriegelICML 2022 · 7 citations
- Learning Prescriptive ReLU NetworksWei Sun, Asterios TsiourvasICML 2023 · 3 citations
- SMoSE: Sparse Mixture of Shallow Experts for Interpretable Reinforcement Learning in Continuous Control TasksMátyás Vincze, Laura Ferrarotti, Leonardo Lucio Custode, Bruno Lepri et al.AAAI 2025 · 3 citations
- Treatment Policy Learning in Multiobjective Settings with Fully Observed OutcomesSoorajnath Boominathan, Michael Oberst, Helen Zhou, Sanjat Kanjilal et al.KDD 2020 · 1 citation
- Inverse Contextual Bandits: Learning How Behavior Evolves over TimeAlihan Hüyük, Daniel Jarrett, Mihaela van der SchaarICML 2022 · 14 citations
