POETREE: Interpretable Policy Learning with Adaptive Decision Trees
Alizée Pace, Alex J. Chan, Mihaela van der Schaar
Abstract
Building models of human decision-making from observed behaviour is critical to better understand, diagnose and support real-world policies such as clinical care. As established policy learning approaches remain focused on imitation performance, they fall short of explaining the demonstrated decision-making process. Policy Extraction through decision Trees (POETREE) is a novel framework for interpretable policy learning, compatible with fully-offline and partially-observable clinical decision environments -- and builds probabilistic tree policies determining physician actions based on patients' observations and medical history. Fully-differentiable tree architectures are grown incrementally during optimization to adapt their complexity to the modelling task, and learn a representation of patient history through recurrence, resulting in decision tree policies that adapt over time with patient information. This policy learning method outperforms the state-of-the-art on real and synthetic medical datasets, both in terms of understanding, quantifying and evaluating observed behaviour as well as in accurately replicating it -- with potential to improve future decision support systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8c7f6eb3-8921-4ce7-b6e4-256bd410afc6Cited by top-tier papers6
- Tree Variational AutoencodersLaura Manduchi, Moritz Vandenhirtz, Alain Ryser, Julia E. VogtNeurIPS 2023 · 17 citations
- Delphic Offline Reinforcement Learning under Nonidentifiable Hidden ConfoundingAlizée Pace, Hugo Yèche, Bernhard Schölkopf, Gunnar Rätsch et al.ICLR 2024 · 9 citations
- Inverse Online Learning: Understanding Non-Stationary and Reactionary PoliciesAlex J. Chan, Alicia Curth, Mihaela van der SchaarICLR 2022 · 8 citations
- Synthetic Model Combination: An Instance-wise Approach to Unsupervised Ensemble LearningAlex J. Chan, Mihaela van der SchaarNeurIPS 2022 · 6 citations
- Contextualized Policy Recovery: Modeling and Interpreting Medical Decisions with Adaptive Imitation LearningJannik Deuschel, Caleb Ellington, Yingtao Luo, Benjamin J. Lengerich et al.ICML 2024 · 5 citations
Builds on7
- Imitation Learning via Off-Policy Distribution MatchingIlya Kostrikov, Ofir Nachum, Jonathan TompsonICLR 2020 · 239 citations
- What Did You Think Would Happen? Explaining Agent Behaviour through Intended OutcomesHerman Yau, Chris Russell, Simon HadfieldNeurIPS 2020 · 44 citations
- Explaining by Imitating: Understanding Decisions by Interpretable Policy LearningAlihan Hüyük, Daniel Jarrett, Cem Tekin, Mihaela van der SchaarICLR 2021 · 22 citations
- Inverse Active Sensing: Modeling and Understanding Timely Decision-MakingDaniel Jarrett, Mihaela van der SchaarICML 2020 · 20 citations
- Scalable Bayesian Inverse Reinforcement LearningAlex James Chan, Mihaela van der SchaarICLR 2021 · 11 citations
Related papers
- Scalable Multi-Action Offline Policy Learning with an m-ary TreeShusei EshimaKDD 2026
- Generative Models for Automatic Medical Decision Rule Extraction from TextYuxin He, Buzhou Tang, Xiaoling WangEMNLP 2024 · 2 citations
- Interpretable Off-Policy Learning via Hyperbox SearchDaniel Tschernutter, Tobias Hatt, Stefan FeuerriegelICML 2022 · 7 citations
- Learning Prescriptive ReLU NetworksWei Sun, Asterios TsiourvasICML 2023 · 3 citations
- SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control TasksYongyan Wen, Siyuan Li, Rongchang Zuo, Lei Yuan et al.AAAI 2025 · 4 citations
