Treatment Policy Learning in Multiobjective Settings with Fully Observed Outcomes
Soorajnath Boominathan, Michael Oberst, Helen Zhou, Sanjat Kanjilal, David A. Sontag
Abstract
In several medical decision-making problems, such as antibiotic prescription, laboratory testing can provide precise indications for how a patient will respond to different treatment options. This enables us to "fully observe" all potential treatment outcomes, but while present in historical data, these results are infeasible to produce in real-time at the point of the initial treatment decision. Moreover, treatment policies in these settings often need to trade off between multiple competing objectives, such as effectiveness of treatment and harmful side effects. We present, compare, and evaluate three approaches for learning individualized treatment policies in this setting: First, we consider two indirect approaches, which use predictive models of treatment response to construct policies optimal for different trade-offs between objectives. Second, we consider a direct approach that constructs such a set of policies without intermediate models of outcomes. Using a medical dataset of Urinary Tract Infection (UTI) patients, we show that all approaches learn policies that achieve strictly better performance on all outcomes than clinicians, while also trading off between different objectives. We demonstrate additional benefits of the direct approach, including flexibly incorporating other goals such as deferral to physicians on simple cases. CCS CONCEPTS • Computing methodologies → Supervised learning; • Applied computing → Health care information systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext efcd8818-8b48-43fb-b420-88569d6e16b1Cited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- Towards Safe Policy Learning under Partial Identifiability: A Causal ApproachShalmali Joshi, Junzhe Zhang, Elias BareinboimAAAI 2024 · 10 citations
- Contextualized Policy Recovery: Modeling and Interpreting Medical Decisions with Adaptive Imitation LearningJannik Deuschel, Caleb Ellington, Yingtao Luo, Benjamin J. Lengerich et al.ICML 2024 · 5 citations
- Inferring Lexicographically-Ordered Rewards from PreferencesAlihan Hüyük, William R. Zame, Mihaela van der SchaarAAAI 2022 · 6 citations
- Clinician-in-the-Loop Decision Making: Reinforcement Learning with Near-Optimal Set-Valued PoliciesShengpu Tang, Aditya Modi, Michael W. Sjoding, Jenna WiensICML 2020 · 35 citations
- Zero-shot causal learningHamed Nilforoshan, Michael Moor, Yusuf H. Roohani, Yining Chen et al.NeurIPS 2023 · 25 citations
