Reliable Off-Policy Learning for Dosage Combinations
Jonas Schweisthal, Dennis Frauen, Valentyn Melnychuk, Stefan Feuerriegel
摘要
Decision-making in personalized medicine such as cancer therapy or critical care must often make choices for dosage combinations, i.e., multiple continuous treatments. Existing work for this task has modeled the effect of multiple treatments independently, while estimating the joint effect has received little attention but comes with non-trivial challenges. In this paper, we propose a novel method for reliable off-policy learning for dosage combinations. Our method proceeds along three steps: (1) We develop a tailored neural network that estimates the individualized dose-response function while accounting for the joint effect of multiple dependent dosages. (2) We estimate the generalized propensity score using conditional normalizing flows in order to detect regions with limited overlap in the shared covariate-treatment space. (3) We present a gradient-based learning algorithm to find the optimal, individualized dosage combinations. Here, we ensure reliable estimation of the policy value by avoiding regions with limited overlap. We finally perform an extensive evaluation of our method to show its effectiveness. To the best of our knowledge, ours is the first work to provide a method for reliable off-policy learning for optimal dosage combinations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Foundation Models for Causal Inference via Prior-Data Fitted NetworksYuchen Ma, Dennis Frauen, Emil Javurek, Stefan FeuerriegelICLR 2026 · 被引用 37 次
- Bayesian Neural Controlled Differential Equations for Treatment Effect EstimationKonstantin Hess, Valentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICLR 2024 · 被引用 27 次
- Normalizing Flows for Interventional Density EstimationValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICML 2023 · 被引用 25 次
- Conformal Prediction for Causal Effects of Continuous TreatmentsMaresa Schröder, Dennis Frauen, Jonas Schweisthal, Konstantin Hess 等NeurIPS 2025 · 被引用 21 次
- A Neural Framework for Generalized Causal Sensitivity AnalysisDennis Frauen, Fergus Imrie, Alicia Curth, Valentyn Melnychuk 等ICLR 2024 · 被引用 15 次
它引用的顶会 Paper21
- Conservative Q-Learning for Offline Reinforcement LearningAviral Kumar, Aurick Zhou, George Tucker, Sergey LevineNeurIPS 2020 · 被引用 2,881 次
- On Gradient Descent Ascent for Nonconvex-Concave Minimax ProblemsTianyi Lin, Chi Jin, Michael I. JordanICML 2020 · 被引用 587 次
- Learning Counterfactual Representations for Estimating Individual Dose-Response CurvesPatrick Schwab, Lorenz Linhardt, Stefan Bauer, Joachim M. Buhmann 等AAAI 2020 · 被引用 159 次
- Causal Transformer for Estimating Counterfactual OutcomesValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICML 2022 · 被引用 146 次
- Estimating the Effects of Continuous-valued Interventions using Generative Adversarial NetworksIoana Bica, James Jordon, Mihaela van der SchaarNeurIPS 2020 · 被引用 137 次
相关 Paper
- Deep Jump Learning for Off-Policy Evaluation in Continuous Treatment SettingsHengrui Cai, Chengchun Shi, Rui Song, Wenbin LuNeurIPS 2021 · 被引用 18 次
- IGC-Net for conditional average potential outcome estimation over timeKonstantin Hess, Dennis Frauen, Valentyn Melnychuk, Stefan FeuerriegelICLR 2026 · 被引用 8 次
- An Orthogonal Learner for Individualized Outcomes in Markov Decision ProcessesEmil Javurek, Valentyn Melnychuk, Jonas Schweisthal, Konstantin Hess 等ICLR 2026 · 被引用 2 次
- DoseSurv: Predicting Personalized Survival Outcomes under Continuous-Valued TreatmentsMoritz Gögl, Yu Liu, Christopher Yau, Peter J. Watkinson 等NeurIPS 2025 · 被引用 1 次
- VCNet and Functional Targeted Regularization For Learning Causal Effects of Continuous TreatmentsLizhen Nie, Mao Ye, Qiang Liu, Dan NicolaeICLR 2021 · 被引用 81 次
