Active Policy Optimization for Individualized Dosing via Gradient Variance Minimization
Yi Wan, Xin Wang, Huanhuan Chen
摘要
In domains such as healthcare and marketing, learning optimal individualized dosing policies to maximize utility is crucial, yet high experimental costs impose strict budget constraints, necessitating efficient active policy learning. Existing active learning methods in causal inference primarily focus on binary treatments and effect estimation, leaving continuous dosing and policy optimization underexplored. To address this gap, we propose an active learning framework tailored for optimal policy learning. Exploiting the inherent structure of dose-response curves, we theoretically show that the policy optimization regret is bounded by the expected posterior gradient variance at the estimated optimal doses. Motivated by this result, we introduce Gradient Variance Active Learning for Individualized Dosing (GVALID), a batch acquisition strategy that greedily selects samples to minimize target gradient variance for efficient policy learning. Experiments demonstrate that GVALID achieves superior performance under strict budget constraints 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Learning Counterfactual Representations for Estimating Individual Dose-Response CurvesPatrick Schwab, Lorenz Linhardt, Stefan Bauer, Joachim M. Buhmann 等AAAI 2020 · 被引用 159 次
- Time Series Deconfounder: Estimating Treatment Effects over Time in the Presence of Hidden ConfoundersIoana Bica, Ahmed M. Alaa, Mihaela van der SchaarICML 2020 · 被引用 133 次
- Optimal Best-arm Identification in Linear BanditsYassir Jedra, Alexandre ProutièreNeurIPS 2020 · 被引用 99 次
- Regret Bounds for Batched BanditsHossein Esfandiari, Amin Karbasi, Abbas Mehrabian, Vahab S. MirrokniAAAI 2021 · 被引用 74 次
- Active Bayesian Causal InferenceChristian Toth, Lars Lorch, Christian Knoll, Andreas Krause 等NeurIPS 2022 · 被引用 52 次
相关 Paper
- Progressive Generalization Risk Reduction for Data-Efficient Causal Effect EstimationHechuan Wen, Tong Chen, Guanhua Ye, Li Kheng Chai 等KDD 2025 · 被引用 1 次
- ABC3: Active Bayesian Causal Inference with Cohn Criteria in Randomized ExperimentsTaehun Cha, Donghun LeeAAAI 2025
- Budgeted Active Experimentation for Treatment Effect Estimation from Observational and Randomized DataJiacan Gao, Xinyan Su, Mingyuan Ma, Yiyan HUANG 等ICML 2026 · 被引用 1 次
- Learning to search efficiently for causally near-optimal treatmentsSamuel Håkansson, Viktor Lindblom, Omer Gottesman, Fredrik D. JohanssonNeurIPS 2020 · 被引用 7 次
- Differentiable Multi-Target Causal Bayesian Experimental DesignPanagiotis Tigas, Yashas Annadani, Desi R. Ivanova, Andrew Jesson 等ICML 2023 · 被引用 15 次
