Active Policy Optimization for Individualized Dosing via Gradient Variance Minimization
Yi Wan, Xin Wang, Huanhuan Chen
Abstract
In domains such as healthcare and marketing, learning optimal individualized dosing policies to maximize utility is crucial, yet high experimental costs impose strict budget constraints, necessitating efficient active policy learning. Existing active learning methods in causal inference primarily focus on binary treatments and effect estimation, leaving continuous dosing and policy optimization underexplored. To address this gap, we propose an active learning framework tailored for optimal policy learning. Exploiting the inherent structure of dose-response curves, we theoretically show that the policy optimization regret is bounded by the expected posterior gradient variance at the estimated optimal doses. Motivated by this result, we introduce Gradient Variance Active Learning for Individualized Dosing (GVALID), a batch acquisition strategy that greedily selects samples to minimize target gradient variance for efficient policy learning. Experiments demonstrate that GVALID achieves superior performance under strict budget constraints 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8b36e0a4-3337-4f39-b3ee-2c92ec3adc21Builds on16
- Learning Counterfactual Representations for Estimating Individual Dose-Response CurvesPatrick Schwab, Lorenz Linhardt, Stefan Bauer, Joachim M. Buhmann et al.AAAI 2020 · 159 citations
- Time Series Deconfounder: Estimating Treatment Effects over Time in the Presence of Hidden ConfoundersIoana Bica, Ahmed M. Alaa, Mihaela van der SchaarICML 2020 · 133 citations
- Optimal Best-arm Identification in Linear BanditsYassir Jedra, Alexandre ProutièreNeurIPS 2020 · 99 citations
- Regret Bounds for Batched BanditsHossein Esfandiari, Amin Karbasi, Abbas Mehrabian, Vahab S. MirrokniAAAI 2021 · 74 citations
- Active Bayesian Causal InferenceChristian Toth, Lars Lorch, Christian Knoll, Andreas Krause et al.NeurIPS 2022 · 52 citations
Related papers
- Progressive Generalization Risk Reduction for Data-Efficient Causal Effect EstimationHechuan Wen, Tong Chen, Guanhua Ye, Li Kheng Chai et al.KDD 2025 · 1 citation
- ABC3: Active Bayesian Causal Inference with Cohn Criteria in Randomized ExperimentsTaehun Cha, Donghun LeeAAAI 2025
- Budgeted Active Experimentation for Treatment Effect Estimation from Observational and Randomized DataJiacan Gao, Xinyan Su, Mingyuan Ma, Yiyan HUANG et al.ICML 2026 · 1 citation
- Learning to search efficiently for causally near-optimal treatmentsSamuel Håkansson, Viktor Lindblom, Omer Gottesman, Fredrik D. JohanssonNeurIPS 2020 · 7 citations
- Differentiable Multi-Target Causal Bayesian Experimental DesignPanagiotis Tigas, Yashas Annadani, Desi R. Ivanova, Andrew Jesson et al.ICML 2023 · 15 citations
