Off-Policy Evaluation with Policy-Dependent Optimization Response
Wenshuo Guo, Michael I. Jordan, Angela Zhou
摘要
The intersection of causal inference and machine learning for decision-making is rapidly expanding, but the default decision criterion remains an average of individual causal outcomes across a population. In practice, various operational restrictions ensure that a decision-maker's utility is not realized as an average but rather as an output of a downstream decision-making problem (such as matching, assignment, network flow, minimizing predictive risk). In this work, we develop a new framework for off-policy evaluation with policy-dependent linear optimization responses: causal outcomes introduce stochasticity in objective function coefficients. Under this framework, a decision-maker's utility depends on the policy-dependent optimization, which introduces a fundamental challenge of optimization bias even for the case of policy evaluation. We construct unbiased estimators for the policy-dependent estimand by a perturbation method, and discuss asymptotic variance properties for a set of adjusted plug-in estimators. Lastly, attaining unbiased policy evaluation allows for policy optimization: we provide a general algorithm for optimizing causal interventions. We corroborate our theoretical results with numerical simulations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Decision-Focused Learning with Directional GradientsMichael Huang, Vishal GuptaNeurIPS 2024 · 被引用 25 次
- Empirical Gateaux Derivatives for Causal InferenceMichael I. Jordan, Yixin Wang, Angela ZhouNeurIPS 2022 · 被引用 13 次
它引用的顶会 Paper6
- Performative PredictionJuan C. Perdomo, Tijana Zrnic, Celestine Mendler-Dünner, Moritz HardtICML 2020 · 被引用 422 次
- Universal Off-Policy EvaluationYash Chandak, Scott Niekum, Bruno C. da Silva, Erik G. Learned-Miller 等NeurIPS 2021 · 被引用 64 次
- RieszNet and ForestRiesz: Automatic Debiased Machine Learning with Neural Nets and Random ForestsVictor Chernozhukov, Whitney Newey, Victor Quintas-Martinez, Vasilis SyrgkanisICML 2022 · 被引用 61 次
- Interference, Bias, and Variance in Two-Sided Marketplace Experimentation: Guidance for PlatformsHannah Li, Geng Zhao, Ramesh Johari, Gabriel Y. WeintraubWWW 2022 · 被引用 46 次
- Cost-Effective Incentive Allocation via Structured Counterfactual InferenceRomain Lopez, Chenchen Li, Xiang Yan, Junwu Xiong 等AAAI 2020 · 被引用 21 次
相关 Paper
- Causal Strategic Linear RegressionYonadav Shavit, Benjamin L. Edelman, Brian AxelrodICML 2020 · 被引用 91 次
- Efficient and Sharp Off-Policy Learning under Unobserved ConfoundingKonstantin Hess, Dennis Frauen, Valentyn Melnychuk, Stefan FeuerriegelICLR 2026 · 被引用 5 次
- Strategic Instrumental Variable Regression: Recovering Causal Relationships From Strategic ResponsesKeegan Harris, Dung Daniel T. Ngo, Logan Stapleton, Hoda Heidari 等ICML 2022 · 被引用 37 次
- Causal Modeling for Fairness In Dynamical SystemsElliot Creager, David Madras, Toniann Pitassi, Richard S. ZemelICML 2020 · 被引用 72 次
- Estimation of Bounds on Potential Outcomes For Decision MakingMaggie Makar, Fredrik D. Johansson, John V. Guttag, David A. SontagICML 2020 · 被引用 11 次
