Deep Proxy Causal Learning and its Application to Confounded Bandit Policy Evaluation
Liyuan Xu, Heishiro Kanagawa, Arthur Gretton
摘要
Proxy causal learning (PCL) is a method for estimating the causal effect of treatments on outcomes in the presence of unobserved confounding, using proxies (structured side information) for the confounder. This is achieved via two-stage regression: in the first stage, we model relations among the treatment and proxies; in the second stage, we use this model to learn the effect of treatment on the outcome, given the context provided by the proxies. PCL guarantees recovery of the true causal effect, subject to identifiability conditions. We propose a novel method for PCL, the deep feature proxy variable method (DFPV), to address the case where the proxies, treatments, and outcomes are high-dimensional and have nonlinear complex relationships, as represented by deep neural network features. We show that DFPV outperforms recent state-of-the-art PCL methods on challenging synthetic benchmarks, including settings involving high dimensional image data. Furthermore, we show that PCL can be applied to off-policy evaluation for the confounded bandit problem, in which DFPV also exhibits competitive performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Provably Efficient Reinforcement Learning in Partially Observable Dynamical SystemsMasatoshi Uehara, Ayush Sekhari, Jason D. Lee, Nathan Kallus 等NeurIPS 2022 · 被引用 48 次
- A Minimax Learning Approach to Off-Policy Evaluation in Confounded Partially Observable Markov Decision ProcessesChengchun Shi, Masatoshi Uehara, Jiawei Huang, Nan JiangICML 2022 · 被引用 31 次
- Functional Bilevel Optimization for Machine LearningIeva Petrulionyte, Julien Mairal, Michael ArbelNeurIPS 2024 · 被引用 27 次
- An Instrumental Variable Approach to Confounded Off-Policy EvaluationYang Xu, Jin Zhu, Chengchun Shi, Shikai Luo 等ICML 2023 · 被引用 24 次
- Deep Learning Methods for Proximal Inference via Maximum Moment RestrictionBenjamin Kompa, David R. Bellamy, Thomas Kolokotrones, James M. Robins 等NeurIPS 2022 · 被引用 22 次
它引用的顶会 Paper5
- Off-Policy Evaluation in Partially Observable EnvironmentsGuy Tennenholtz, Uri Shalit, Shie MannorAAAI 2020 · 被引用 91 次
- Learning Deep Features in Instrumental Variable RegressionLiyuan Xu, Yutian Chen, Siddarth Srinivasan, Nando de Freitas 等ICLR 2021 · 被引用 85 次
- Proximal Causal Learning with Kernels: Two-Stage Estimation and Moment RestrictionAfsaneh Mastouri, Yuchen Zhu, Limor Gultchin, Anna Korba 等ICML 2021 · 被引用 78 次
- Provably Efficient Neural Estimation of Structural Equation Models: An Adversarial ApproachLuofeng Liao, You-Lin Chen, Zhuoran Yang, Bo Dai 等NeurIPS 2020 · 被引用 40 次
- A Proxy Variable View of Shared ConfoundingYixin Wang, David M. BleiICML 2021 · 被引用 14 次
相关 Paper
- Automating the Selection of Proxy Variables of Unmeasured ConfoundersFeng Xie, Zhengming Chen, Shanshan Luo, Wang Miao 等ICML 2024 · 被引用 5 次
- Density Ratio-Free Doubly Robust Proxy Causal LearningBariscan Bozkurt, Houssam Zenati, Dimitri Meunier, Liyuan Xu 等NeurIPS 2025 · 被引用 2 次
- Causal Inference with Conditional Front-Door Adjustment and Identifiable Variational AutoencoderZiqi Xu, Debo Cheng, Jiuyong Li, Jixue Liu 等ICLR 2024 · 被引用 26 次
- Deep Multi-Modal Structural Equations For Causal Effect Estimation With Unstructured ProxiesShachi Deshpande, Kaiwen Wang, Dhruv Sreenivas, Zheng Li 等NeurIPS 2022 · 被引用 15 次
- Learning Decision Policies with Instrumental Variables through Double Machine LearningDaqian Shao, Ashkan Soleymani, Francesco Quinzan, Marta KwiatkowskaICML 2024 · 被引用 4 次
