Deep Proxy Causal Learning and its Application to Confounded Bandit Policy Evaluation
Liyuan Xu, Heishiro Kanagawa, Arthur Gretton
Abstract
Proxy causal learning (PCL) is a method for estimating the causal effect of treatments on outcomes in the presence of unobserved confounding, using proxies (structured side information) for the confounder. This is achieved via two-stage regression: in the first stage, we model relations among the treatment and proxies; in the second stage, we use this model to learn the effect of treatment on the outcome, given the context provided by the proxies. PCL guarantees recovery of the true causal effect, subject to identifiability conditions. We propose a novel method for PCL, the deep feature proxy variable method (DFPV), to address the case where the proxies, treatments, and outcomes are high-dimensional and have nonlinear complex relationships, as represented by deep neural network features. We show that DFPV outperforms recent state-of-the-art PCL methods on challenging synthetic benchmarks, including settings involving high dimensional image data. Furthermore, we show that PCL can be applied to off-policy evaluation for the confounded bandit problem, in which DFPV also exhibits competitive performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3607b745-6e27-4712-bdf6-e87eabf205d5Cited by top-tier papers20
- Provably Efficient Reinforcement Learning in Partially Observable Dynamical SystemsMasatoshi Uehara, Ayush Sekhari, Jason D. Lee, Nathan Kallus et al.NeurIPS 2022 · 48 citations
- A Minimax Learning Approach to Off-Policy Evaluation in Confounded Partially Observable Markov Decision ProcessesChengchun Shi, Masatoshi Uehara, Jiawei Huang, Nan JiangICML 2022 · 31 citations
- Functional Bilevel Optimization for Machine LearningIeva Petrulionyte, Julien Mairal, Michael ArbelNeurIPS 2024 · 27 citations
- An Instrumental Variable Approach to Confounded Off-Policy EvaluationYang Xu, Jin Zhu, Chengchun Shi, Shikai Luo et al.ICML 2023 · 24 citations
- Deep Learning Methods for Proximal Inference via Maximum Moment RestrictionBenjamin Kompa, David R. Bellamy, Thomas Kolokotrones, James M. Robins et al.NeurIPS 2022 · 22 citations
Builds on5
- Off-Policy Evaluation in Partially Observable EnvironmentsGuy Tennenholtz, Uri Shalit, Shie MannorAAAI 2020 · 91 citations
- Learning Deep Features in Instrumental Variable RegressionLiyuan Xu, Yutian Chen, Siddarth Srinivasan, Nando de Freitas et al.ICLR 2021 · 85 citations
- Proximal Causal Learning with Kernels: Two-Stage Estimation and Moment RestrictionAfsaneh Mastouri, Yuchen Zhu, Limor Gultchin, Anna Korba et al.ICML 2021 · 78 citations
- Provably Efficient Neural Estimation of Structural Equation Models: An Adversarial ApproachLuofeng Liao, You-Lin Chen, Zhuoran Yang, Bo Dai et al.NeurIPS 2020 · 40 citations
- A Proxy Variable View of Shared ConfoundingYixin Wang, David M. BleiICML 2021 · 14 citations
Related papers
- Automating the Selection of Proxy Variables of Unmeasured ConfoundersFeng Xie, Zhengming Chen, Shanshan Luo, Wang Miao et al.ICML 2024 · 5 citations
- Density Ratio-Free Doubly Robust Proxy Causal LearningBariscan Bozkurt, Houssam Zenati, Dimitri Meunier, Liyuan Xu et al.NeurIPS 2025 · 2 citations
- Causal Inference with Conditional Front-Door Adjustment and Identifiable Variational AutoencoderZiqi Xu, Debo Cheng, Jiuyong Li, Jixue Liu et al.ICLR 2024 · 26 citations
- Deep Multi-Modal Structural Equations For Causal Effect Estimation With Unstructured ProxiesShachi Deshpande, Kaiwen Wang, Dhruv Sreenivas, Zheng Li et al.NeurIPS 2022 · 15 citations
- Learning Decision Policies with Instrumental Variables through Double Machine LearningDaqian Shao, Ashkan Soleymani, Francesco Quinzan, Marta KwiatkowskaICML 2024 · 4 citations
