GDR-learners: Orthogonal Learning of Generative Models for Potential Outcomes
Valentyn Melnychuk, Stefan Feuerriegel
Abstract
Various deep generative models have been proposed to estimate potential outcomes distributions from observational data. However, none of them have the favorable theoretical property of general Neyman-orthogonality and, associated with it, quasioracle efficiency and double robustness. In this paper, we introduce a general suite of generative Neyman-orthogonal (doubly-robust) learners that estimate the conditional distributions of potential outcomes. Our proposed generative doubly-robust learners (GDR-learners) are flexible and can be instantiated with many state-ofthe-art deep generative models. In particular, we develop GDR-learners based on (a) conditional normalizing flows (which we call GDR-CNFs), (b) conditional generative adversarial networks (GDR-CGANs), (c) conditional variational autoencoders (GDR-CVAEs), and (d) conditional diffusion models (GDR-CDMs). Unlike the existing methods, our GDR-learners possess the properties of quasi-oracle efficiency and rate double robustness, and are thus asymptotically optimal. In a series of (semi-)synthetic experiments, we demonstrate that our GDR-learners are very effective and outperform the existing methods in estimating the conditional distributions of potential outcomes. INTRODUCTION Causal machine learning (ML) is widely used to predict potential outcomes (POs), namely, the outcome after an intervention. In medicine, for example, an accurate prediction of POs can guide the choice of the optimal treatment from several available treatment options (Feuerriegel et al., 2024) ). The POs have a central role in the causal ML as they define various causal quantities, such as treatment effects (Curth & van der Schaar, 2021) or a policy value (Qian & Murphy, 2011; Frauen et al., 2025b). Recently, many works have suggested departing from estimating simple conditional averages of the POs (CAPOs) and rather aim at the whole conditional distributions (
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7eba5b79-d94f-42c1-b116-ec915b2859ffBuilds on20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Deep Structural Causal Models for Tractable Counterfactual InferenceNick Pawlowski, Daniel Coelho de Castro, Ben GlockerNeurIPS 2020 · 353 citations
- Counterfactual Generative NetworksAxel Sauer, Andreas GeigerICLR 2021 · 145 citations
- Treatment Effect Estimation with Disentangled Latent FactorsWeijia Zhang, Lin Liu, Jiuyong LiAAAI 2021 · 115 citations
- High Fidelity Image Counterfactuals with Probabilistic Causal ModelsFabio De Sousa Ribeiro, Tian Xia, Miguel Monteiro, Nick Pawlowski et al.ICML 2023 · 68 citations
Related papers
- An Orthogonal Learner for Individualized Outcomes in Markov Decision ProcessesEmil Javurek, Valentyn Melnychuk, Jonas Schweisthal, Konstantin Hess et al.ICLR 2026 · 2 citations
- Normalizing Flows for Interventional Density EstimationValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICML 2023 · 25 citations
- Quantifying Aleatoric Uncertainty of the Treatment Effect: A Novel Orthogonal LearnerValentyn Melnychuk, Stefan Feuerriegel, Mihaela van der SchaarNeurIPS 2024 · 13 citations
- DiffPO: A causal diffusion model for learning distributions of potential outcomesYuchen Ma, Valentyn Melnychuk, Jonas Schweisthal, Stefan FeuerriegelNeurIPS 2024 · 23 citations
- Conformal Meta-learners for Predictive Inference of Individual Treatment EffectsAhmed M. Alaa, Zaid Ahmad, Mark J. van der LaanNeurIPS 2023 · 32 citations
