Validating Causal Inference Methods
Harsh Parikh, Carlos Varjao, Louise Xu, Eric Tchetgen Tchetgen
摘要
The fundamental challenge of drawing causal inference is that counterfactual outcomes are not fully observed for any unit. Furthermore, in observational studies, treatment assignment is likely to be confounded. Many statistical methods have emerged for causal inference under unconfoundedness conditions given pre-treatment covariates, including propensity score-based methods, prognostic score-based methods, and doubly robust methods. Unfortunately for applied researchers, there is no `one-size-fits-all' causal method that can perform optimally universally. In practice, causal methods are primarily evaluated quantitatively on handcrafted simulated data. Such data-generative procedures can be of limited value because they are typically stylized models of reality. They are simplified for tractability and lack the complexities of real-world data. For applied researchers, it is critical to understand how well a method performs for the data at hand. Our work introduces a deep generative model-based framework, Credence, to validate causal inference methods. The framework's novelty stems from its ability to generate synthetic data anchored at the empirical distribution for the observed sample, and therefore virtually indistinguishable from the latter. The approach allows the user to specify ground truth for the form and magnitude of causal effects and confounding bias as functions of covariates. Thus simulated data sets are used to evaluate the potential performance of various causal estimation methods when applied to data similar to the observed sample. We demonstrate Credence's ability to accurately assess the relative performance of causal estimation techniques in an extensive simulation study and two real-world data applications from Lalonde and Project STAR studies.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Can You Rely on Your Model Evaluation? Improving Model Evaluation with Synthetic Test DataBoris van Breugel, Nabeel Seedat, Fergus Imrie, Mihaela van der SchaarNeurIPS 2023 · 被引用 51 次
- In Search of Insights, Not Magic Bullets: Towards Demystification of the Model Selection Dilemma in Heterogeneous Treatment Effect EstimationAlicia Curth, Mihaela van der SchaarICML 2023 · 被引用 36 次
- Deep Copula-Based Survival Analysis for Dependent Censoring with Identifiability GuaranteesWeijia Zhang, Chun Kai Ling, Xuanhui ZhangAAAI 2024 · 被引用 12 次
- Marginal Causal Flows for Validation and InferenceDaniel de Vassimon Manela, Laura Battaglia, Robin J. EvansNeurIPS 2024 · 被引用 10 次
- Generative Conditional Distributions by Neural (Entropic) Optimal TransportBao Nguyen, Binh Nguyen, Hieu Trung Nguyen, Viet Anh NguyenICML 2024 · 被引用 2 次
它引用的顶会 Paper1
相关 Paper
- Deep Multi-Modal Structural Equations For Causal Effect Estimation With Unstructured ProxiesShachi Deshpande, Kaiwen Wang, Dhruv Sreenivas, Zheng Li 等NeurIPS 2022 · 被引用 15 次
- Data Fusion for Partial Identification of Causal EffectsQuinn Lanners, Cynthia Rudin, Alexander Volfovsky, Harsh ParikhNeurIPS 2025 · 被引用 5 次
- Causal Inference with Conditional Front-Door Adjustment and Identifiable Variational AutoencoderZiqi Xu, Debo Cheng, Jiuyong Li, Jixue Liu 等ICLR 2024 · 被引用 26 次
- Falsification before Extrapolation in Causal Effect EstimationZeshan M. Hussain, Michael Oberst, Ming-Chieh Shih, David A. SontagNeurIPS 2022 · 被引用 11 次
- Improving the Generation and Evaluation of Synthetic Data for Downstream Medical Causal InferenceHarry Amad, Zhaozhi Qian, Dennis Frauen, Julianna Piskorz 等NeurIPS 2025 · 被引用 6 次
