How and Why to Use Experimental Data to Evaluate Methods for Observational Causal Inference
Amanda Gentzel, Purva Pruthi, David D. Jensen
摘要
Methods that infer causal dependence from observational data are central to many areas of science, including medicine, economics, and the social sciences. A variety of theoretical properties of these methods have been proven, but empirical evaluation remains a challenge, largely due to the lack of observational data sets for which treatment effect is known. We describe and analyze observational sampling from randomized controlled trials (OSRCT), a method for evaluating causal inference methods using data from randomized controlled trials (RCTs). This method can be used to create constructed observational data sets with corresponding unbiased estimates of treatment effect, substantially increasing the number of data sets available for empirical evaluation of causal inference methods. We show that, in expectation, OSRCT creates data sets that are equivalent to those produced by randomly sampling from empirical data sets in which all potential outcomes are available. We then perform a large-scale evaluation of seven causal inference methods over 37 data sets, drawn from RCTs, as well as simulators, real-world computational systems, and observational data sets augmented with a synthetic response variable. We find notable performance differences when comparing across data from different sources, demonstrating the importance of using data from a variety of sources when evaluating any causal inference method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Validating Causal Inference MethodsHarsh Parikh, Carlos Varjao, Louise Xu, Eric Tchetgen TchetgenICML 2022 · 被引用 36 次
- Intervention Generalization: A View from Factor Graph ModelsGecia Bravo Hermsdorff, David S. Watson, Jialin Yu, Jakob Zeitler 等NeurIPS 2023 · 被引用 7 次
- Estimating Causal Effects Identifiable from a Combination of Observations and ExperimentsYonghan Jung, Ivan Diaz, Jin Tian, Elias BareinboimNeurIPS 2023 · 被引用 7 次
- SpaCE: The Spatial Confounding EnvironmentMauricio Tec, Ana Trisovic, Michelle Audirac, Sophie Woodward 等ICLR 2024 · 被引用 6 次
- Estimating Joint Treatment Effects by Combining Multiple ExperimentsYonghan Jung, Jin Tian, Elias BareinboimICML 2023 · 被引用 5 次
它引用的顶会 Paper3
- Bounding Causal Effects on Continuous OutcomeJunzhe Zhang, Elias BareinboimAAAI 2021 · 被引用 49 次
- Counterfactual Cross-Validation: Stable Model Selection Procedure for Causal Inference ModelsYuta Saito, Shota YasuiICML 2020 · 被引用 34 次
- Causal Inference using Gaussian Processes with Structured Latent ConfoundersSam Witty, Kenta Takatsu, David D. Jensen, Vikash MansinghkaICML 2020 · 被引用 21 次
相关 Paper
- Do Contemporary Causal Inference Models Capture Real-World Heterogeneity? Findings from a Large-Scale BenchmarkHaining Yu, Yizhou SunICLR 2025
- Falsification before Extrapolation in Causal Effect EstimationZeshan M. Hussain, Michael Oberst, Ming-Chieh Shih, David A. SontagNeurIPS 2022 · 被引用 11 次
- Prediction-powered Generalization of Causal InferencesIlker Demirel, Ahmed M. Alaa, Anthony Philippakis, David A. SontagICML 2024 · 被引用 18 次
- Task-specific experimental design for treatment effect estimationBethany Connolly, Kim Moore, Tobias Schwedes, Alexander Adam 等ICML 2023 · 被引用 4 次
- Language Models as Causal Effect GeneratorsLucius E. J. Bynum, Kyunghyun ChoEMNLP 2025
