How and Why to Use Experimental Data to Evaluate Methods for Observational Causal Inference
Amanda Gentzel, Purva Pruthi, David D. Jensen
Abstract
Methods that infer causal dependence from observational data are central to many areas of science, including medicine, economics, and the social sciences. A variety of theoretical properties of these methods have been proven, but empirical evaluation remains a challenge, largely due to the lack of observational data sets for which treatment effect is known. We describe and analyze observational sampling from randomized controlled trials (OSRCT), a method for evaluating causal inference methods using data from randomized controlled trials (RCTs). This method can be used to create constructed observational data sets with corresponding unbiased estimates of treatment effect, substantially increasing the number of data sets available for empirical evaluation of causal inference methods. We show that, in expectation, OSRCT creates data sets that are equivalent to those produced by randomly sampling from empirical data sets in which all potential outcomes are available. We then perform a large-scale evaluation of seven causal inference methods over 37 data sets, drawn from RCTs, as well as simulators, real-world computational systems, and observational data sets augmented with a synthetic response variable. We find notable performance differences when comparing across data from different sources, demonstrating the importance of using data from a variety of sources when evaluating any causal inference method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 63f6b56b-4dd3-4218-bf11-bb566f5bcafcCited by top-tier papers7
- Validating Causal Inference MethodsHarsh Parikh, Carlos Varjao, Louise Xu, Eric Tchetgen TchetgenICML 2022 · 36 citations
- Intervention Generalization: A View from Factor Graph ModelsGecia Bravo Hermsdorff, David S. Watson, Jialin Yu, Jakob Zeitler et al.NeurIPS 2023 · 7 citations
- Estimating Causal Effects Identifiable from a Combination of Observations and ExperimentsYonghan Jung, Ivan Diaz, Jin Tian, Elias BareinboimNeurIPS 2023 · 7 citations
- SpaCE: The Spatial Confounding EnvironmentMauricio Tec, Ana Trisovic, Michelle Audirac, Sophie Woodward et al.ICLR 2024 · 6 citations
- Estimating Joint Treatment Effects by Combining Multiple ExperimentsYonghan Jung, Jin Tian, Elias BareinboimICML 2023 · 5 citations
Builds on3
- Bounding Causal Effects on Continuous OutcomeJunzhe Zhang, Elias BareinboimAAAI 2021 · 49 citations
- Counterfactual Cross-Validation: Stable Model Selection Procedure for Causal Inference ModelsYuta Saito, Shota YasuiICML 2020 · 34 citations
- Causal Inference using Gaussian Processes with Structured Latent ConfoundersSam Witty, Kenta Takatsu, David D. Jensen, Vikash MansinghkaICML 2020 · 21 citations
Related papers
- Do Contemporary Causal Inference Models Capture Real-World Heterogeneity? Findings from a Large-Scale BenchmarkHaining Yu, Yizhou SunICLR 2025
- Falsification before Extrapolation in Causal Effect EstimationZeshan M. Hussain, Michael Oberst, Ming-Chieh Shih, David A. SontagNeurIPS 2022 · 11 citations
- Prediction-powered Generalization of Causal InferencesIlker Demirel, Ahmed M. Alaa, Anthony Philippakis, David A. SontagICML 2024 · 18 citations
- Task-specific experimental design for treatment effect estimationBethany Connolly, Kim Moore, Tobias Schwedes, Alexander Adam et al.ICML 2023 · 4 citations
- Language Models as Causal Effect GeneratorsLucius E. J. Bynum, Kyunghyun ChoEMNLP 2025
