Falsification before Extrapolation in Causal Effect Estimation
Zeshan M. Hussain, Michael Oberst, Ming-Chieh Shih, David A. Sontag
Abstract
Randomized Controlled Trials (RCTs) represent a gold standard when developing policy guidelines. However, RCTs are often narrow, and lack data on broader populations of interest. Causal effects in these populations are often estimated using observational datasets, which may suffer from unobserved confounding and selection bias. Given a set of observational estimates (e.g. from multiple studies), we propose a meta-algorithm that attempts to reject observational estimates that are biased. We do so using validation effects, causal effects that can be inferred from both RCT and observational data. After rejecting estimators that do not pass this test, we generate conservative confidence intervals on the extrapolated causal effects for subgroups not observed in the RCT. Under the assumption that at least one observational estimator is asymptotically normal and consistent for both the validation and extrapolated effects, we provide guarantees on the coverage probability of the intervals output by our algorithm. To facilitate hypothesis testing in settings where causal effect transportation across datasets is necessary, we give conditions under which a doubly-robust estimator of group average treatment effects is asymptotically normal, even when flexible machine learning methods are used for estimation of nuisance parameters. We illustrate the properties of our approach on semi-synthetic and real world datasets, and show that it compares favorably to standard meta-analysis techniques.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 189f0fea-4d97-47b2-a32b-b85d76b25d57Cited by top-tier papers5
- Prediction-powered Generalization of Causal InferencesIlker Demirel, Ahmed M. Alaa, Anthony Philippakis, David A. SontagICML 2024 · 18 citations
- Addressing Hidden Confounding with Heterogeneous Observational Datasets for RecommendationYanghao Xiao, Haoxuan Li, Yongqiang Tang, Wensheng ZhangNeurIPS 2024 · 15 citations
- Falsification of Unconfoundedness by Testing Independence of Causal MechanismsRickard Karlsson, Jesse H. KrijtheICML 2025
- Doubly robust identification of treatment effects from multiple environmentsPiersilvio De Bartolomeis, Julia Kostin, Javier Abad, Yixin Wang et al.ICLR 2025
- Uncovering Bias Mechanisms in Observational StudiesIlker Demirel, Zeshan Hussain, Piersilvio De Bartolomeis, David SontagICML 2026
Builds on1
Related papers
- Comparison of meta-learners for estimating multi-valued treatment heterogeneous effectsNaoufal Acharki, Ramiro Lugo, Antoine Bertoncello, Josselin GarnierICML 2023 · 18 citations
- Meta-Learners for Partially-Identified Treatment Effects Across Multiple EnvironmentsJonas Schweisthal, Dennis Frauen, Mihaela van der Schaar, Stefan FeuerriegelICML 2024 · 10 citations
- Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect EstimatorsYiyan Huang, Cheuk Hang Leung, Siyi Wang, Yijun Li et al.NeurIPS 2024 · 2 citations
- Quantifying Ignorance in Individual-Level Causal-Effect Estimates under Hidden ConfoundingAndrew Jesson, Sören Mindermann, Yarin Gal, Uri ShalitICML 2021 · 66 citations
- Conformal Meta-learners for Predictive Inference of Individual Treatment EffectsAhmed M. Alaa, Zaid Ahmad, Mark J. van der LaanNeurIPS 2023 · 32 citations
