Observationally Informed Adaptive Causal Experimental Design
Erdun Gao, Liang Zhang, Jake Fawkes, Aoqi Zuo, Wenqin Liu, Haoxuan Li, Mingming Gong, Dino Sejdinovic
摘要
Randomized Controlled Trials (RCTs) represent the gold standard for causal inference yet remain a scarce resource. While large-scale observational data is often available, it is typically used only for retrospective fusion, and remains discarded in prospective trial design due to bias concerns. We argue that this ''tabula rasa'' data acquisition strategy is inefficient when observational models contain useful structural information. In this work, we propose Active Residual Learning, a new paradigm that leverages the observational model as an informative but biased prior. This shifts the experimental focus from learning target causal quantities from scratch to estimating residual corrections that debias the observational model. To operationalize this, we introduce the R-Design framework. Theoretically, we characterize two key advantages: (1) a conditional structural efficiency gap, showing that estimating lower-complexity residual contrasts can admit faster convergence rates than reconstructing full outcomes; and (2) information efficiency, where we quantify the redundancy in standard parameter-based acquisition, demonstrating that such baselines can waste budget on task-irrelevant nuisance uncertainty. We propose R-EPIG (Residual Expected Predictive Information Gain), a unified criterion that directly targets the downstream causal quantity, reducing residual uncertainty for estimation or clarifying decision boundaries for policy. Experiments on synthetic and semi-synthetic benchmarks show that R-Design significantly outperforms baselines in the intended informative-but-biased regime, supporting the effectiveness of correcting a biased model rather than learning from scratch.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper18
- TabPFN: A Transformer That Solves Small Tabular Classification Problems in a SecondNoah Hollmann, Samuel Müller, Katharina Eggensperger, Frank HutterICLR 2023 · 被引用 96 次
- Active Bayesian Causal InferenceChristian Toth, Lars Lorch, Christian Knoll, Andreas Krause 等NeurIPS 2022 · 被引用 52 次
- CausalPFN: Amortized Causal Effect Estimation via In-Context LearningVahid Balazadeh Meresht, Hamidreza Kamkari, Valentin Thomas, Junwei Ma 等NeurIPS 2025 · 被引用 52 次
- Causal-BALD: Deep Bayesian Active Learning of Outcomes to Infer Treatment-Effects from Observational DataAndrew Jesson, Panagiotis Tigas, Joost van Amersfoort, Andreas Kirsch 等NeurIPS 2021 · 被引用 42 次
- Transductive Active Learning: Theory and ApplicationsJonas Hübotter, Bhavya Sukhija, Lenart Treven, Yarden As 等NeurIPS 2024 · 被引用 24 次
相关 Paper
- Causal-EPIG: Causally Aligned Active CATE EstimationErdun Gao, Jake Fawkes, Dino SejdinovicICML 2026 · 被引用 3 次
- ActiveCQ: Active Estimation of Causal QuantitiesErdun Gao, Dino SejdinovicICLR 2026 · 被引用 1 次
- Budgeted Active Experimentation for Treatment Effect Estimation from Observational and Randomized DataJiacan Gao, Xinyan Su, Mingyuan Ma, Yiyan HUANG 等ICML 2026 · 被引用 1 次
- ABC3: Active Bayesian Causal Inference with Cohn Criteria in Randomized ExperimentsTaehun Cha, Donghun LeeAAAI 2025
- Learning-To-Measure: In-Context Active Feature AcquisitionYuta Kobayashi, Zilin Jing, Jiayu Yao, Hongseok Namkoong 等ICML 2026 · 被引用 2 次
