Efficient Randomized Experiments Using Foundation Models
Piersilvio De Bartolomeis, Javier Abad, Guanbo Wang, Konstantin Donhauser, Raymond M. Duch, Fanny Yang, Issa J. Dahabreh
Abstract
Randomized experiments are the preferred approach for evaluating the effects of interventions, but they are costly and often yield estimates with substantial uncertainty. On the other hand, in silico experiments leveraging foundation models offer a cost-effective alternative that can potentially attain higher statistical precision. However, the benefits of in silico experiments come with a significant risk: statistical inferences are not valid if the models fail to accurately predict experimental responses to interventions. In this paper, we propose a novel approach that integrates the predictions from multiple foundation models with experimental data while preserving valid statistical inference. Our estimator is consistent and asymptotically normal, with asymptotic variance no larger than the standard estimator based on experimental data alone. Importantly, these statistical properties hold even when model predictions are arbitrarily biased. Empirical results across several randomized experiments show that our estimator offers substantial precision gains, equivalent to a reduction of up to 20% in the sample size needed to match the same precision as the standard estimator based on experimental data alone 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext caca396e-0ba9-40ed-8a66-d4a9a9d2c9c2Cited by top-tier papers5
- Prediction-Powered Causal InferencesRiccardo Cadei, Ilker Demirel, Piersilvio De Bartolomeis, Lukas Lindorfer et al.NeurIPS 2025 · 9 citations
- General Synthetic-Powered InferenceMeshi Bashari, Yonghoon Lee, Roy Lotan, Edgar Dobriban et al.ICML 2026 · 5 citations
- Revisiting Active Sequential Prediction-Powered Mean EstimationMaria-Eleni Sfyraki, Jun-Kun WangICLR 2026 · 4 citations
- AI-Assisted Variance Reduction in Randomized ExperimentsDavid Arbour, Eli Ben-Michael, Avi Feller, Apoorva Lal et al.KDD 2026 · 4 citations
- How can we assess human-agent interactions? Case studies in software agent designValerie Chen, Rohit Malhotra, Xingyao Wang, Juan Michelini et al.ICML 2026
Builds on5
- Using Imperfect Surrogates for Downstream Inference: Design-based Supervised Learning for Social Science Applications of Large Language ModelsNaoki Egami, Musashi Hinck, Brandon M. Stewart, Hanying WeiNeurIPS 2023 · 74 citations
- End-To-End Causal Effect Estimation from Unstructured Natural Language DataNikita Dhawan, Leonardo Cotta, Karen Ullrich, Rahul G. Krishnan et al.NeurIPS 2024 · 24 citations
- Prediction-powered Generalization of Causal InferencesIlker Demirel, Ahmed M. Alaa, Anthony Philippakis, David A. SontagICML 2024 · 18 citations
- No Free Lunch: Non-Asymptotic Analysis of Prediction-Powered InferencePranav Mani, Peng Xu, Zachary Lipton, Michael OberstICML 2026 · 8 citations
- Limits to scalable evaluation at the frontier: LLM as judge won't beat twice the dataFlorian E. Dorner, Vivian Yvonne Nastl, Moritz HardtICLR 2025
Related papers
- Constructing Confidence Intervals for Average Treatment Effects from Multiple DatasetsYuxin Wang, Maresa Schröder, Dennis Frauen, Jonas Schweisthal et al.ICLR 2025
- Estimate Level Adjustment For Inference With Proxies Under Random Distribution ShiftsSteven Wilkins-Reeves, Alexandra N. M. Darmon, Deeksha SinhaKDD 2026 · 1 citation
- Estimating Joint Treatment Effects by Combining Multiple ExperimentsYonghan Jung, Jin Tian, Elias BareinboimICML 2023 · 5 citations
- Bridging Domain Expertise and Generalization for Performance EstimationShuxuan Li, Zhilin Zhao, Quyu Kong, Wei-Shi ZhengCVPR 2026 · 1 citation
- Estimating Distributional Treatment Effects in Randomized Experiments: Machine Learning for Variance ReductionUndral Byambadalai, Tatsushi Oka, Shota YasuiICML 2024 · 7 citations
