Machine Learning for Variance Reduction in Online Experiments
Yongyi Guo, Dominic Coey, Mikael Konutgan, Wenting Li, Chris Schoener, Matt Goldman
Abstract
We consider the problem of variance reduction in randomized controlled trials, through the use of covariates correlated with the outcome but independent of the treatment. We propose a machine learning regression-adjusted treatment effect estimator, which we call MLRATE. MLRATE uses machine learning predictors of the outcome to reduce estimator variance. It employs cross-fitting to avoid overfitting biases, and we prove consistency and asymptotic normality under general conditions. MLRATE is robust to poor predictions from the machine learning step: if the predictions are uncorrelated with the outcomes, the estimator performs asymptotically no worse than the standard difference-in-means estimator, while if predictions are highly correlated with outcomes, the efficiency gains are large. In A/A tests, for a set of 48 outcome metrics commonly monitored in Facebook experiments the estimator has over 70% lower variance than the simple difference-in-means estimator, and about 19% lower variance than the common univariate procedure which adjusts only for pre-experiment values of the outcome.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ce300c11-efbe-4a50-93fa-b49550f2d00dCited by top-tier papers5
- Estimating Distributional Treatment Effects in Randomized Experiments: Machine Learning for Variance ReductionUndral Byambadalai, Tatsushi Oka, Shota YasuiICML 2024 · 7 citations
- AI-Assisted Variance Reduction in Randomized ExperimentsDavid Arbour, Eli Ben-Michael, Avi Feller, Apoorva Lal et al.KDD 2026 · 4 citations
- A More Accurate Algorithm Comparison through A/B Testing using Offline Evaluation MethodsKoki Konishi, Masataka Ushiku, Yuta SaitoKDD 2026 · 1 citation
- Strategic A/B testing via Maximum Probability-driven Two-armed BanditYu Zhang, Shanshan Zhao, Bokui Wan, Jinjuan Wang et al.ICML 2025
- On Efficient Estimation of Distributional Treatment Effects under Covariate-Adaptive RandomizationUndral Byambadalai, Tomu Hirata, Tatsushi Oka, Shota YasuiICML 2025
Related papers
- Modeling Covariate Transition for Efficient Estimation of Longitudinal Treatment Effects in Randomized ExperimentsNaoki Chihara, Tatsushi Oka, Yasuko Matsubara, Yasushi Sakurai et al.ICML 2026
- Coordinated Double Machine LearningNitai Fingerhut, Matteo Sesia, Yaniv RomanoICML 2022 · 5 citations
- Estimating individual treatment effects under unobserved confounding using binary instrumentsDennis Frauen, Stefan FeuerriegelICLR 2023 · 3 citations
- Comparison of meta-learners for estimating multi-valued treatment heterogeneous effectsNaoufal Acharki, Ramiro Lugo, Antoine Bertoncello, Josselin GarnierICML 2023 · 18 citations
- A Meta-learner for Heterogeneous Effects in Difference-in-DifferencesHui Lan, Haoge Chang, Eleanor Wiske Dillon, Vasilis SyrgkanisICML 2025
