Estimate Level Adjustment For Inference With Proxies Under Random Distribution Shifts
Steven Wilkins-Reeves, Alexandra N. M. Darmon, Deeksha Sinha
摘要
In many scientific domains, including experimentation, researchers rely on measurements of proxy outcomes to achieve faster and more frequent reads, especially when the primary outcome of interest is challenging to measure directly. While proxies offer a more readily accessible observation for inference, the ultimate goal is to draw statistical inferences about the primary outcome parameter; and proxy data are typically imperfect in some ways. To correct for these imperfections, current statistical inference methods often depend on strict identifying assumptions (such as surrogacy, covariate/label shift, or missingness assumptions). These assumptions can be difficult to validate and may be violated by various additional sources of distribution shift, potentially leading to biased parameter estimates and miscalibrated uncertainty quantification. We introduce an estimate-level framework, inspired by domain adaptation methods, to empirically calibrate proxy-based inference. This framework models the proxy–primary metric discrepancy as a random effect at the parameter level, estimating its distribution from aggregated historical observations across past domains (e.g., experiments, time periods, or distinct segments). This method avoids the requirement for retaining individual-level response data. Additionally, this adjustment can be layered on top of existing proxy-correction methods (such as prediction-powered inference or importance weighting) to account for additional biases not addressed by those corrections. To manage uncertainty when the number of historical domains is limited, we provide both a method-of-moments estimator and a domain bootstrap procedure. Through extensive simulations, our adjusted intervals demonstrate improved calibration and more reliable coverage across various proxy procedures and deviations from standard covariate shift mechanisms. We further validate this approach using publicly available datasets and real-world experiments. Code for replicating simulations and experiments on public datasets are available at https://www.github.com/facebookresearch/estimate-level-adjustment.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Synthetic-powered predictive inferenceMeshi Bashari, Roy Maor Lotan, Yonghoon Lee, Edgar Dobriban 等NeurIPS 2025 · 被引用 12 次
- Inferring the Long-Term Causal Effects of Long-Term Treatments from Short-Term ExperimentsAllen Tran, Aurélien Bibaut, Nathan KallusICML 2024 · 被引用 11 次
- Learning the Covariance of Treatment Effects Across Many Weak ExperimentsAurélien Bibaut, Winston Chou, Simon Ejdemyr, Nathan KallusKDD 2024 · 被引用 2 次
相关 Paper
- Using Surrogates in Covariate-adjusted Response-adaptive Randomization Experiments with Delayed OutcomesLei Shi, Waverly Wei, Jingshen WangNeurIPS 2024 · 被引用 4 次
- Diffusion-Based Probabilistic Uncertainty Estimation for Active Domain AdaptationZhekai Du, Jingjing LiNeurIPS 2023 · 被引用 32 次
- PAC Prediction Sets Under Label ShiftWenwen Si, Sangdon Park, Insup Lee, Edgar Dobriban 等ICLR 2024 · 被引用 15 次
- MEC: Machine-Learning-Assisted Generalized Entropy Calibration for Semi-Supervised Mean EstimationSe Yoon Lee, Jae-kwang KimICML 2026 · 被引用 1 次
- IW-GAE: Importance weighted group accuracy estimation for improved calibration and model selection in unsupervised domain adaptationTaejong Joo, Diego KlabjanICML 2024 · 被引用 1 次
