Revisiting Active Sequential Prediction-Powered Mean Estimation
Maria-Eleni Sfyraki, Jun-Kun Wang
摘要
In this work, we revisit the problem of active sequential prediction-powered mean estimation, where at each round one must decide the query probability of the ground-truth label upon observing the covariates of a sample. Furthermore, if the label is not queried, the prediction from a machine learning model is used instead. Prior work proposed an elegant scheme that determines the query probability by combining an uncertainty-based suggestion with a constant probability that encodes a soft constraint on the query probability. We explored different values of the mixing parameter and observed an intriguing empirical pattern: the smallest confidence width tends to occur when the weight on the constant probability is close to one, thereby reducing the influence of the uncertainty-based component. Motivated by this observation, we develop a non-asymptotic analysis of the estimator and establish a data-dependent bound on its confidence interval. Our analysis further suggests that when a no-regret learning approach is used to determine the query probability and control this bound, the query probability converges to the constraint of the max value of the query probability when it is chosen obliviously to the current covariates. We also conduct simulations that corroborate these theoretical findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper24
- Using Imperfect Surrogates for Downstream Inference: Design-based Supervised Learning for Social Science Applications of Large Language ModelsNaoki Egami, Musashi Hinck, Brandon M. Stewart, Hanying WeiNeurIPS 2023 · 被引用 74 次
- Active Statistical InferenceTijana Zrnic, Emmanuel J. CandèsICML 2024 · 被引用 34 次
- Prediction-Powered Ranking of Large Language ModelsIvi Chatzi, Eleni Straitouri, Suhas Thejaswi, Manuel Gomez RodriguezNeurIPS 2024 · 被引用 34 次
- High-dimensional Robust Mean Estimation via Gradient DescentYu Cheng, Ilias Diakonikolas, Rong Ge, Mahdi SoltanolkotabiICML 2020 · 被引用 33 次
- Stratified Prediction-Powered Inference for Effective Hybrid Evaluation of Language ModelsAdam Fisch, Joshua Maynez, R. Alex Hofer, Bhuwan Dhingra 等NeurIPS 2024 · 被引用 27 次
相关 Paper
- Active, anytime-valid risk controlling prediction setsZiyu Xu, Nikos Karampatziakis, Paul MineiroNeurIPS 2024 · 被引用 19 次
- Agnostic Continuous-Time Online LearningPramith Devulapalli, Changlong Wu, Ananth Grama, Wojciech SzpankowskiNeurIPS 2025 · 被引用 2 次
- Confidence-Budget Matching for Sequential Budgeted LearningYonathan Efroni, Nadav Merlis, Aadirupa Saha, Shie MannorICML 2021 · 被引用 14 次
- Multinomial Logit Contextual Bandits: Provable Optimality and PracticalityMin-hwan Oh, Garud IyengarAAAI 2021 · 被引用 29 次
- Sequential Mode Estimation with Oracle QueriesDhruti Shah, Tuhinangshu Choudhury, Nikhil Karamchandani, Aditya GopalanAAAI 2020 · 被引用 7 次
