Adaptive Sampling for Estimating Probability Distributions
Shubhanshu Shekhar, Tara Javidi, Mohammad Ghavamzadeh
摘要
We consider the problem of allocating a fixed budget of samples to a finite set of discrete distributions to learn them uniformly well (minimizing the maximum error) in terms of four common distance measures: 2 2 , 1 , f -divergence, and separation distance. To present a unified treatment of these distances, we first propose a general optimistic tracking algorithm and analyze its sample allocation performance w.r.t. an oracle. We then instantiate this algorithm for the four distance measures and derive bounds on their regret. We also show that the allocation performance of the proposed algorithm cannot, in general, be improved, by deriving lower-bounds on the expected deviation from the oracle allocation for any adaptive scheme. We verify our theoretical findings through some experiments. Finally, we show that the techniques developed in the paper can be easily extended to learn some classes of continuous distributions as well as to the related setting of minimizing the average error (rather than the maximum error) in learning a set of distributions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Effective Dimension in Bandit Problems under CensorshipGauthier Guinet, Saurabh Amin, Patrick JailletNeurIPS 2022 · 被引用 3 次
- Exploration-free Algorithms for Multi-group Mean EstimationZiyi Wei, Huaiyang Zhong, Xiaocheng LiICML 2026
相关 Paper
- Online Sign Identification: Minimization of the Number of Errors in Thresholding BanditsReda Ouhamma, Odalric-Ambrym Maillard, Vianney PerchetNeurIPS 2021 · 被引用 4 次
- Instance-Optimal Uniformity Testing and TrackingGuy Blanc, Clément L. Canonne, Erik WaingartenFOCS 2025 · 被引用 3 次
- From Finite to Countable-Armed BanditsAnand Kalvit, Assaf ZeeviNeurIPS 2020 · 被引用 15 次
- On-Demand Sampling: Learning Optimally from Multiple DistributionsNika Haghtalab, Michael I. Jordan, Eric ZhaoNeurIPS 2022 · 被引用 57 次
- Beyond the Best: Distribution Functional Estimation in Infinite-Armed BanditsYifei Wang, Tavor Z. Baharav, Yanjun Han, Jiantao Jiao 等NeurIPS 2022 · 被引用 2 次
