Nearly Optimal Catoni's M-estimator for Infinite Variance
Sujay Bhatt, Guanhua Fang, Ping Li, Gennady Samorodnitsky
Abstract
1 In this paper, we extend the remarkable Mestimator of Catoni (2012) to situations where the variance is infinite. In particular, given a sequence of i.i.d random variables X i n i=1 from distribution D over R with mean µ, we only assume the existence of a known upper bound υ ε > 0 on the (1 + ε) th central moment of the random variables, namely, for ε ∈ (0, 1] The extension is non-trivial owing to the difficulty in characterizing the roots of certain polynomials of degree smaller than 2. The proposed estimator has the same order of magnitude and the same asymptotic constant as in Catoni ( 2012 ), but for the case of bounded moments. We further propose a version of the estimator that does not require even the knowledge of υ ε , but adapts the moment bound in a data-driven manner. Finally, to illustrate the usefulness of the derived non-asymptotic confidence bounds, we consider an application in multi-armed bandits and propose best arm identification algorithms, in the fixed confidence setting, that outperform the state of the art.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d755433b-7958-40bc-be37-0916e5fd2df6Cited by top-tier papers3
- Tackling Heavy-Tailed Rewards in Reinforcement Learning with Function Approximation: Minimax Optimal and Instance-Dependent Regret BoundsJiayi Huang, Han Zhong, Liwei Wang, Lin YangNeurIPS 2023 · 16 citations
- Minimax M-estimation under Adversarial ContaminationSujay Bhatt, Guanhua Fang, Ping Li, Gennady SamorodnitskyICML 2022 · 9 citations
- A Simple and Optimal Policy Design for Online Learning with Safety against Heavy-tailed RiskDavid Simchi-Levi, Zeyu Zheng, Feng ZhuNeurIPS 2022 · 7 citations
Related papers
- Best Arm Identification in Contaminated Stochastic BanditsArpan Mukherjee, Ali Tajer, Pin-Yu Chen, Payel DasNeurIPS 2021 · 1 citation
- Covariance-adaptive best arm identificationEl Mehdi Saad, Gilles Blanchard, Nicolas VerzelenNeurIPS 2023 · 1 citation
- Optimal Estimation of the Best Mean in Multi-Armed BanditsTakayuki Osogami, Junya Honda, Junpei KomiyamaNeurIPS 2025
- Optimal Best-Arm Identification Methods for Tail-Risk MeasuresShubhada Agrawal, Wouter M. Koolen, Sandeep JunejaNeurIPS 2021 · 34 citations
- Concentration bounds for CVaR estimation: The cases of light-tailed and heavy-tailed distributionsPrashanth L. A., Krishna P. Jagannathan, Ravi Kumar KollaICML 2020 · 53 citations
