Refined Convergence Rates for Maximum Likelihood Estimation under Finite Mixture Models
Tudor A. Manole, Nhat Ho
摘要
We revisit the classical problem of deriving convergence rates for the maximum likelihood estimator (MLE) in finite mixture models. The Wasserstein distance has become a standard loss function for the analysis of parameter estimation in these models, due in part to its ability to circumvent label switching and to accurately characterize the behaviour of fitted mixture components with vanishing weights. However, the Wasserstein distance is only able to capture the worst-case convergence rate among the remaining fitted mixture components. We demonstrate that when the log-likelihood function is penalized to discourage vanishing mixing weights, stronger loss functions can be derived to resolve this shortcoming of the Wasserstein distance. These new loss functions accurately capture the heterogeneity in convergence rates of fitted mixture components, and we use them to sharpen existing pointwise and uniform convergence rates in various classes of mixture models. In particular, these results imply that a subset of the components of the penalized MLE typically converge significantly faster than could have been anticipated from past work. We further show that some of these conclusions extend to the traditional MLE. Our theoretical findings are supported by a simulation study to illustrate these improved convergence rates.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- FuseMoE: Mixture-of-Experts Transformers for Fleximodal FusionXing Han, Huy Nguyen, Carl Harris, Nhat Ho 等NeurIPS 2024 · 被引用 129 次
- Mixture of Experts Meets Prompt-Based Continual LearningMinh Le, An Nguyen The, Huy Nguyen, Trang Nguyen 等NeurIPS 2024 · 被引用 57 次
- Demystifying Softmax Gating Function in Gaussian Mixture of ExpertsHuy Nguyen, TrungTin Nguyen, Nhat HoNeurIPS 2023 · 被引用 44 次
- Sigmoid Gating is More Sample Efficient than Softmax Gating in Mixture of ExpertsHuy Nguyen, Nhat Ho, Alessandro RinaldoNeurIPS 2024 · 被引用 35 次
- Statistical Perspective of Top-K Sparse Softmax Gating Mixture of ExpertsHuy Nguyen, Pedram Akbarian, Fanqi Yan, Nhat HoICLR 2024 · 被引用 29 次
相关 Paper
- Solving General Elliptical Mixture Models through an Approximate Wasserstein ManifoldShengxi Li, Zeyang Yu, Min Xiang, Danilo P. MandicAAAI 2020 · 被引用 6 次
- On Excess Mass Behavior in Gaussian Mixture Models with Orlicz-Wasserstein DistancesAritra Guha, Nhat Ho, XuanLong NguyenICML 2023 · 被引用 8 次
- Beyond black box densities: Parameter learning for the deviated componentsDat Do, Nhat Ho, XuanLong NguyenNeurIPS 2022 · 被引用 2 次
- Quantifying the Empirical Wasserstein Distance to a Set of Measures: Beating the Curse of DimensionalityNian Si, Jose H. Blanchet, Soumyadip Ghosh, Mark S. SquillanteNeurIPS 2020 · 被引用 16 次
- Normalized Wasserstein for Mixture Distributions With Applications in Adversarial Learning and Domain AdaptationYogesh Balaji, Rama Chellappa, Soheil FeiziICCV 2019 · 被引用 53 次
