Bregman Power k-Means for Clustering Exponential Family Data
Adithya Vellal, Saptarshi Chakraborty, Jason Q. Xu
摘要
Recent progress in center-based clustering algorithms combats poor local minima by implicit annealing, using a family of generalized means. These methods are variations of Lloyd's celebrated -means algorithm, and are most appropriate for spherical clusters such as those arising from Gaussian data. In this paper, we bridge these algorithmic advances to classical work on hard clustering under Bregman divergences, which enjoy a bijection to exponential family distributions and are thus well-suited for clustering objects arising from a breadth of data generating mechanisms. The elegant properties of Bregman divergences allow us to maintain closed form updates in a simple and transparent algorithm, and moreover lead to new theoretical arguments for establishing finite sample bounds that relax the bounded support assumption made in the existing state of the art. Additionally, we consider thorough empirical analyses on simulated experiments and a case study on rainfall data, finding that the proposed method outperforms existing peer methods in a variety of non-Gaussian data settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Efficient constrained sampling via the mirror-Langevin algorithmKwangjun Ahn, Sinho ChewiNeurIPS 2021 · 被引用 77 次
- Concentration inequalities under sub-Gaussian and sub-exponential conditionsAndreas Maurer, Massimiliano PontilNeurIPS 2021 · 被引用 39 次
- Uniform Concentration Bounds toward a Unified Framework for Robust ClusteringDebolina Paul, Saptarshi Chakraborty, Swagatam Das, Jason Q. XuNeurIPS 2021 · 被引用 19 次
相关 Paper
- Modified K-means Algorithm with Local Optimality GuaranteesMingyi Li, Michael R. Metel, Akiko TakedaICML 2025
- Gradient Based ClusteringAleksandar Armacki, Dragana Bajovic, Dusan Jakovetic, Soummya KarICML 2022 · 被引用 11 次
- Efficient Algorithms for Sum-Of-Minimum OptimizationLisang Ding, Ziang Chen, Xinshang Wang, Wotao YinICML 2024 · 被引用 7 次
- Achieving Optimal Clustering in Gaussian Mixture Models with Anisotropic Covariance StructuresXin Chen, Anderson Ye ZhangNeurIPS 2024 · 被引用 15 次
- Likelihood Adjusted Semidefinite Programs for Clustering Heterogeneous DataYubo Zhuang, Xiaohui Chen, Yun YangICML 2023 · 被引用 2 次
