An Evaluative Measure of Clustering Methods Incorporating Hyperparameter Sensitivity
Siddhartha Mishra, Nicholas Monath, Michael Boratko, Ariel Kobren, Andrew McCallum
Abstract
Clustering algorithms are often evaluated using metrics which compare with ground-truth cluster assignments, such as Rand index and NMI. Algorithm performance may vary widely for different hyperparameters, however, and thus model selection based on optimal performance for these metrics is discordant with how these algorithms are applied in practice, where labels are unavailable and tuning is often more art than science. It is therefore desirable to compare clustering algorithms not only on their optimally tuned performance, but also some notion of how realistic it would be to obtain this performance in practice. We propose an evaluation of clustering methods capturing this ease-of-tuning by modeling the expected best clustering score under a given computation budget. To encourage the adoption of the proposed metric alongside classic clustering evaluations, we provide an extensible benchmarking framework. We perform an extensive empirical evaluation of our proposed metric on popular clustering algorithms over a large collection of datasets from different domains, and observe that our new metric leads to several noteworthy observations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext db6f8b44-80eb-4b30-952d-807420778d4cBuilds on2
- MPNet: Masked and Permuted Pre-training for Language UnderstandingKaitao Song, Xu Tan, Tao Qin, Jianfeng Lu et al.NeurIPS 2020 · 1,957 citations
- BoTorch: A Framework for Efficient Monte-Carlo Bayesian OptimizationMaximilian Balandat, Brian Karrer, Daniel R. Jiang, Samuel Daulton et al.NeurIPS 2020 · 686 citations
Related papers
- p-value Adjustment for Monotonous, Unbiased, and Fast Clustering ComparisonKai Klede, Thomas Altstidl, Dario Zanca, Bjoern M. EskofierNeurIPS 2023 · 2 citations
- CHB: A Diagnostic Toolkit for Hardness-Aware Clustering EvaluationWalid Durani, Philipp Jahn, Collin Leiber, David B. Hoffmann et al.ICML 2026
- Systematic Analysis of Cluster Similarity Indices: How to Validate Validation MeasuresMartijn Gösgens, Alexey Tikhonov, Liudmila ProkhorenkovaICML 2021 · 27 citations
- FastAMI - a Monte Carlo Approach to the Adjustment for Chance in Clustering Comparison MetricsKai Klede, Leo Schwinn, Dario Zanca, Björn M. EskofierAAAI 2023 · 2 citations
- Showing Your Offline Reinforcement Learning Work: Online Evaluation Budget MattersVladislav Kurenkov, Sergey KolesnikovICML 2022 · 25 citations
