CHB: A Diagnostic Toolkit for Hardness-Aware Clustering Evaluation
Walid Durani, Philipp Jahn, Collin Leiber, David B. Hoffmann, Thomas Seidl, Claudia Plant, Christian Böhm
摘要
Clustering methods are commonly compared through leaderboards that collapse performance into a single aggregated ranking. Such summaries do not reveal why methods succeed, which data properties align with failure, and how conclusions shift under representation changes and realistic tuning constraints. We present the Clustering Hardness Benchmark (CHB), a diagnostic toolkit for hardness-aware clustering via external evaluation. CHB maps each dataset to an interpretable hardness fingerprint capturing (i) separation, (ii) cohesion and scale heterogeneity, and (iii) topology. Using this diagnostic space, CHB evaluates clustering algorithms under standardized default configurations and budgeted hyperparameter tuning. Conditioning results on hardness coordinates turns comparison into diagnosis: across a broad range of datasets and their representations, CHB reveals reproducible structural regimes, uncovers regime-dependent ranking across method families, and surfaces robustness signatures, including topology-linked breakdowns. CHB further enables representation auditing by attributing gains to measurable shifts in the hardness fingerprint rather than just external performance changes. We release CHB as an open, extensible artifact for evaluating new clustering methods and embeddings within a shared diagnostic framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- An Evaluative Measure of Clustering Methods Incorporating Hyperparameter SensitivitySiddhartha Mishra, Nicholas Monath, Michael Boratko, Ariel Kobren 等AAAI 2022 · 被引用 6 次
- The ParClusterers Benchmark Suite (PCBS): A Fine-Grained Analysis of Scalable Graph ClusteringShangdi Yu, Jessica Shi, Jamison Meindl, David Eisenstat 等VLDB 2025 · 被引用 1 次
- Dissecting Sample Hardness: A Fine-Grained Analysis of Hardness Characterization Methods for Data-Centric AINabeel Seedat, Fergus Imrie, Mihaela van der SchaarICLR 2024 · 被引用 16 次
- Beyond Supervised vs. Unsupervised: Representative Benchmarking and Analysis of Image Representation LearningMatthew Gwilliam, Abhinav ShrivastavaCVPR 2022 · 被引用 14 次
- Interactive Deep Clustering via Value MiningHonglin Liu, Peng Hu, Changqing Zhang, Yunfan Li 等NeurIPS 2024 · 被引用 24 次
