Enhanced Denesity Peak Clustering for High-Dimensional Data
Zhongli Wang, Jie Yang, Junyi Guan, Chenglong Zhang, Xinyan Liang, Bingbing Jiang, Weiguo Sheng
摘要
As a foundational clustering paradigm, Density Peak Clustering (DPC) partitions samples into clusters based on their density peaks, garnering widespread attention. However, traditional DPC methods usually focus on high-density regions, neglecting representative peaks in relatively low-density areas, particularly in datasets with varying densities and multiple peaks. Moreover, existing DPC variants struggle to identify clusters correctly in high-dimensional spaces due to the indistinct distance differences among samples and sparse data distributions. Additionally, existing methods typically adopt a one-step label assignment strategy, making them prone to cascading errors when initial misassignments occur. To address these challenges, we propose an Enhanced Density Peak Clustering (EDPC) method, which creatively incorporates multilayer perceptron (MLP)-based dimensionality reduction and a hierarchical label assignment strategy to significantly improve clustering performance in high-dimensional scenarios. Specifically, we introduce an effective selection condition that combines average densities and density-related distances to generate potential cluster centers, ensuring that peaks across different density regions are considered simultaneously. Furthermore, an MLP, guided by pseudo-labels from sub-clusters, is designed to learn low-dimensional embeddings for high-dimensional data, preserving data locality while enhancing clusterability. Extensive experiments demonstrate the effectiveness and superiority of EDPC against state-of-the-art DPC methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Self-Enhanced Density Clustering for High Dimension and Low Sample Size DataBingbing Jiang, Zhongli Wang, Jie Yang, Guangkui Xu 等KDD 2026
- Fast Density-Peaks Clustering: Multicore-based Parallelization ApproachDaichi Amagata, Takahiro HaraSIGMOD 2021 · 被引用 21 次
- Samples Are Not Equal: A Sample Selection Approach for Deep ClusteringZhengxing Jiao, Yaxin Hou, Jun Ma, Yuhang Li 等ICLR 2026
- Self-reconstructive evidential clustering for high-dimensional dataChaoyu Gong, Yongbin Liu, Di Fu, Yong Liu 等ICDE 2022 · 被引用 6 次
- Top-Down Deep Clustering with Multi-Generator GANsDaniel P. M. de Mello, Renato M. Assunção, Fabricio MuraiAAAI 2022 · 被引用 22 次
