Enhanced Denesity Peak Clustering for High-Dimensional Data
Zhongli Wang, Jie Yang, Junyi Guan, Chenglong Zhang, Xinyan Liang, Bingbing Jiang, Weiguo Sheng
Abstract
As a foundational clustering paradigm, Density Peak Clustering (DPC) partitions samples into clusters based on their density peaks, garnering widespread attention. However, traditional DPC methods usually focus on high-density regions, neglecting representative peaks in relatively low-density areas, particularly in datasets with varying densities and multiple peaks. Moreover, existing DPC variants struggle to identify clusters correctly in high-dimensional spaces due to the indistinct distance differences among samples and sparse data distributions. Additionally, existing methods typically adopt a one-step label assignment strategy, making them prone to cascading errors when initial misassignments occur. To address these challenges, we propose an Enhanced Density Peak Clustering (EDPC) method, which creatively incorporates multilayer perceptron (MLP)-based dimensionality reduction and a hierarchical label assignment strategy to significantly improve clustering performance in high-dimensional scenarios. Specifically, we introduce an effective selection condition that combines average densities and density-related distances to generate potential cluster centers, ensuring that peaks across different density regions are considered simultaneously. Furthermore, an MLP, guided by pseudo-labels from sub-clusters, is designed to learn low-dimensional embeddings for high-dimensional data, preserving data locality while enhancing clusterability. Extensive experiments demonstrate the effectiveness and superiority of EDPC against state-of-the-art DPC methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b052cb39-f37b-4c7b-ada7-8f11e396affbBuilds on1
Related papers
- Self-Enhanced Density Clustering for High Dimension and Low Sample Size DataBingbing Jiang, Zhongli Wang, Jie Yang, Guangkui Xu et al.KDD 2026
- Fast Density-Peaks Clustering: Multicore-based Parallelization ApproachDaichi Amagata, Takahiro HaraSIGMOD 2021 · 21 citations
- Samples Are Not Equal: A Sample Selection Approach for Deep ClusteringZhengxing Jiao, Yaxin Hou, Jun Ma, Yuhang Li et al.ICLR 2026
- Self-reconstructive evidential clustering for high-dimensional dataChaoyu Gong, Yongbin Liu, Di Fu, Yong Liu et al.ICDE 2022 · 6 citations
- Top-Down Deep Clustering with Multi-Generator GANsDaniel P. M. de Mello, Renato M. Assunção, Fabricio MuraiAAAI 2022 · 22 citations
