Towards Calibrated Deep Clustering Network
Yuheng Jia, Jianhong Cheng, Hui Liu, Junhui Hou
摘要
Deep clustering has exhibited remarkable performance; however, the overconfidence problem, i.e., the estimated confidence for a sample belonging to a particular cluster greatly exceeds its actual prediction accuracy, has been overlooked in prior research. To tackle this critical issue, we pioneer the development of a calibrated deep clustering framework. Specifically, we propose a novel dualhead (calibration head and clustering head) deep clustering model that can effectively calibrate the estimated confidence and the actual accuracy. The calibration head adjusts the overconfident predictions of the clustering head, generating prediction confidence that matches the model learning status. Then, the clustering head dynamically selects reliable high-confidence samples estimated by the calibration head for pseudo-label self-training. Additionally, we introduce an effective network initialization strategy that enhances both training speed and network robustness. The effectiveness of the proposed calibration approach and initialization strategy are both endorsed with solid theoretical guarantees. Extensive experiments demonstrate the proposed calibrated deep clustering model not only surpasses the state-of-the-art deep clustering methods by 5× on average in terms of expected calibration error, but also significantly outperforms them in terms of clustering accuracy. Code is available at https://github.com/ChengJianH/CDC.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Mini-cluster Guided Long-tailed Deep ClusteringZhixin Li, Yuheng Jia, Guanliang Chen, Hui Liu 等ICLR 2026 · 被引用 11 次
- You Can Trust Your Clustering Model: A Parameter-free Self-Boosting Plug-in for Deep ClusteringHanyang Li, Yuheng Jia, Hui Liu, Junhui HouNeurIPS 2025 · 被引用 2 次
- ESMC: MLLM-Based Embedding Selection for Explainable Multiple ClusteringXinyue Wang, Yuheng Jia, Hui Liu, Junhui HouAAAI 2026 · 被引用 1 次
- DiCaP: Distribution-Calibrated Pseudo-labeling for Semi-Supervised Multi-Label LearningBo Han, Zhuoming Li, Xiaoyu Wang, Yaxin Hou 等AAAI 2026
- Samples Are Not Equal: A Sample Selection Approach for Deep ClusteringZhengxing Jiao, Yaxin Hou, Jun Ma, Yuhang Li 等ICLR 2026
它引用的顶会 Paper19
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo LabelingBowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu 等NeurIPS 2021 · 被引用 1,389 次
- Contrastive ClusteringYunfan Li, Peng Hu, Jerry Zitao Liu, Dezhong Peng 等AAAI 2021 · 被引用 798 次
- Calibrating Deep Neural Networks using Focal LossJishnu Mukhoti, Viveka Kulharia, Amartya Sanyal, Stuart Golodetz 等NeurIPS 2020 · 被引用 674 次
相关 Paper
- Deep Semantic Clustering by Partition Confidence MaximisationJiabo Huang, Shaogang Gong, Xiatian ZhuCVPR 2020
- Calibrated Information Bottleneck for Trusted Multi-modal ClusteringShizhe Hu, Zhangwen Gou, Shuaiju Li, Jin Qin 等ICLR 2026
- Improving Unsupervised Image Clustering With Robust LearningSungwon Park, Sungwon Han, Sundong Kim, Danu Kim 等CVPR 2021
- Deep Graph Clustering via Dual Correlation ReductionYue Liu, Wenxuan Tu, Sihang Zhou, Xinwang Liu 等AAAI 2022 · 被引用 300 次
- Be Confident! Towards Trustworthy Graph Neural Networks via Confidence CalibrationXiao Wang, Hongrui Liu, Chuan Shi, Cheng YangNeurIPS 2021 · 被引用 158 次
