DeepDPM: Deep Clustering With an Unknown Number of Clusters
Meitar Ronen, Shahaf E. Finder, Oren Freifeld
摘要
Deep Learning (DL) has shown great promise in the unsupervised task of clustering. That said, while in classical (i.e., non-deep) clustering the benefits of the nonparametric approach are well known, most deep-clustering methods are parametric: namely, they require a predefined and fixed number of clusters, denoted by K. When <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> is unknown, however, using model-selection criteria to choose its optimal value might become computationally expensive, especially in DL as the training process would have to be repeated numerous times. In this work, we bridge this gap by introducing an effective deep-clustering method that does not require knowing the value of <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> as it infers it during the learning. Using a split/merge framework, a dynamic architecture that adapts to the changing <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> , and a novel loss, our proposed method outperforms existing nonparametric methods (both classical and deep ones). While the very few existing deep nonparametric methods lack scalability, we demonstrate ours by being the first to report the performance of such a method on ImageNet. We also demonstrate the importance of inferring <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> by showing how methods that fix it deteriorate in performance when their assumed <tex xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"></tex> value gets further from the ground-truth one, especially on imbalanced datasets. Our code is available at https://github.com/BGU-CS-VIL/DeepDPM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Learning Semi-supervised Gaussian Mixture Models for Generalized Category DiscoveryBingchen Zhao, Xin Wen, Kai HanICCV 2023 · 被引用 109 次
- End-to-end Learnable Clustering for Intent Learning in RecommendationYue Liu, Shihao Zhu, Jun Xia, Yingwei Ma 等NeurIPS 2024 · 被引用 56 次
- Visual Recognition with Deep Nearest CentroidsWenguan Wang, Cheng Han, Tianfei Zhou, Dongfang LiuICLR 2023 · 被引用 45 次
- Reinforcement Graph Clustering with Unknown Cluster NumberYue Liu, Ke Liang, Jun Xia, Xihong Yang 等ACM MM 2023 · 被引用 30 次
- Unified 3D Segmenter As Prototypical ClassifiersZheyun Qin, Cheng Han, Qifan Wang, Xiushan Nie 等NeurIPS 2023 · 被引用 27 次
它引用的顶会 Paper9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Unsupervised Pre-Training of Image Features on Non-Curated DataMathilde Caron, Piotr Bojanowski, Julien Mairal, Armand JoulinICCV 2019 · 被引用 254 次
- Deep Clustering by Gaussian Mixture Variational Autoencoders With Graph EmbeddingLinxiao Yang, Ngai-Man Cheung, Jiaying Li, Jun FangICCV 2019 · 被引用 149 次
- You Only Train Once: Loss-Conditional Training of Deep NetworksAlexey Dosovitskiy, Josip DjolongaICLR 2020 · 被引用 96 次
- Video Face Clustering With Unknown Number of ClustersMakarand Tapaswi, Marc T. Law, Sanja FidlerICCV 2019 · 被引用 63 次
相关 Paper
- Dynamic Deep Clustering of High-Dimensional Directional Data via Hyperspherical Embeddings with Bayesian Nonparametric MixturesZhiwen Luo, Wentao Fan, Manar Amayri, Nizar BouguilaKDD 2025 · 被引用 4 次
- Adaptive Prototypical Contrastive Learning for Time Series ClusteringWei LiKDD 2026
- P2OT: Progressive Partial Optimal Transport for Deep Imbalanced ClusteringChuyu Zhang, Hui Ren, Xuming HeICLR 2024 · 被引用 12 次
- Generalised Mutual Information for Discriminative ClusteringLouis Ohl, Pierre-Alexandre Mattei, Charles Bouveyron, Warith Harchaoui 等NeurIPS 2022 · 被引用 10 次
- Learning to Discover Novel Visual Categories via Deep Transfer ClusteringKai Han, Andrea Vedaldi, Andrew ZissermanICCV 2019 · 被引用 378 次
