Clustering by Maximizing Mutual Information Across Views
Kien Do, Truyen Tran, Svetha Venkatesh
Abstract
We propose a novel framework for image clustering that incorporates joint representation learning and clustering. Our method consists of two heads that share the same backbone network - a "representation learning" head and a "clustering" head. The "representation learning" head captures fine-grained patterns of objects at the instance level which serve as clues for the "clustering" head to extract coarse-grain information that separates objects into clusters. The whole model is trained in an end-to-end manner by minimizing the weighted sum of two sample-oriented contrastive losses applied to the outputs of the two heads. To ensure that the contrastive loss corresponding to the "clustering" head is optimal, we introduce a novel critic function called "log-of-dot-product". Extensive experimental results demonstrate that our method significantly outperforms state-of-the-art single-stage clustering methods across a variety of image datasets, improving over the best baseline by about 5-7% in accuracy on CIFAR10/20, STL10, and ImageNet-Dogs. Further, the "two-stage" variant of our method also achieves better results than baselines on three challenging ImageNet subsets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Multimodal Variational Auto-encoder based Audio-Visual SegmentationYuxin Mao, Jing Zhang, Mochu Xiang, Yiran Zhong et al.ICCV 2023 · 57 citations
- Dual Mutual Information Constraints for Discriminative ClusteringHongyu Li, Lefei Zhang, Kehua SuAAAI 2023 · 17 citations
- Generalised Mutual Information for Discriminative ClusteringLouis Ohl, Pierre-Alexandre Mattei, Charles Bouveyron, Warith Harchaoui et al.NeurIPS 2022 · 10 citations
- FreeCOS: Self-Supervised Learning from Fractals and Unlabeled Images for Curvilinear Object SegmentationTianyi Shi, Xiaohuan Ding, Liang Zhang, Xin YangICCV 2023 · 10 citations
- An Information-theoretical Framework for Understanding Out-of-distribution Detection with Pretrained Vision-Language ModelsBo Peng, Jie Lu, Guangquan Zhang, Zhen FangNeurIPS 2025 · 9 citations
Builds on17
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
- Invariant Information Clustering for Unsupervised Image Classification and SegmentationXu Ji, Andrea Vedaldi, João F. HenriquesICCV 2019 · 956 citations
Related papers
- Contrastive ClusteringYunfan Li, Peng Hu, Jerry Zitao Liu, Dezhong Peng et al.AAAI 2021 · 798 citations
- You Never Cluster AloneYuming Shen, Ziyi Shen, Menghan Wang, Jie Qin et al.NeurIPS 2021 · 69 citations
- Clustering-friendly Representation Learning via Instance Discrimination and Feature DecorrelationYaling Tao, Kentaro Takagi, Kouta NakataICLR 2021 · 115 citations
- Stable Cluster Discrimination for Deep ClusteringQi QianICCV 2023 · 38 citations
- Graph Contrastive ClusteringHuasong Zhong, Jianlong Wu, Chong Chen, Jianqiang Huang et al.ICCV 2021 · 163 citations
