Multi-aspect Self-guided Deep Information Bottleneck for Multi-modal Clustering
Shizhe Hu, Jiahao Fan, Guoliang Zou, Yangdong Ye
Abstract
Deep multi-modal clustering can extract useful information among modals, thus benefiting the final clustering and many related fields. However, existing multi-modal clustering methods have two major limitations. First, they often ignore different levels of guiding information from both the feature representations and cluster assignments, which thus are difficult in learning discriminative representations. Second, most methods fail to effectively eliminate redundant information between multi-modal data, negatively affecting clustering results. In this paper, we propose a novel multi-aspect self-guided deep information bottleneck (MSDIB) method for multi-modal clustering, which can effectively employ different aspects of guiding information for learning cluster-friendly information among modals. MSDIB mainly contains two parts: information compression and information preservation. In information compression, we extract from the private information of each modality to obtain the compact representation and meanwhile conduct mutual compression between them. In information preservation, the aim is to preserve the shared information among modals and the self-supervised information from the clustering results in each iteration. In the above process, there are mainly three aspects of self-guiding information, the modality-private information, the modality-shared information and the self-supervised pseudo label information. By minimizing the mutual information based objective function with a variational optimization method, we can fully extract useful discriminative information while eliminating the irrelevant parts. Extensive experimental results demonstrate that our method outperforms state-of-the-art multi-modal clustering methods, showcasing its superior performance and broad application prospects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 03d3e08e-5416-4e2a-a64e-ab9a7796da84Cited by top-tier papers4
- A Peer-review Look on Multi-modal Clustering: An Information Bottleneck Realization MethodZhengzheng Lou, Hang Xue, Chaoyang Zhang, Shizhe HuICML 2025
- Interest-driven Deep Multi-modal ClusteringGuoliang Zou, Tongji Chen, Sijia Li, Jin Qin et al.AAAI 2026
- Calibrated Information Bottleneck for Trusted Multi-modal ClusteringShizhe Hu, Zhangwen Gou, Shuaiju Li, Jin Qin et al.ICLR 2026
- DMCAR: Disentangled Mixture-of-Experts with Context-Aware Routing for Multi-View ClusteringBaili Xiao, Ke Liang, Jiaqi Jin, Jun Wang et al.AAAI 2026
Builds on9
- Multi-level Feature Learning for Contrastive Multi-view ClusteringJie Xu, Huayi Tang, Yazhou Ren, Liang Peng et al.CVPR 2022 · 335 citations
- DealMVC: Dual Contrastive Calibration for Multi-view ClusteringXihong Yang, Jiaqi Jin, Siwei Wang, Ke Liang et al.ACM MM 2023 · 138 citations
- Incomplete Contrastive Multi-View Clustering with High-Confidence GuidingGuoqing Chao, Yi Jiang, Dianhui ChuAAAI 2024 · 135 citations
- Decoupled Contrastive Multi-View Clustering with High-Order Random WalksYiding Lu, Yijie Lin, Mouxing Yang, Dezhong Peng et al.AAAI 2024 · 107 citations
- Deep Mutual Information Maximin for Cross-Modal ClusteringYiqiao Mao, Xiaoqiang Yan, Qiang Guo, Yangdong YeAAAI 2021 · 58 citations
Related papers
- Super Deep Contrastive Information Bottleneck for Multi-modal ClusteringZhengzheng Lou, Ke Zhang, Yucong Wu, Shizhe HuICML 2025
- A Parameter-free Multi-view Information Bottleneck Clustering Method by Cross-view WeightingShizhe Hu, Ruilin Geng, Zhaoxu Cheng, Chaoyang Zhang et al.ACM MM 2022 · 6 citations
- Learning Optimal Multimodal Information Bottleneck RepresentationsQilong Wu, Yiyang Shao, Jun Wang, Xiaobo SunICML 2025
- CDIB: Consistency Discovery-guided Information Bottleneck for Multi-modal Knowledge Graph ReasoningHaichuan Fang, Haoran Zhang, Yulin Du, Qiang Guo et al.ACM MM 2025
- Differentiable Information Bottleneck for Deterministic Multi-View ClusteringXiaoqiang Yan, Zhixiang Jin, Fengshou Han, Yangdong YeCVPR 2024 · 19 citations
