A Peer-review Look on Multi-modal Clustering: An Information Bottleneck Realization Method
Zhengzheng Lou, Hang Xue, Chaoyang Zhang, Shizhe Hu
摘要
Despite the superior capability in complementary information exploration and consistent clustering structure learning, most current weight-based multi-modal clustering methods still contain three limitations: 1) lack of trustworthiness in learned weights; 2) isolated view weight learning; 3) extra weight parameters. Motivated by the peer-review mechanism in the academia, we in this paper give a new peer-review look on the multi-modal clustering problem and propose to iteratively treat one modality as "author" and the remaining modalities as "reviewers" so as to reach a peer-review score for each modality. It essentially explores the underlying relationships among modalities. To improve the trustworthiness, we further design a new trustworthy score with a self-supervision working mechanism. Following that, we propose a novel Peer-review Trustworthy Information Bottleneck (PTIB) method for weighted multi-modal clustering, where both the above scores are simultaneously taken into account for accurate and parameter-free modality weight learning. Extensive experiments on eight multi-modal datasets suggest that PTIB can outperform the state-of-theart multi-modal clustering methods. A specific example Reviewer 2 Author review result result Accept Revise Reject Accept Revise Reject General peer-review Author /Reviewer Author /Reviewer Author /Reviewer "Peer-review" look on multimodal clustering Modality 3 Modality 1 Modality 2 A specific case Modality 2 "review" Modality 1 result result 0 1 The score 0 1 The score Modality 3
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Calibrated Information Bottleneck for Trusted Multi-modal ClusteringShizhe Hu, Zhangwen Gou, Shuaiju Li, Jin Qin 等ICLR 2026
- Information-Theoretic Disentangled Latent Modeling with Conditional Diffusion for Incomplete Multi-View ClusteringWenlan Chen, Lu Gao, Daoyuan Wang, Cheng Liang 等ICML 2026
它引用的顶会 Paper11
- Learning Robust Representations via Multi-View Information BottleneckMarco Federici, Anjan Dutta, Patrick Forré, Nate Kushman 等ICLR 2020 · 被引用 330 次
- One Pass Late Fusion Multi-view ClusteringXinwang Liu, Li Liu, Qing Liao, Siwei Wang 等ICML 2021 · 被引用 119 次
- Self-Supervised Graph Attention Networks for Deep Weighted Multi-View ClusteringZongmo Huang, Yazhou Ren, Xiaorong Pu, Shudong Huang 等AAAI 2023 · 被引用 50 次
- Multi-Level Confidence Learning for Trustworthy Multimodal ClassificationXiao Zheng, Chang Tang, Zhiguo Wan, Chengyu Hu 等AAAI 2023 · 被引用 41 次
- DPNET: Dynamic Poly-attention Network for Trustworthy Multi-modal ClassificationXin Zou, Chang Tang, Xiao Zheng, Zhenglai Li 等ACM MM 2023 · 被引用 16 次
相关 Paper
- A Parameter-free Multi-view Information Bottleneck Clustering Method by Cross-view WeightingShizhe Hu, Ruilin Geng, Zhaoxu Cheng, Chaoyang Zhang 等ACM MM 2022 · 被引用 6 次
- Multi-aspect Self-guided Deep Information Bottleneck for Multi-modal ClusteringShizhe Hu, Jiahao Fan, Guoliang Zou, Yangdong YeAAAI 2025 · 被引用 5 次
- Super Deep Contrastive Information Bottleneck for Multi-modal ClusteringZhengzheng Lou, Ke Zhang, Yucong Wu, Shizhe HuICML 2025
- Multi-View Information-Bottleneck Representation LearningZhibin Wan, Changqing Zhang, Pengfei Zhu, Qinghua HuAAAI 2021 · 被引用 116 次
- Learning Optimal Multimodal Information Bottleneck RepresentationsQilong Wu, Yiyang Shao, Jun Wang, Xiaobo SunICML 2025
