Mamba-Driven Multi-View Discriminative Clustering via Global-Local Cross-View Sequence Modeling
Yuanyang Zhang, Xinhang Wan, Chao Zhang, Jie Xu, Cunjian Chen, Tien-Tsin Wong, Li Yao, Yijie Lin
Abstract
Multi-view clustering (MVC) has recently garnered increasing attention for its ability to partition unlabeled samples into distinct clusters by leveraging complementary and consistent information from different views. Existing MVC methods primarily combine deep neural networks with contrastive learning for cross-view representation learning, yet often overlook the inherent global-local structural relationships among samples. While GNN-based methods capture local structures, they struggle to model global dependencies, leading to inferior inter-cluster separability. In contrast, Transformer-based methods excel at global aggregation but suffer from quadratic complexity, and their attention smoothing effect weakens fine-grained local structures, resulting in suboptimal intra-cluster compactness. To address these limitations, we propose a novel end-to-end MVC framework called Mamba-Driven Multi-View Discriminative Clustering via Global-Local Cross-View Sequence Modeling (MGLC). By flexibly constructing multi-view sequences, MGLC fully exploits the efficient sequence modeling capabilities of Mamba to jointly model cross-view dependencies and global-local structural relationships among samples. Furthermore, MGLC introduces a Cross-Mamba Fusion module to dynamically integrate cross-view and global-local structural representations. Additionally, MGLC incorporates a Dual Calibration Contrastive Learning module, guided by high-confidence pseudo-labels, that adaptively refines both feature and semantic representations while mitigating false negatives among semantically similar samples. Extensive comparative experiments and ablation studies demonstrate the effectiveness of MGLC.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 47133d23-caff-4ca5-a3e3-8cf68e6705e5Builds on18
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 3,482 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- Multi-level Feature Learning for Contrastive Multi-view ClusteringJie Xu, Huayi Tang, Yazhou Ren, Liang Peng et al.CVPR 2022 · 335 citations
- DealMVC: Dual Contrastive Calibration for Multi-view ClusteringXihong Yang, Jiaqi Jin, Siwei Wang, Ke Liang et al.ACM MM 2023 · 138 citations
- Self-Weighted Contrastive Learning among Multiple Views for Mitigating Representation DegenerationJie Xu, Shuo Chen, Yazhou Ren, Xiaoshuang Shi et al.NeurIPS 2023 · 71 citations
Related papers
- Deep Multiview Clustering by Contrasting Cluster AssignmentsJie Chen, Hua Mao, Wai Lok Woo, Xi PengICCV 2023 · 142 citations
- Graph based Consistency Learning for Contrastive Multi-View ClusteringBinbin Xu, Jun Yin, Nan ZhangACM MM 2024 · 4 citations
- GCFAgg: Global and Cross-View Feature Aggregation for Multi-View ClusteringWeiqing Yan, Yuanyang Zhang, Chenlei Lv, Chang Tang et al.CVPR 2023
- Dual-stage Contrastive Learning-enhanced Multi-view Variational ClusteringYanxi Liu, Yipin Hu, Fangxi Liu, Yanwei Yu et al.ICML 2026
- Multi-view Granular-ball Contrastive ClusteringPeng Su, Shudong Huang, Weihong Ma, Deng Xiong et al.AAAI 2025 · 13 citations
