Disentangling Multi-view Representations Beyond Inductive Bias
Guanzhou Ke, Yang Yu, Guoqing Chao, Xiaoli Wang, Chenyang Xu, Shengfeng He
Abstract
Multi-view (or -modality) representation learning aims to understand the relationships between different view representations. Existing methods disentangle multi-view representations into consistent and view-specific representations by introducing strong inductive biases, which can limit their generalization ability. In this paper, we propose a novel multi-view representation disentangling method that aims to go beyond inductive biases, ensuring both interpretability and generalizability of the resulting representations. Our method is based on the observation that discovering multi-view consistency in advance can determine the disentangling information boundary, leading to a decoupled learning objective. We also found that the consistency can be easily extracted by maximizing the transformation invariance and clustering consistency between views. These observations drive us to propose a two-stage framework. In the first stage, we obtain multi-view consistency by training a consistent encoder to produce semantically-consistent representations across views as well as their corresponding pseudo-labels. In the second stage, we disentangle specificity from comprehensive representations by minimizing the upper bound of mutual information between consistent and comprehensive representations. Finally, we reconstruct the original data by concatenating pseudo-labels and view-specific representations. Our experiments on four multi-view datasets demonstrate that our proposed method outperforms 12 comparison methods in terms of clustering and classification performance. The visualization results also show that the extracted consistency and specificity are compact and interpretable. Our code can be found at https://github.com/Guanzhou-Ke/DMRIB.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 39d3f8e9-7de6-4024-9409-55d555e4d781Cited by top-tier papers4
- Incomplete Contrastive Multi-View Clustering with High-Confidence GuidingGuoqing Chao, Yi Jiang, Dianhui ChuAAAI 2024 · 135 citations
- Regularized Contrastive Partial Multi-view Outlier DetectionYijia Wang, Qianqian Xu, Yangbangyan Jiang, Siran Dai et al.ACM MM 2024 · 8 citations
- Enhancing Multi-view Open-set Learning via Ambiguity Uncertainty Calibration and View-wise DebiasingZihan Fang, Zhiyong Xu, Lan Du, Shide Du et al.ACM MM 2025 · 1 citation
- Hierarchical Cross-View Alignment for Multi-View Clustering via Decoupled Information DistillationTaichun Zhou, Siwei Wang, Zhibin Dong, Jiaqi Jin et al.AAAI 2026
Builds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Learning Robust Representations via Multi-View Information BottleneckMarco Federici, Anjan Dutta, Patrick Forré, Nate Kushman et al.ICLR 2020 · 330 citations
- Multi-VAE: Learning Disentangled View-common and View-peculiar Visual Representations for Multi-view ClusteringJie Xu, Yazhou Ren, Huayi Tang, Xiaorong Pu et al.ICCV 2021 · 158 citations
- Multi-View Information-Bottleneck Representation LearningZhibin Wan, Changqing Zhang, Pengfei Zhu, Qinghua HuAAAI 2021 · 116 citations
Related papers
- Incomplete Multi-View Multi-label Learning via Disentangled Representation and Label Semantic EmbeddingXu Yan, Jun Yin, Jie WenCVPR 2025
- Generalized Information-theoretic Multi-view ClusteringWeitian Huang, Sirui Yang, Hongmin CaiNeurIPS 2023 · 17 citations
- Deep Multiview Clustering by Contrasting Cluster AssignmentsJie Chen, Hua Mao, Wai Lok Woo, Xi PengICCV 2023 · 142 citations
- COMPLETER: Incomplete Multi-View Clustering via Contrastive PredictionYijie Lin, Yuanbiao Gou, Zitao Liu, Boyun Li et al.CVPR 2021
- Graph based Consistency Learning for Contrastive Multi-View ClusteringBinbin Xu, Jun Yin, Nan ZhangACM MM 2024 · 4 citations
