Rethinking Multi-View Representation Learning via Distilled Disentangling
Guanzhou Ke, Bo Wang, Xiaoli Wang, Shengfeng He
Abstract
We utilized the Mutual Information Neural Estimator (MINE) 1 [3] as a mutual information estimator to independently assess the mutual information between viewconsistent representations and view-specific representations proposed by CONAN 2 [11], DVIB 3 [2], Multi-VAE 4 [25], and our approach. To ensure a fair comparison, we standardized the representation dimensions of all comparative methods to 10. For constructing the MINE estimator, we employed fully connected layers with Rectified Linear Unit (ReLU) activation, specifying the network architecture as 20-100-100-100-1. We use Adam with the learning rate of 1×10 -4 and the batch size of 128 to train the model for 500 epochs. To mitigate randomness, we executed the MINE procedure 10 times and recorded the average results. B. Related Work Multi-view Representation Learning. The goal of MvRL is to extract both shared and view-specific information from multiple data sources, integrating them into a cohesive representation that is advantageous for predictive tasks [5, 13, 16] . Existing approaches in this field generally fall into three categories: statistic-based, deep learningbased, and hybrid methods. Statistic-based methods, employing techniques like canonical correlation analysis [6, 15] , non-negative matrix factorization [14, 23] , and subspace methods [4, 22] , excel in deriving interpretable models. However, they struggle with datasets that are high-dimensional or large-scale. In contrast, deep learning-based methods have gained prominence, especially in unsupervised settings, where generative models such as autoencoders [1, 21, 27] and generative adversarial networks [29] are used to learn latent representations. Although effective, these methods face the challenge
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Incomplete Multi-view Clustering via Diffusion Contrastive GenerationYuanyang Zhang, Yijie Lin, Weiqing Yan, Li Yao et al.AAAI 2025 · 19 citations
- D3still: Decoupled Differential Distillation for Asymmetric Image RetrievalYi Xie, Yihong Lin, Wenjie Cai, Xuemiao Xu et al.CVPR 2024 · 10 citations
- Hypergraph-Enhanced Contrastive Learning for Multi-View Clustering with Hyper-Laplacian RegularizationZhibin Gu, Weili WangNeurIPS 2025 · 7 citations
- LightBSR: Towards Lightweight Blind Super-Resolution via Discriminative Implicit Degradation Representation LearningJiang Yuan, Ji Ma, Bo Wang, Guanzhou Ke et al.ICCV 2025 · 2 citations
- Imputation-free and Alignment-free: Incomplete Multi-view Clustering Driven by Consensus Semantic LearningYuzhuo Dai, Jiaqi Jin, Zhibin Dong, Siwei Wang et al.CVPR 2025
Builds on4
- Learning Robust Representations via Multi-View Information BottleneckMarco Federici, Anjan Dutta, Patrick Forré, Nate Kushman et al.ICLR 2020 · 330 citations
- Multi-VAE: Learning Disentangled View-common and View-peculiar Visual Representations for Multi-view ClusteringJie Xu, Yazhou Ren, Huayi Tang, Xiaorong Pu et al.ICCV 2021 · 158 citations
- Progressive Deep Multi-View Comprehensive Representation LearningCai Xu, Wei Zhao, Jinglong Zhao, Ziyu Guan et al.AAAI 2023 · 23 citations
- End-to-End Adversarial-Attention Network for Multi-Modal ClusteringRunwu Zhou, Yi-Dong ShenCVPR 2020
Related papers
- Generalized Information-theoretic Multi-view ClusteringWeitian Huang, Sirui Yang, Hongmin CaiNeurIPS 2023 · 17 citations
- Multi-View Representation Learning via Total Correlation ObjectiveHyeongJoo Hwang, Geon-Hyeong Kim, Seunghoon Hong, Kee-Eung KimNeurIPS 2021 · 63 citations
- Differentiable Information Bottleneck for Deterministic Multi-View ClusteringXiaoqiang Yan, Zhixiang Jin, Fengshou Han, Yangdong YeCVPR 2024 · 19 citations
- Deep Multiview Clustering by Contrasting Cluster AssignmentsJie Chen, Hua Mao, Wai Lok Woo, Xi PengICCV 2023 · 142 citations
- Incomplete Multi-View Multi-label Learning via Disentangled Representation and Label Semantic EmbeddingXu Yan, Jun Yin, Jie WenCVPR 2025
