Cross-Modal Subspace Clustering via Deep Canonical Correlation Analysis
Quanxue Gao, Huanhuan Lian, Qianqian Wang, Gan Sun
Abstract
For cross-modal subspace clustering, the key point is how to exploit the correlation information between cross-modal data. However, most hierarchical and structural correlation information among cross-modal data cannot be well exploited due to its high-dimensional non-linear property. To tackle this problem, in this paper, we propose an unsupervised framework named Cross-Modal Subspace Clustering via Deep Canonical Correlation Analysis (CMSC-DCCA), which incorporates the correlation constraint with a self-expressive layer to make full use of information among the inter-modal data and the intra-modal data. More specifically, the proposed model consists of three components: 1) deep canonical correlation analysis (Deep CCA) model; 2) self-expressive layer; 3) Deep CCA decoders. The Deep CCA model consists of convolutional encoders and correlation constraint. Convolutional encoders are used to obtain the latent representations of cross-modal data, while adding the correlation constraint for the latent representations can make full use of the information of the inter-modal data. Furthermore, self-expressive layer works on latent representations and constrain it perform self-expression properties, which makes the shared coefficient matrix could capture the hierarchical intra-modal correlations of each modality. Then Deep CCA decoders reconstruct data to ensure that the encoded features can preserve the structure of the original data. Experimental results on several real-world datasets demonstrate the proposed method outperforms the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 81dd485e-c141-41bf-949a-e3f1731ebc0eCited by top-tier papers8
- Deep Mutual Information Maximin for Cross-Modal ClusteringYiqiao Mao, Xiaoqiang Yan, Qiang Guo, Yangdong YeAAAI 2021 · 58 citations
- Partial Multi-View Clustering via Self-Supervised NetworkWei Feng, Guoshuai Sheng, Qianqian Wang, Quanxue Gao et al.AAAI 2024 · 15 citations
- Learning Canonical F-Correlation Projection for Compact Multiview RepresentationYun-Hao Yuan, Jin Li, Yun Li, Jipeng Qiang et al.CVPR 2022 · 9 citations
- Multi-Level Cross-Modal Alignment for Image ClusteringLiping Qiu, Qin Zhang, Xiaojun Chen, Shaotian CaiAAAI 2024 · 8 citations
- Contrastive Multi-view Subspace Clustering via Tensor Transformers AutoencoderQianqian Wang, Zihao Zhang, Wei Feng, Zhiqiang Tao et al.AAAI 2025 · 5 citations
Related papers
- Deep Self-Supervised t-SNE for Multi-modal Subspace ClusteringQianqian Wang, Wei Xia, Zhiqiang Tao, Quanxue Gao et al.ACM MM 2021 · 19 citations
- Multi-Scale Fusion Subspace Clustering Using Similarity ConstraintZhiyuan Dang, Cheng Deng, Xu Yang, Heng HuangCVPR 2020
- Multi-view Self-Expressive Subspace Clustering NetworkJinrong Cui, Yuting Li, Yulu Fu, Jie WenACM MM 2023 · 11 citations
- Exploring a Principled Framework for Deep Subspace ClusteringXianghan Meng, Zhiyuan Huang, Wei He, Xianbiao Qi et al.ICLR 2025
- A Critique of Self-Expressive Deep Subspace ClusteringBenjamin David Haeffele, Chong You, René VidalICLR 2021 · 35 citations
