Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
Jie Xu, Na Zhao, Gang Niu, Masashi Sugiyama, Xiaofeng Zhu
Abstract
Recently, multi-view learning (MVL) has garnered significant attention due to its ability to fuse discriminative information from multiple views. However, real-world multi-view datasets are often heterogeneous and imperfect, which usually causes MVL methods designed for specific combinations of views to lack application potential and limits their effectiveness. To address this issue, we propose a novel robust MVL method (namely RML) with simultaneous representation fusion and alignment. Specifically, we introduce a simple yet effective multi-view transformer fusion network where we transform heterogeneous multi-view data into homogeneous word embeddings, and then integrate multiple views by the sample-level attention mechanism to obtain a fused representation. Furthermore, we propose a simulated perturbation based multi-view contrastive learning framework that dynamically generates the noise and unusable perturbations for simulating imperfect data conditions. The simulated noisy and unusable data obtain two distinct fused representations, and we utilize contrastive learning to align them for learning discriminative and robust representations. Our RML is self-supervised and can also be applied for downstream tasks as a regularization. In experiments, we employ it in multi-view unsupervised clustering, noise-label classification, and as a plug-and-play module for cross-modal hashing retrieval. Extensive comparison experiments and ablation studies validate RML's effectiveness. Code is available at https://github.com/SubmissionsIn/RML .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f02e980d-f8fa-4d9b-8092-c3cc4b4d8c87Cited by top-tier papers4
- Mamba-Driven Multi-View Discriminative Clustering via Global-Local Cross-View Sequence ModelingYuanyang Zhang, Xinhang Wan, Chao Zhang, Jie Xu et al.AAAI 2026
- Multi-Label Learning with Contrastive Cluster Self-Supervision for 3D Hierarchical Semantic SegmentationShuyu Cao, Chongshou Li, Jie Xu, Tianrui Li et al.ICML 2026
- Bootstrapping Multi-view Learning for Test-time Noisy CorrespondenceChanghao He, Di Xue, Shuxian Li, Yanji Hao et al.CVPR 2026
- Views Attention Fusion of Granular-ball Fuzzy Representations Split for Improved Multi-view ClusteringShuaiyu Liu, Song Wu, Jie Xu, Yazhou Ren et al.AAAI 2026
Builds on27
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty et al.NeurIPS 2021 · 2,985 citations
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 1,663 citations
Related papers
- Partially View-Aligned Representation Learning With Noise-Robust Contrastive LossMouxing Yang, Yunfan Li, Zhenyu Huang, Zitao Liu et al.CVPR 2021
- Learning Cross-Modal Retrieval With Noisy LabelsPeng Hu, Xi Peng, Hongyuan Zhu, Liangli Zhen et al.CVPR 2021
- URRL-IMVC: Unified and Robust Representation Learning for Incomplete Multi-View ClusteringGe Teng, Ting Mao, Chen Shen, Xiang Tian et al.KDD 2024 · 3 citations
- RAC-DMVC: Reliability-Aware Contrastive Deep Multi-View Clustering Under Multi-Source NoiseShihao Dong, Yue Liu, Xiaotong Zhou, Yuhui Zheng et al.AAAI 2026 · 1 citation
- Dual-stage Contrastive Learning-enhanced Multi-view Variational ClusteringYanxi Liu, Yipin Hu, Fangxi Liu, Yanwei Yu et al.ICML 2026
