Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
Jie Xu, Na Zhao, Gang Niu, Masashi Sugiyama, Xiaofeng Zhu
摘要
Recently, multi-view learning (MVL) has garnered significant attention due to its ability to fuse discriminative information from multiple views. However, real-world multi-view datasets are often heterogeneous and imperfect, which usually causes MVL methods designed for specific combinations of views to lack application potential and limits their effectiveness. To address this issue, we propose a novel robust MVL method (namely RML) with simultaneous representation fusion and alignment. Specifically, we introduce a simple yet effective multi-view transformer fusion network where we transform heterogeneous multi-view data into homogeneous word embeddings, and then integrate multiple views by the sample-level attention mechanism to obtain a fused representation. Furthermore, we propose a simulated perturbation based multi-view contrastive learning framework that dynamically generates the noise and unusable perturbations for simulating imperfect data conditions. The simulated noisy and unusable data obtain two distinct fused representations, and we utilize contrastive learning to align them for learning discriminative and robust representations. Our RML is self-supervised and can also be applied for downstream tasks as a regularization. In experiments, we employ it in multi-view unsupervised clustering, noise-label classification, and as a plug-and-play module for cross-modal hashing retrieval. Extensive comparison experiments and ablation studies validate RML's effectiveness. Code is available at https://github.com/SubmissionsIn/RML .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Mamba-Driven Multi-View Discriminative Clustering via Global-Local Cross-View Sequence ModelingYuanyang Zhang, Xinhang Wan, Chao Zhang, Jie Xu 等AAAI 2026
- Multi-Label Learning with Contrastive Cluster Self-Supervision for 3D Hierarchical Semantic SegmentationShuyu Cao, Chongshou Li, Jie Xu, Tianrui Li 等ICML 2026
- Bootstrapping Multi-view Learning for Test-time Noisy CorrespondenceChanghao He, Di Xue, Shuxian Li, Yanji Hao 等CVPR 2026
- Views Attention Fusion of Granular-ball Fuzzy Representations Split for Improved Multi-view ClusteringShuaiyu Liu, Song Wu, Jie Xu, Yazhou Ren 等AAAI 2026
它引用的顶会 Paper27
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty 等NeurIPS 2021 · 被引用 2,985 次
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 被引用 1,663 次
相关 Paper
- Partially View-Aligned Representation Learning With Noise-Robust Contrastive LossMouxing Yang, Yunfan Li, Zhenyu Huang, Zitao Liu 等CVPR 2021
- Learning Cross-Modal Retrieval With Noisy LabelsPeng Hu, Xi Peng, Hongyuan Zhu, Liangli Zhen 等CVPR 2021
- URRL-IMVC: Unified and Robust Representation Learning for Incomplete Multi-View ClusteringGe Teng, Ting Mao, Chen Shen, Xiang Tian 等KDD 2024 · 被引用 3 次
- RAC-DMVC: Reliability-Aware Contrastive Deep Multi-View Clustering Under Multi-Source NoiseShihao Dong, Yue Liu, Xiaotong Zhou, Yuhui Zheng 等AAAI 2026 · 被引用 1 次
- Dual-stage Contrastive Learning-enhanced Multi-view Variational ClusteringYanxi Liu, Yipin Hu, Fangxi Liu, Yanwei Yu 等ICML 2026
