MetaViewer: Towards A Unified Multi-View Representation
Ren Wang, Haoliang Sun, Yuling Ma, Xiaoming Xi, Yilong Yin
Abstract
Existing multi-view representation learning methods typically follow a specific-to-uniform pipeline, extracting latent features from each view and then fusing or aligning them to obtain the unified object representation. However, the manually pre-specified fusion functions and aligning criteria could potentially degrade the quality of the derived representation. To overcome them, we propose a novel uniform-tospecific multi-view learning framework from a meta-learning perspective, where the unified representation no longer involves manual manipulation but is automatically derived from a meta-learner named MetaViewer. Specifically, we formulated the extraction and fusion of view-specific latent features as a nested optimization problem and solved it by using a bi-level optimization scheme. In this way, MetaViewer automatically fuses view-specific features into a unified one and learns the optimal fusion scheme by observing reconstruction processes from the unified to the specific over all views. Extensive experimental results in downstream classification and clustering tasks demonstrate the efficiency and effectiveness of the proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6fe9704e-2969-4c93-bfdb-559f45664e24Builds on14
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 1,663 citations
- MetaPruning: Meta Learning for Automatic Neural Network Channel PruningZechun Liu, Haoyuan Mu, Xiangyu Zhang, Zichao Guo et al.ICCV 2019 · 633 citations
- Learning Relationships between Text, Audio, and Video via Deep Canonical Correlation for Multimodal Language AnalysisZhongkai Sun, Prathusha Kameswara Sarma, William A. Sethares, Yingyu LiangAAAI 2020 · 419 citations
- Multi-level Feature Learning for Contrastive Multi-view ClusteringJie Xu, Huayi Tang, Yazhou Ren, Liang Peng et al.CVPR 2022 · 335 citations
Related papers
- Disentangling Multi-view Representations Beyond Inductive BiasGuanzhou Ke, Yang Yu, Guoqing Chao, Xiaoli Wang et al.ACM MM 2023 · 15 citations
- SeqMvRL: A Sequential Fusion Framework for Multi-view Representation LearningRen Wang, Haoliang Sun, Yuxiu Lin, Chuanhui Zuo et al.CVPR 2025
- Deep Embedded Complementary and Interactive Information for Multi-View ClassificationJinglin Xu, Wenbin Li, Xinwang Liu, Dingwen Zhang et al.AAAI 2020 · 65 citations
- Adaptive Evolutionary Fusion for Multi-View ClusteringYunxiao Zhao, Liang Bai, Xian YangAAAI 2026
- Learn to Merge: Meta-Learning for Adaptive Multi-Task Model MergingJun Chen, Qin Zhang, Weizhi Zhang, Xiao Luo et al.ICML 2026
