Understanding Latent Correlation-Based Multiview Learning and Self-Supervision: An Identifiability Perspective
Qi Lyu, Xiao Fu, Weiran Wang, Songtao Lu
摘要
Multiple views of data, both naturally acquired (e.g., image and audio) and artificially produced (e.g., via adding different noise to data samples), have proven useful in enhancing representation learning. Natural views are often handled by multiview analysis tools, e.g., (deep) canonical correlation analysis [(D)CCA], while the artificial ones are frequently used in self-supervised learning (SSL) paradigms, e.g., BYOL and Barlow Twins. Both types of approaches often involve learning neural feature extractors such that the embeddings of data exhibit high cross-view correlations. Although intuitive, the effectiveness of correlation-based neural embedding is mostly empirically validated. This work aims to understand latent correlation maximization-based deep multiview learning from a latent component identification viewpoint. An intuitive generative model of multiview data is adopted, where the views are different nonlinear mixtures of shared and private components. Since the shared components are view/distortion-invariant, representing the data using such components is believed to reveal the identity of the samples effectively and robustly. Under this model, latent correlation maximization is shown to guarantee the extraction of the shared components across views (up to certain ambiguities). In addition, it is further shown that the private information in each view can be provably disentangled from the shared using proper regularization design. A finite sample analysis, which has been rare in nonlinear mixture identifiability study, is also presented. The theoretical results and newly designed regularization are tested on a series of tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Multi-View Causal Representation Learning with Partial ObservabilityDingling Yao, Danru Xu, Sébastien Lachapelle, Sara Magliacane 等ICLR 2024 · 被引用 70 次
- Causal Component AnalysisWendong Liang, Armin Kekic, Julius von Kügelgen, Simon Buchholz 等NeurIPS 2023 · 被引用 65 次
- Disentangled Multiplex Graph Representation LearningYujie Mo, Yajie Lei, Jialie Shen, Xiaoshuang Shi 等ICML 2023 · 被引用 35 次
- Identification of Nonlinear Latent Hierarchical ModelsLingjing Kong, Biwei Huang, Feng Xie, Eric P. Xing 等NeurIPS 2023 · 被引用 33 次
- Marrying Causal Representation Learning with Dynamical Systems for ScienceDingling Yao, Caroline Muller, Francesco LocatelloNeurIPS 2024 · 被引用 29 次
它引用的顶会 Paper9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- Self-Supervised Learning with Data Augmentations Provably Isolates Content from StyleJulius von Kügelgen, Yash Sharma, Luigi Gresele, Wieland Brendel 等NeurIPS 2021 · 被引用 421 次
相关 Paper
- Deep Probabilistic Canonical Correlation AnalysisMahdi Karami, Dale SchuurmansAAAI 2021 · 被引用 8 次
- Shared Generative Latent Representation Learning for Multi-View ClusteringMing Yin, Weitian Huang, Junbin GaoAAAI 2020 · 被引用 78 次
- COPER: Correlation-based Permutations for Multi-View ClusteringRan Eisenberg, Jonathan Svirsky, Ofir LindenbaumICLR 2025
- Unconstrained Stochastic CCA: Unifying Multiview and Self-Supervised LearningJames Chapman, Lennie Wells, Ana Lawry AguilaICLR 2024 · 被引用 2 次
- Preventing Model Collapse in Deep Canonical Correlation Analysis by Noise RegularizationJunlin He, Jinxiao Du, Susu Xu, Wei MaNeurIPS 2024 · 被引用 5 次
