Unconstrained Stochastic CCA: Unifying Multiview and Self-Supervised Learning
James Chapman, Lennie Wells, Ana Lawry Aguila
Abstract
The Canonical Correlation Analysis (CCA) family of methods is foundational in multiview learning. Regularised linear CCA methods can be seen to generalise Partial Least Squares (PLS) and be unified with a Generalized Eigenvalue Problem (GEP) framework. However, classical algorithms for these linear methods are computationally infeasible for large-scale data. Extensions to Deep CCA show great promise, but current training procedures are slow and complicated. First we propose a novel unconstrained objective that characterizes the top subspace of GEPs. Our core contribution is a family of fast algorithms for stochastic PLS, stochastic CCA, and Deep CCA, simply obtained by applying stochastic gradient descent (SGD) to the corresponding CCA objectives. Our algorithms show far faster convergence and recover higher correlations than the previous state-of-theart on all standard CCA and Deep CCA benchmarks. These improvements allow us to perform a first-of-its-kind PLS analysis of an extremely large biomedical dataset from the UK Biobank, with over 33,000 individuals and 500,000 features. Finally, we apply our algorithms to match the performance of 'CCA-family' Self-Supervised Learning (SSL) methods on CIFAR-10 and CIFAR-100 with minimal hyper-parameter tuning, and also present theory to clarify the links between these methods and classical CCA, laying the groundwork for future insights. * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d8455f97-73ea-4da5-861d-0fd5e9c099a2Cited by top-tier papers2
- Contrastive Predictive Coding Done Right for Mutual Information EstimationJongha Ryu, Pavan Yeddanapudi, Xiangxiang Xu, Gregory W. WornellICLR 2026 · 1 citation
- Revisiting Orbital Minimization Method for Neural Operator DecompositionJongha Ryu, Samuel Zhou, Gregory W. WornellNeurIPS 2025 · 1 citation
Builds on7
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 1,226 citations
- Understanding Dimensional Collapse in Contrastive Self-supervised LearningLi Jing, Pascal Vincent, Yann LeCun, Yuandong TianICLR 2022 · 467 citations
- Contrastive and Non-Contrastive Self-Supervised Learning Recover Global and Local Spectral Embedding MethodsRandall Balestriero, Yann LeCunNeurIPS 2022 · 189 citations
- EigenGame: PCA as a Nash EquilibriumIan Gemp, Brian McWilliams, Claire Vernade, Thore GraepelICLR 2021 · 56 citations
Related papers
- Deep Probabilistic Canonical Correlation AnalysisMahdi Karami, Dale SchuurmansAAAI 2021 · 8 citations
- Understanding Latent Correlation-Based Multiview Learning and Self-Supervision: An Identifiability PerspectiveQi Lyu, Xiao Fu, Weiran Wang, Songtao LuICLR 2022 · 37 citations
- L0-Sparse Canonical Correlation AnalysisOfir Lindenbaum, Moshe Salhov, Amir Averbuch, Yuval KlugerICLR 2022 · 20 citations
- Preventing Model Collapse in Deep Canonical Correlation Analysis by Noise RegularizationJunlin He, Jinxiao Du, Susu Xu, Wei MaNeurIPS 2024 · 5 citations
- Best of Both Worlds: Multimodal Contrastive Learning with Tabular and Imaging DataPaul Hager, Martin J. Menten, Daniel RueckertCVPR 2023
