Learning in Feature Spaces via Coupled Covariances: Asymmetric Kernel SVD and Nyström method
Qinghua Tao, Francesco Tonin, Alex Lambert, Yingyi Chen, Panagiotis Patrinos, Johan A. K. Suykens
Abstract
In contrast with Mercer kernel-based approaches as used e.g., in Kernel Principal Component Analysis (KPCA), it was previously shown that Singular Value Decomposition (SVD) inherently relates to asymmetric kernels and Asymmetric Kernel Singular Value Decomposition (KSVD) has been proposed. However, the existing formulation to KSVD cannot work with infinite-dimensional feature mappings, the variational objective can be unbounded, and needs further numerical evaluation and exploration towards machine learning. In this work, i) we introduce a new asymmetric learning paradigm based on coupled covariance eigenproblem (CCE) through covariance operators, allowing infinite-dimensional feature maps. The solution to CCE is ultimately obtained from the SVD of the induced asymmetric kernel matrix, providing links to KSVD. ii) Starting from the integral equations corresponding to a pair of coupled adjoint eigenfunctions, we formalize the asymmetric Nyström method through a finite sample approximation to speed up training. iii) We provide the first empirical evaluations verifying the practical utility and benefits of KSVD and compare with methods resorting to symmetrization or linear SVD across multiple tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6b33cccc-6ab4-4a98-a3ec-9bd84431c488Builds on5
- Nyströmformer: A Nyström-based Algorithm for Approximating Self-AttentionYunyang Xiong, Zhanpeng Zeng, Rudrasis Chakraborty, Mingxing Tan et al.AAAI 2021 · 675 citations
- Kernel Methods Through the Roof: Handling Billions of Points EfficientlyGiacomo Meanti, Luigi Carratino, Lorenzo Rosasco, Alessandro RudiNeurIPS 2020 · 138 citations
- Directed Graph Auto-EncodersGeorgios Kollias, Vasileios Kalantzis, Tsuyoshi Idé, Aurélie C. Lozano et al.AAAI 2022 · 49 citations
- Primal-Attention: Self-attention through Asymmetric Kernel SVD in Primal RepresentationYingyi Chen, Qinghua Tao, Francesco Tonin, Johan A. K. SuykensNeurIPS 2023 · 42 citations
- Efficient and Effective Optimal Transport-Based BiclusteringChakib Fettal, Lazhar Labiod, Mohamed NadifNeurIPS 2022 · 9 citations
Related papers
- Extending Kernel PCA through Dualization: Sparsity, Robustness and Fast AlgorithmsFrancesco Tonin, Alex Lambert, Panagiotis Patrinos, Johan A. K. SuykensICML 2023 · 3 citations
- NeuralEF: Deconstructing Kernels by Deep Neural NetworksZhijie Deng, Jiaxin Shi, Jun ZhuICML 2022 · 29 citations
- Nyström-Accelerated Primal LS-SVMs: Breaking the Complexity Bottleneck for Scalable ODEs LearningWeikuo Wang, Yue Liao, Huan LuoNeurIPS 2025
- Contrastive Learning Can Find An Optimal Basis For Approximately View-Invariant FunctionsDaniel D. Johnson, Ayoub El Hanchi, Chris J. MaddisonICLR 2023 · 1 citation
- Learning to Decompose Asymmetric Channel Kernels for Generalized Eigenwave MultiplexingZhibin Zou, Iresha Amarasekara, Aveek DuttaINFOCOM 2024 · 9 citations
