Lune

NeurIPS2024Top-tier venue

Disentangling Interpretable Factors with Supervised Independent Subspace Principal Component Analysis

Jiayu Su, David A. Knowles, Raúl Rabadán

2024Year
5Citations

Abstract

The success of machine learning models relies heavily on effectively representing high-dimensional data. However, ensuring data representations capture human-understandable concepts remains difficult, often requiring the incorporation of prior knowledge and decomposition of data into multiple subspaces. Traditional linear methods fall short in modeling more than one space, while more expressive deep learning approaches lack interpretability. Here, we introduce Supervised Independent Subspace Principal Component Analysis (sisPCA\texttt{sisPCA}), a PCA extension designed for multi-subspace learning. Leveraging the Hilbert-Schmidt Independence Criterion (HSIC), sisPCA\texttt{sisPCA} incorporates supervision and simultaneously ensures subspace disentanglement. We demonstrate sisPCA\texttt{sisPCA}'s connections with autoencoders and regularized linear regression and showcase its ability to identify and separate hidden data structures through extensive applications, including breast cancer diagnosis from image features, learning aging-associated DNA methylation changes, and single-cell analysis of malaria infection. Our results reveal distinct functional pathways associated with malaria colonization, underscoring the essentiality of explainable representation in high-dimensional data analysis.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext f9d91ae5-4982-4382-b64d-3b22b213f35f

Builds on2

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines