Modelling Cellular Perturbations with the Sparse Additive Mechanism Shift Variational Autoencoder
Michael Bereket, Theofanis Karaletsos
摘要
Generative models of observations under interventions have been a vibrant topic of interest across machine learning and the sciences in recent years. For example, in drug discovery, there is a need to model the effects of diverse interventions on cells in order to characterize unknown biological mechanisms of action. We propose the Sparse Additive Mechanism Shift Variational Autoencoder, SAMS-VAE, to combine compositionality, disentanglement, and interpretability for perturbation models. SAMS-VAE models the latent state of a perturbed sample as the sum of a local latent variable capturing sample-specific variation and sparse global variables of latent intervention effects. Crucially, SAMS-VAE sparsifies these global latent variables for individual perturbations to identify disentangled, perturbation-specific latent subspaces that are flexibly composable. We evaluate SAMS-VAE both quantitatively and qualitatively on a range of tasks using two popular single cell sequencing datasets. In order to measure perturbation-specific model-properties, we also introduce a framework for evaluation of perturbation models based on average treatment effects with links to posterior predictive checks. SAMS-VAE outperforms comparable models in terms of generalization across in-distribution and out-of-distribution tasks, including a combinatorial reasoning task under resource paucity, and yields interpretable latent structures which correlate strongly to known biological mechanisms. Our results suggest SAMS-VAE is an interesting addition to the modeling toolkit for machine learning-driven scientific discovery. * Research supporting this publication conducted while authors were employed at insitro 37th Conference on Neural Information Processing Systems (NeurIPS 2023).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Learning Identifiable Factorized Causal Representations of Cellular ResponsesHaiyi Mao, Romain Lopez, Kai Liu, Jan-Christian Huetter 等NeurIPS 2024 · 被引用 10 次
- Scalable Single-Cell Gene Expression Generation with Latent Diffusion ModelsGiovanni Palla, Sudarshan Babu, Payam Dibaeinia, James Pearce 等ICML 2026 · 被引用 7 次
- Cradle-VAE: Enhancing Single-Cell Gene Perturbation Modeling with Counterfactual Reasoning-based Artifact DisentanglementSeungheun Baek, Soyon Park, Yan Ting Chok, Junhyun Lee 等AAAI 2025 · 被引用 5 次
- Doloris: Dual Conditional Diffusion Implicit Bridges with Sparsity Masking Strategy for Unpaired Single-Cell Perturbation EstimationChangxi Chi, Jun Xia, Yufei Huang, Zhuoli Ouyang 等ICLR 2026 · 被引用 4 次
- Statistical and structural identifiability in representation learningWalter Nelson, Marco Fumero, Theofanis Karaletsos, Francesco LocatelloICLR 2026 · 被引用 4 次
相关 Paper
- What Makes a Representation Good for Single-Cell Perturbation Prediction?Wenkang Jiang, Yuhang Liu, Yichao Cai, Erdun Gao 等ICML 2026 · 被引用 2 次
- Interpretable Causal Representation Learning for Biological Data in the Pathway SpaceJesus de la Fuente Cedeño, Robert Lehmann, Carlos Ruiz-Arenas, Jan Voges 等ICLR 2025
- Identifiability Guarantees for Causal Disentanglement from Soft InterventionsJiaqi Zhang, Kristjan H. Greenewald, Chandler Squires, Akash Srivastava 等NeurIPS 2023 · 被引用 120 次
- Generative Intervention Models for Causal Perturbation ModelingNora Schneider, Lars Lorch, Niki Kilbertus, Bernhard Schölkopf 等ICML 2025
- ProtSAE: Disentangling and Interpreting Protein Language Models via Semantically-Guided Sparse AutoencodersXiangyu Liu, Haodi Lei, Yi Liu, Yang Liu 等AAAI 2026 · 被引用 2 次
