What Makes a Representation Good for Single-Cell Perturbation Prediction?
Wenkang Jiang, Yuhang Liu, Yichao Cai, Erdun Gao, Jiayi Dong, Ehsan Abbasnejad, Lina Yao, Javen Qinfeng Shi
摘要
Single-cell perturbation modeling is fundamental for understanding and predicting cellular responses to genetic perturbations. However, existing approaches, from causal representation learning to foundation models, often struggle with an overlooked challenge: gene expression is dominated by perturbation-invariant information, while perturbation-specific signals are intrinsically sparse. As a result, learned representations either entangle invariant and perturbation-specific information, leading to spurious and non-generalizable predictors, or suppress perturbation-specific signals altogether, rendering them ineffective for prediction. To address this, we propose PerturbedVAE, a general framework designed to resolve this signal imbalance. The framework explicitly separates perturbation-specific information from dominant invariant structure and recovers causal representations to effectively utilize such information for prediction. We further provide an identifiability analysis that characterizes the conditions under which sparse perturbation effects can be reliably recovered, thereby clarifying how the framework can be concretely specified under such conditions. Empirically, PerturbedVAE achieves state-of-the-art performance on a widely used benchmark across multiple evaluation settings, yielding significant gains on out-of-distribution combinatorial predictions and uncovering interpretable perturbation-response programs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- On Mutual Information Maximization for Representation LearningMichael Tschannen, Josip Djolonga, Paul K. Rubenstein, Sylvain Gelly 等ICLR 2020 · 被引用 559 次
- Self-Supervised Learning with Data Augmentations Provably Isolates Content from StyleJulius von Kügelgen, Yash Sharma, Luigi Gresele, Wieland Brendel 等NeurIPS 2021 · 被引用 421 次
- Interventional Causal Representation LearningKartik Ahuja, Divyat Mahajan, Yixin Wang, Yoshua BengioICML 2023 · 被引用 143 次
- CITRIS: Causal Identifiability from Temporal Intervened SequencesPhillip Lippe, Sara Magliacane, Sindy Löwe, Yuki M. Asano 等ICML 2022 · 被引用 136 次
相关 Paper
- Cradle-VAE: Enhancing Single-Cell Gene Perturbation Modeling with Counterfactual Reasoning-based Artifact DisentanglementSeungheun Baek, Soyon Park, Yan Ting Chok, Junhyun Lee 等AAAI 2025 · 被引用 5 次
- Beyond Independent Genes: Learning Module-Inductive Representations for Single-Cell Gene Perturbation PredictionJiafa Ruan, Ruijie Quan, Liyang Xu, Zongxin Yang 等ICML 2026 · 被引用 3 次
- Modelling Cellular Perturbations with the Sparse Additive Mechanism Shift Variational AutoencoderMichael Bereket, Theofanis KaraletsosNeurIPS 2023 · 被引用 59 次
- scDFM: Distributional Flow Matching Model for Robust Single-Cell Perturbation PredictionChenglei Yu, Chuanrui Wang, Bangyan Liao, Tailin WuICLR 2026 · 被引用 15 次
- Learning Cross-Domain Representations for Transferable Drug Perturbations on Single-Cell Transcriptional ResponsesHui Liu, Shikai JinAAAI 2025 · 被引用 1 次
