Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics
Jiayuan Chen, Ruoqi Liu, Zishan Gu, Ping Zhang
Abstract
Microscopy-based phenotypic profiling is scalable for drug discovery but lacks the mechanistic depth of transcriptomics, which remains costly and scarce. Existing multimodal approaches either use images to support other modalities or naively align representations by sample identity, ignoring cell-type and dose variations in weakly paired datalimiting generalization to unseen interventions. In this paper, we introduce an intervention-aware distillation framework that leverages perturbational transcriptomics to guide image representation learning. A transcriptome-conditioned teacher integrates gene expression and intervention metadata to produce soft distributions over a chemistry-aware codebook organized by drug similarity. The teacher employs a fine-tuned single-cell foundation model to encode cell-type context and disentangle dose effects. An image-only student learns to predict these distributions from microscopy alone, distilling mechanistic knowledge while operating independently at test time. This design emphasizes intervention semantics rather than identity alignment and explicitly handles dose and cell-type mismatches. We provide theoretical guarantees showing that transcriptomic guidance tightens the risk bound for image-based prediction. On Cell Painting and RxRx datasets paired with L1000, our method significantly improves one-shot transfer to unseen interventions and drugtarget gene discovery compared to self-supervised and alignment baselines 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- How Molecules Impact Cells: Unlocking Contrastive PhenoMolecular RetrievalPhilip Fradkin, Puria Azadi Moghadam, Karush Suri, Frederik Wenkel et al.NeurIPS 2024 · 14 citations
- Removing Biases from Molecular Representations via Information MaximizationChenyu Wang, Sharut Gupta, Caroline Uhler, Tommi S. JaakkolaICLR 2024 · 11 citations
- Learning Molecular Representation in a CellGang Liu, Srijit Seal, John Arevalo, Zhenwen Liang et al.ICLR 2025
Related papers
- Integrating Biological Knowledge for Robust Microscopy Image Profiling on De Novo Cell LinesJiayuan Chen, Thai-Hoang Pham, Yuanlong Wang, Ping ZhangICCV 2025
- A Cross Modal Knowledge Distillation & Data Augmentation Recipe for Improving Transcriptomics Representations through Morphological FeaturesIhab Bendidi, Yassir El Mesbahi, Alisandra Kaye Denton, Karush Suri et al.ICML 2025
- Learning Cross-Domain Representations for Transferable Drug Perturbations on Single-Cell Transcriptional ResponsesHui Liu, Shikai JinAAAI 2025 · 1 citation
- scDFM: Distributional Flow Matching Model for Robust Single-Cell Perturbation PredictionChenglei Yu, Chuanrui Wang, Bangyan Liao, Tailin WuICLR 2026 · 15 citations
- PETRI: Learning Unified Cell Embeddings from Unpaired Modalities via Early-Fusion Joint ReconstructionRyan W Conrad, Ethan Weinberger, Saradha Venkatachalapathy, Yuwen Chen et al.ICLR 2026
