Adapting a Pre-trained Single-Cell Foundation Model to Spatial Gene Expression Generation from Histology Images
Donghai Fang, Yongheng Li, Zhen WANG, Yuansong Zeng, Wenwen Min
Abstract
Spatial transcriptomics (ST) enables spot-level in situ expression profiling, but its high cost and limited throughput motivate predicting expression directly from H&E-stained histology. Recent advances explore using score- or flow-based generative models to estimate the conditional distribution of gene expression from histology, offering a flexible alternative to deterministic regression approaches. However, most existing generative approaches omit explicit modeling of gene–gene dependencies, undermining biological coherence. Single-cell foundation models (sc-FMs), pre-trained across diverse cell populations, capture these critical gene relationships that histology alone cannot reveal. Yet, applying expression-only sc-FMs to histology-conditioned expression modeling is nontrivial due to the absence of a visual pathway, a mismatch between their pre-training and conditional ST objectives, and the scarcity of mixed-cell ST supervision. To address these challenges, we propose HINGE ( HI stology-co N ditioned GE neration), which retrofits a pre-trained sc-FM into a conditional expression generator while mostly preserving its learned gene relationships. We achieve this by introducing SoftAdaLN , a lightweight, identity-initialized modulation that injects layer-wise visual context into the backbone, coupled with an expression-space masked diffusion objective and a warm-start curriculum to ensure objective alignment and training stability. Evaluated on three ST datasets, HINGE outperforms state-of-the-art baselines on mean Pearson correlation and yields more accurate spatial marker expression patterns and higher pairwise co-expression consistency, establishing a practical route to adapt pre-trained sc-FMs for histology-conditioned spatial expression generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f30fd978-4833-4da9-a922-a1e7a988eea9Cited by top-tier papers1
Ask how each one uses itBuilds on10
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Spatially Resolved Gene Expression Prediction from Histology Images via Bi-modal Contrastive LearningRonald Xie, Kuan Pang, Sai Chung, Catia Perciani et al.NeurIPS 2023 · 125 citations
- Flow Matching for Generative ModelingYaron Lipman, Ricky T. Q. Chen, Heli Ben-Hamu, Maximilian Nickel et al.ICLR 2023 · 87 citations
- M2OST: Many-to-one Regression for Predicting Spatial Transcriptomics from Digital Pathology ImagesHongyi Wang, Xiuju Du, Jing Liu, Shuyi Ouyang et al.AAAI 2025 · 12 citations
- Scalable Generation of Spatial Transcriptomics from Histology Images via Whole-Slide Flow MatchingTinglin Huang, Tianyu Liu, Mehrtash Babadi, Wengong Jin et al.ICML 2025
Related papers
- Diffusion Generative Modeling for Spatially Resolved Gene Expression Inference from Histology ImagesSichen Zhu, Yuchen Zhu, Molei Tao, Peng QiuICLR 2025
- HEXST: Hexagonal Shifted-Window Transformer for Spatial Transcriptomics Gene Expression PredictionKeunho Byeon, Jin Tae KwakICML 2026 · 1 citation
- HiFusion: Hierarchical Intra-Spot Alignment and Regional Context Fusion for Spatial Gene Expression Prediction from HistopathologyZiqiao Weng, Yaoyu Fang, Jiahe Qian, Xinkun Wang et al.AAAI 2026
- FLAG: Foundation model representation with Latent diffusion Alignment via Graph for spatial gene expression predictionQi Si, Penglei Wang, Yushuai Wu, Yifeng Jiao et al.ICML 2026 · 1 citation
- SToFM: a Multi-scale Foundation Model for Spatial TranscriptomicsSuyuan Zhao, Yizhen Luo, Ganbo Yang, Yan Zhong et al.ICML 2025
