Predicting Spatial Transcriptomics from Histology Images via High-Order Multi-Cell Interaction Modeling
Youhan Sun, Jiahua Rao, Kangrui Du, Jiancong Xie, Yuedong Yang
Abstract
Spatial transcriptomics (ST) links gene expression to tissue architecture and enables predicting spatial expression from H&E-stained whole-slide images (WSIs). However, existing spot-or slide-level predictors focus on single-spot features or pairwise relations, failing to capture high-order, manyto-many cross-cell interactions. As a result, they miss synergistic and antagonistic effects among multiple neighboring cells. Here, we introduce MCToGene, a scalable and accurate framework that explicitly models multi-cell interactions via many-body attention with hierarchical coupling to predict spatial gene expression. MCToGene employs a many-body attention module to encode high-order, manyto-many cross-cell dependencies, enabling context-aware microenvironment modeling. To mitigate the combinatorial burden of many-body modeling, we design a hierarchical interaction module that couples pairwise and many-body representations for feature aggregation and prediction, preserving many-body expressiveness while controlling computation and memory. On HEST-1k and STImage-1K4M, MCToGene surpasses state-of-the-art baselines with 7.85% relative improvement. Ablations confirm that explicit highorder, many-to-many modeling drives these gains, and visualizations demonstrate that multi-cell interactions are essential for biologically coherent spatial predictions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b535213d-ebb1-4144-8972-dbba5c4f2f2bBuilds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Spatially Resolved Gene Expression Prediction from Histology Images via Bi-modal Contrastive LearningRonald Xie, Kuan Pang, Sai Chung, Catia Perciani et al.NeurIPS 2023 · 125 citations
- MedM2G: Unifying Medical Multi-Modal Generation via Cross-Guided Diffusion with Visual InvariantChenlu Zhan, Yu Lin, Gaoang Wang, Hongwei Wang et al.CVPR 2024 · 20 citations
- Diffusion Models for Multi-Task Generative ModelingChangyou Chen, Han Ding, Bunyamin Sisman, Yi Xu et al.ICLR 2024 · 11 citations
Related papers
- HiFusion: Hierarchical Intra-Spot Alignment and Regional Context Fusion for Spatial Gene Expression Prediction from HistopathologyZiqiao Weng, Yaoyu Fang, Jiahe Qian, Xinkun Wang et al.AAAI 2026
- Scalable Generation of Spatial Transcriptomics from Histology Images via Whole-Slide Flow MatchingTinglin Huang, Tianyu Liu, Mehrtash Babadi, Wengong Jin et al.ICML 2025
- M2OST: Many-to-one Regression for Predicting Spatial Transcriptomics from Digital Pathology ImagesHongyi Wang, Xiuju Du, Jing Liu, Shuyi Ouyang et al.AAAI 2025 · 12 citations
- SPATIA: Multimodal Generation and Prediction of Spatial Cell PhenotypesZhenglun Kong, Mufan Qiu, John Boesen, xiang lin et al.ICML 2026 · 1 citation
- FEAST: Fully Connected Expressive Attention for Spatial TranscriptomicsTaejin Jeong, Joohyeok Kim, Jinyeong Kim, Chanyoung Kim et al.CVPR 2026 · 1 citation
