HEIST: A Graph Foundation Model for Spatial Transcriptomics and Proteomics Data
Hiren Madhu, João Felipe Rocha, Tinglin Huang, Siddharth Viswanath, Smita Krishnaswamy, Rex Ying
摘要
Single-cell transcriptomics and proteomics have become a great source for data-driven insights into biology, enabling the use of advanced deep learning methods to understand cellular heterogeneity and gene expression at the single-cell level. With the advent of spatial-omics data, we have the promise of characterizing cells within their tissue context as it provides both spatial coordinates and intra-cellular transcriptional or protein counts. Beyond transcriptomics, proteomics offers a complementary view by directly measuring proteins, which are the primary effectors of cellular function and key therapeutic targets. However, existing models either ignore the spatial information or the complex genetic and proteomic programs within cells. Thus they cannot infer how cell internal regulation adapts to microenvironmental cues. Furthermore, these models often utilize fixed gene vocabularies, hindering their generalizability to datasets with different genes than pretraining. In this paper, we introduce HEIST, a hierarchical graph transformer foundation model for spatial transcriptomics and proteomics. HEIST models tissues as hierarchical graphs. The higher level graph is a spatial cell graph, and each cell in turn, is represented by its lower level gene co-expression network graph. Rather than using a fixed gene vocabulary, HEIST computes gene embeddings from its co-expression network and cellular context. HEIST achieves this by performing both intra-level and cross-level message passing to utilize the hierarchy in its embeddings and can thus generalize to novel datatypes including spatial proteomics without retraining. HEIST is pretrained on 22.3M cells from 124 tissues across 15 organs using spatially-aware contrastive and masked autoencoding objectives. Unsupervised analysis of HEIST embeddings reveals spatially informed subpopulations missed by prior models. Downstream evaluations demonstrate generalizability to proteomics data and state-of-the-art performance in clinical outcome prediction, cell type annotation, and gene imputation across multiple technologies.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- LLM4Cell: Taxonomy and Evaluation of LLM and Agentic Models for Single-Cell BiologySajib Acharjee Dip, Adrika Zafor, Bikash Kumar Paul, Uddip Acharjee Shuvo 等ACL 2026
- Robust Integrative Analysis of Multi-omics Datasets via Nuclear-norm MaximizationMeng-zhu Wang, Yu Zhang, Hongxing ZhangAAAI 2026
它引用的顶会 Paper6
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng 等NeurIPS 2021 · 被引用 1,632 次
- From Canonical Correlation Analysis to Self-supervised Graph Neural NetworksHengrui Zhang, Qitian Wu, Junchi Yan, David Wipf 等NeurIPS 2021 · 被引用 319 次
- CellPLM: Pre-training of Cell Language Model Beyond Single CellsHongzhi Wen, Wenzhuo Tang, Xinnan Dai, Jiayuan Ding 等ICLR 2024 · 被引用 76 次
- Efficient Learning of Mesh-Based Physical Simulation with Bi-Stride Multi-Scale Graph Neural NetworkYadi Cao, Menglei Chai, Minchen Li, Chenfanfu JiangICML 2023 · 被引用 47 次
- An iterative clustering algorithm for the Contextual Stochastic Block Model with optimality guaranteesGuillaume Braun, Hemant Tyagi, Christophe BiernackiICML 2022 · 被引用 16 次
相关 Paper
- SToFM: a Multi-scale Foundation Model for Spatial TranscriptomicsSuyuan Zhao, Yizhen Luo, Ganbo Yang, Yan Zhong 等ICML 2025
- HiST: A Hierarchical Sparse Transformer for Cross-Modal Spatial Transcriptomics ModelingWeiyi Wu, Xinwen Xu, Xingjian Diao, Siting Li 等ICML 2026
- Cross-Slice Knowledge Transfer via Masked Multi-Modal Heterogeneous Graph Contrastive Learning for Spatial Gene Expression InferenceZhiceng Shi, Changmiao Wang, Jun Wan, Wenwen MinCVPR 2026 · 被引用 1 次
- ST-LLM: Spatial Transcriptomics Embedding with Large Language ModelsZhetao Xu, Xiaohua Wan, Le Li, Shuang Feng 等AAAI 2026
- Adapting a Pre-trained Single-Cell Foundation Model to Spatial Gene Expression Generation from Histology ImagesDonghai Fang, Yongheng Li, Zhen WANG, Yuansong Zeng 等CVPR 2026 · 被引用 2 次
