TopoSlide: Topologically-Informed Histopathology Whole Slide Image Representation Learning
Shahira Abousamra, Asmita Sood, Sylvia Plevritis
Abstract
Histopathology whole slide images (WSIs) are gigapixel images that present significant challenges in generating effective representations that capture both local histological features and their global spatial organization. Current pathology foundation models focus primarily on local patch-level features while neglecting the complex spatial relationships that pathologists rely on for diagnosis and prognosis. We introduce TopoSlide, a novel self-supervised representation learning framework that leverages persistent homology from topological data analysis to capture the global spatial organization of tissue architecture in WSIs.
Our method decomposes slides into histologically meaningful clusters using patch-level embeddings, then characterizes their spatial arrangement through topological descriptors. We train a vision transformer to predict cluster topology from slide-level embeddings using a conditional multitask objective that integrates local patch features with their topological attributes. Evaluated across lung adenocarcinoma and breast cancer cohorts, TopoSlide achieves superior performance improving histologic pattern retrieval by up to 15% in majority voting macro F1 score, and competitive survival and gene mutation predictions, while training on only hundreds of slides compared to hundreds of thousands for foundation models. Our results demonstrate that topology-aware learning provides a powerful inductive bias for pathology representation learning, enabling both improved performance and novel topology-based conditional retrieval capabilities for clinical applications. Our code and models are publicly available. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on6
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals et al.ICML 2021 · 1,399 citations
- Image BERT Pre-training with Online TokenizerJinghao Zhou, Chen Wei, Huiyu Wang, Wei Shen et al.ICLR 2022 · 287 citations
- Localization in the Crowd with Topological ConstraintsShahira Abousamra, Minh Hoai, Dimitris Samaras, Chao ChenAAAI 2021 · 160 citations
- Topologically Faithful Image Segmentation via Induced Matching of Persistence BarcodesNico Stucki, Johannes C. Paetzold, Suprosanna Shit, Bjoern H. Menze et al.ICML 2023 · 72 citations
- TopoDiffusionNet: A Topology-aware Diffusion ModelSaumya Gupta, Dimitris Samaras, Chao ChenICLR 2025
Related papers
- Rotation-Agnostic Image Representation Learning for Digital PathologySaghir Alfasly, Abubakr Shafique, Peyman Nejat, Jibran A. Khan et al.CVPR 2024
- Unsupervised Foundation Model-Agnostic Slide-Level Representation LearningTim Lenz, Peter Neidlinger, Marta Ligero, Georg Wölflein et al.CVPR 2025
- TopoImages: Incorporating Local Topology Encoding into Deep Learning Models for Medical Image ClassificationPengfei Gu, Hongxiao Wang, Yejia Zhang, Huimin Li et al.ACM MM 2025 · 4 citations
- Hierarchical Discriminative Learning Improves Visual Representations of Biomedical MicroscopyCheng Jiang, Xinhai Hou, Akhil Kondepudi, Asadur Chowdury et al.CVPR 2023
- Multimodal Optimal Transport-based Co-Attention Transformer with Global Structure Consistency for Survival PredictionYingxue Xu, Hao ChenICCV 2023 · 132 citations
