Promptable Representation Distribution Learning and Data Augmentation for Gigapixel Histopathology WSI Analysis
Kunming Tang, Zhiguo Jiang, Jun Shi, Wei Wang, Haibo Wu, Yushan Zheng
Abstract
Gigapixel image analysis, particularly for whole slide images (WSIs), often relies on multiple instance learning (MIL). Under the paradigm of MIL, patch image representations are extracted and then fixed during the training of the MIL classifiers for efficiency consideration. However, the invariance of representations makes it difficult to perform data augmentation for WSI-level model training, which significantly limits the performance of the downstream WSI analysis. The current data augmentation methods for gigapixel images either introduce additional computational costs or result in a loss of semantic information, which is hard to meet the requirements for efficiency and stability needed for WSI model training. In this paper, we propose a Promptable Representation Distribution Learning framework (PRDL) for both patch-level representation learning and WSI-level data augmentation. Meanwhile, we explore the use of prompts to guide data augmentation in feature space, which achieves promptable data augmentation for training robust WSI-level models. The experimental results have demonstrated that the proposed method stably outperforms state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 97c32fc9-8cae-4fdd-b485-a43800c2689bCited by top-tier papers2
- Federated Distillation for Whole Slide Image via Gaussian-Mixture Feature Alignment and Curriculum IntegrationLuru Jing, Cong Cong, Yanyuan Chen, Yongzhi CaoICML 2026 · 1 citation
- Contrastive Cross-Bag Augmentation for Multiple Instance Learning-based Whole Slide Image ClassificationBo Zhang, Xinan Xu, Shuo Yan, Yu Bai et al.CVPR 2026
Builds on17
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
Related papers
- Controllable Latent Space Augmentation for Digital PathologySofiène Boutaj, Marin Scalbert, Pierre Marza, Florent Couzinie-Devy et al.ICCV 2025 · 2 citations
- LNPL-MIL: Learning from Noisy Pseudo Labels for Promoting Multiple Instance Learning in Whole Slide ImageZhuchen Shao, Yifeng Wang, Yang Chen, Hao Bian et al.ICCV 2023 · 27 citations
- Boosting Multiple Instance Learning Models for Whole Slide Image Classification: A Model-Agnostic Framework Based on Counterfactual InferenceWeiping Lin, Zhenfeng Zhuang, Lequan Yu, Liansheng WangAAAI 2024 · 19 citations
- Task-Specific Fine-Tuning via Variational Information Bottleneck for Weakly-Supervised Pathology Whole Slide Image ClassificationHonglin Li, Chenglu Zhu, Yunlong Zhang, Yuxuan Sun et al.CVPR 2023
- Continual Multiple Instance Learning with Enhanced Localization for Histopathological Whole Slide Image AnalysisByung Hyun Lee, Wongi Jeong, Woojae Han, Kyoungbun Lee et al.ICCV 2025 · 3 citations
