Advancing Multiple Instance Learning with Continual Learning for Whole Slide Imaging
Xianrui Li, Yufei Cui, Jun Li, Antoni B. Chan
Abstract
Advances in medical imaging and deep learning have propelled progress in whole slide image (WSI) analysis, with multiple instance learning (MIL) showing promise for efficient and accurate diagnostics. However, conventional MIL models often lack adaptability to evolving datasets, as they rely on static training that cannot incorporate new information without extensive retraining. Applying continual learning (CL) to MIL models is a possible solution, but often sees limited improvements. In this paper, we analyze CL in the context of attention MIL models and find that the model forgetting is mainly concentrated in the attention layers of the MIL model. Using the results of this analysis we propose two components for improving CL on MIL: Attention Knowledge Distillation (AKD) and the Pseudo-Bag Memory Pool (PMP). AKD mitigates catastrophic forgetting by focusing on retaining attention layer knowledge between learning sessions, while PMP reduces the memory footprint by selectively storing only the most informative patches, or "pseudo-bags" from WSIs. Experimental evaluations demonstrate that our method significantly improves both accuracy and memory efficiency on diverse WSI datasets, outperforming current state-of-the-art CL methods. This work provides a foundation for CL in large-scale, weakly annotated clinical datasets, paving the way for more adaptable and resilient diagnostic models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ab08142c-fddb-4d9f-8a3a-3acad18ddb65Builds on5
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Nyströmformer: A Nyström-based Algorithm for Approximating Self-AttentionYunyang Xiong, Zhanpeng Zeng, Rudrasis Chakraborty, Mingxing Tan et al.AAAI 2021 · 675 citations
- DTFD-MIL: Double-Tier Feature Distillation Multiple Instance Learning for Histopathology Whole Slide Image ClassificationHongrun Zhang, Yanda Meng, Yitian Zhao, Yihong Qiao et al.CVPR 2022 · 402 citations
- ConSlide: Asynchronous Hierarchical Interaction Transformer with Breakup-Reorganize Rehearsal for Continual Whole Slide Image AnalysisYanyan Huang, Weiqin Zhao, Shujun Wang, Yu Fu et al.ICCV 2023 · 32 citations
- Bayes-MIL: A New Probabilistic Perspective on Attention-based Multiple Instance Learning for Whole Slide ImagesYufei Cui, Ziquan Liu, Xiangyu Liu, Xue Liu et al.ICLR 2023
Related papers
- Continual Multiple Instance Learning with Enhanced Localization for Histopathological Whole Slide Image AnalysisByung Hyun Lee, Wongi Jeong, Woojae Han, Kyoungbun Lee et al.ICCV 2025 · 3 citations
- Bi-directional Weakly Supervised Knowledge Distillation for Whole Slide Image ClassificationLinhao Qu, Xiaoyuan Luo, Manning Wang, Zhijian SongNeurIPS 2022 · 88 citations
- OODML: Whole Slide Image Classification Meets Online Pseudo-Supervision and Dynamic Mutual LearningTingting Zheng, Kui Jiang, Hongxun Yao, Yi Xiao et al.AAAI 2025 · 7 citations
- ASMIL: Attention-Stabilized Multiple Instance Learning for Whole-Slide ImagingLinfeng Ye, Shayan Mohajer Hamidi, Zhixiang Chi, Guang Li et al.ICLR 2026 · 9 citations
- Dual-Curriculum Contrastive Multi-Instance Learning for Cancer Prognosis Analysis with Whole Slide ImagesChao Tu, Yu Zhang, Zhenyuan NingNeurIPS 2022 · 24 citations
