M3amba: Memory Mamba is All You Need for Whole Slide Image Classification
Tingting Zheng, Kui Jiang, Yi Xiao, Sicheng Zhao, Hongxun Yao
Abstract
Multi-instance learning (MIL) has demonstrated impressive performance in whole slide image (WSI) analysis. However, existing approaches struggle with undesirable results and unbearable computational overhead due to the quadratic complexity of Transformers. Recently, Mamba has offered a feasible solution for modeling long-range dependencies with linear complexity. However, vanilla Mamba inherently suffers from contextual forgetting issues, making it ill-suited for capturing global dependencies across instances in large-scale WSIs. To address this, we propose a memory-driven Mamba network, dubbed M3amba, to fully explore the global latent relations among instances. Specifically, M3amba retains and iteratively updates historical information with a dynamic memory bank (DMB), thus overcoming the catastrophic forgetting defects of Mamba for long-term context representation. For better feature representation, M3amba involves an intra-group bidirectional Mamba (BiMamba) block to refine local interactions within groups. Meanwhile, we additionally perform cross-attention fusion to incorporate relevant historical information across groups, facilitating richer inter-group connections. The joint learning of inter- and intra-group representations with memory merits enables M3amba with a more powerful capability for achieving accurate and comprehensive WSI representation. Extensive experiments on four datasets demonstrate that M3amba outperforms the state-of-the-art by 6.2% and 7.0% in accuracy on the TCGA BRCA and TCGA Lung datasets while maintaining low computational costs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- GMMamba: Group Masking Mamba for Whole Slide Image ClassificationTingting Zheng, Hongxun Yao, Kui Jiang, Yi Xiao et al.ICCV 2025 · 5 citations
- Dynamic Fractal Mamba: A Neural Renormalization Group Flow for Scale-Invariant Sequence ModelingShenglei Fang, Xianfang Sun, You ZhouICML 2026
- Content-aware Information Compression and Selection for Whole Slide Image AnalysisTingting Zheng, Hongxun Yao, Sicheng Zhao, Yi XiaoAAAI 2026
- Cello: A Universal Cell-wise Feature Aggregation framework for Reliable Pathology Images AnalysisHengrui Lou, Weihan Li, Jiazhen Yang, Lingxiang Jia et al.ICML 2026
Builds on13
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- DTFD-MIL: Double-Tier Feature Distillation Multiple Instance Learning for Histopathology Whole Slide Image ClassificationHongrun Zhang, Yanda Meng, Yitian Zhao, Yihong Qiao et al.CVPR 2022 · 402 citations
- AdaMV-MoE: Adaptive Multi-Task Vision Mixture-of-ExpertsTianlong Chen, Xuxi Chen, Xianzhi Du, Abdullah Rashwan et al.ICCV 2023 · 119 citations
- Multi-modal Gated Mixture of Local-to-Global Experts for Dynamic Image FusionBing Cao, Yiming Sun, Pengfei Zhu, Qinghua HuICCV 2023 · 110 citations
Related papers
- Bridging Local Inductive Bias and Long-Range Dependencies With Pixel-Mamba for End-To-End Whole Slide Image AnalysisZhongwei Qiu, Hanqing Chao, Tiancheng Lin, Wanxing Chang et al.ICCV 2025 · 1 citation
- OODML: Whole Slide Image Classification Meets Online Pseudo-Supervision and Dynamic Mutual LearningTingting Zheng, Kui Jiang, Hongxun Yao, Yi Xiao et al.AAAI 2025 · 7 citations
- FBTA: Enabling Single-GPU End-to-End Gigapixel WSI Classification with Feature Bridging and Translation AlignmentJiuyang Dong, Jiahan Li, Junjun Jiang, Yongbing ZhangCVPR 2026
- SAM-MIL: A Spatial Contextual Aware Multiple Instance Learning Approach for Whole Slide Image ClassificationHeng Fang, Sheng Huang, Wenhao Tang, Luwen Huangfu et al.ACM MM 2024 · 13 citations
- Continual Multiple Instance Learning with Enhanced Localization for Histopathological Whole Slide Image AnalysisByung Hyun Lee, Wongi Jeong, Woojae Han, Kyoungbun Lee et al.ICCV 2025 · 3 citations
