M3AE: Multimodal Representation Learning for Brain Tumor Segmentation with Missing Modalities
Hong Liu, Dong Wei, Donghuan Lu, Jinghan Sun, Liansheng Wang, Yefeng Zheng
Abstract
Multimodal magnetic resonance imaging (MRI) provides complementary information for sub-region analysis of brain tumors. Plenty of methods have been proposed for automatic brain tumor segmentation using four common MRI modalities and achieved remarkable performance. In practice, however, it is common to have one or more modalities missing due to image corruption, artifacts, acquisition protocols, allergy to contrast agents, or simply cost. In this work, we propose a novel two-stage framework for brain tumor segmentation with missing modalities. In the first stage, a multimodal masked autoencoder (M 3 AE) is proposed, where both random modalities (i.e., modality dropout) and random patches of the remaining modalities are masked for a reconstruction task, for self-supervised learning of robust multimodal representations against missing modalities. To this end, we name our framework M 3 AE. Meanwhile, we employ model inversion to optimize a representative full-modal image at marginal extra cost, which will be used to substitute for the missing modalities and boost performance during inference. Then in the second stage, a memory-efficient self distillation is proposed to distill knowledge between heterogenous missing-modal situations while fine-tuning the model for supervised segmentation. Our M 3 AE belongs to the 'catchall' genre where a single model can be applied to all possible subsets of modalities, thus is economic for both training and deployment. Extensive experiments on BraTS 2018 and 2020 datasets demonstrate its superior performance to existing state-of-the-art methods with missing modalities, as well as the efficacy of its components. Our code is available at: https://github.com/ccarliu/m3ae .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8b9eff1d-7e18-4398-8f8a-ac912df0f2d4Cited by top-tier papers23
- TMFormer: Token Merging Transformer for Brain Tumor Segmentation with Missing ModalitiesZheyu Zhang, Gang Yang, Yueyi Zhang, Huanjing Yue et al.AAAI 2024 · 29 citations
- PASSION: Towards Effective Incomplete Multi-Modal Medical Image Segmentation with Imbalanced Missing RatesJunjie Shi, Caozhi Shang, Zhaobin Sun, Li Yu et al.ACM MM 2024 · 20 citations
- SimMLM: A Simple Framework for Multi-Modal Learning with Missing ModalitySijie Li, Chen Chen, Jungong HanICCV 2025 · 14 citations
- Learning Disentangled Representation for Multi-Modal Time-Series Sensing SignalsRuichu Cai, Zhifan Jiang, Kaitao Zheng, Zijian Li et al.WWW 2025 · 8 citations
- Towards a Universal 3D Medical Multi-Modality Generalization via Learning Personalized Invariant RepresentationZhaorui Tan, Xi Yang, Tan Pan, Tianyi Liu et al.ICCV 2025 · 5 citations
Builds on5
- 3D Self-Supervised Methods for Medical ImagingAiham Taleb, Winfried Loetzsch, Noel Danz, Julius Severin et al.NeurIPS 2020 · 281 citations
- RFNet: Region-aware Fusion Network for Incomplete Multi-modal Brain Tumor SegmentationYuhang Ding, Xin Yu, Yi YangICCV 2021 · 160 citations
- Geometric Multimodal Contrastive Representation LearningPetra Poklukar, Miguel Vasco, Hang Yin, Francisco S. Melo et al.ICML 2022 · 68 citations
- Refine Myself by Teaching Myself: Feature Refinement via Self-Knowledge DistillationMingi Ji, Seungjae Shin, Seunghyun Hwang, Gibeom Park et al.CVPR 2021
- Masked Autoencoders Are Scalable Vision LearnersKaiming He, Xinlei Chen, Saining Xie, Yanghao Li et al.CVPR 2022
Related papers
- Uni-Encoder Meets Multi-Encoders: Representation Before Fusion for Brain Tumor Segmentation with Missing ModalitiesPeibo Song, Xiaotian Xue, Jinshuo Zhang, Zihao Wang et al.CVPR 2026
- Virtual Nodes Guided Dynamic Graph Neural Network for Brain Tumor Segmentation with Missing ModalitiesSha Tao, Jiao Pan, Yu Guo, Chao YaoCVPR 2026
- Tackling Dual-stage Missing Modalities in Brain Tumor Segmentation via Robust Modality Reconstruction and Prompt-guided Modality AdaptationYunpeng Zhao, Cheng Chen, Qing You Pang, Yibing Fu et al.AAAI 2026
- Enhancing Modality-Agnostic Representations via Meta-learning for Brain Tumor SegmentationAishik Konwer, Xiaoling Hu, Joseph Bae, Xuan Xu et al.ICCV 2023 · 23 citations
- Scratch Each Other's Back: Incomplete Multi-modal Brain Tumor Segmentation Via Category Aware Group Self-Support LearningYansheng Qiu, Delin Chen, Hongdou Yao, Yongchao Xu et al.ICCV 2023 · 30 citations
