TMFormer: Token Merging Transformer for Brain Tumor Segmentation with Missing Modalities
Zheyu Zhang, Gang Yang, Yueyi Zhang, Huanjing Yue, Aiping Liu, Yunwei Ou, Jian Gong, Xiaoyan Sun
Abstract
Numerous techniques excel in brain tumor segmentation using multi-modal magnetic resonance imaging (MRI) sequences, delivering exceptional results. However, the prevalent absence of modalities in clinical scenarios hampers performance. Current approaches frequently resort to zero maps as substitutes for missing modalities, inadvertently introducing feature bias and redundant computations. To address these issues, we present the Token Merging transFormer (TM-Former) for robust brain tumor segmentation with missing modalities. TMFormer tackles these challenges by extracting and merging accessible modalities into more compact token sequences. The architecture comprises two core components: the Uni-modal Token Merging Block (UMB) and the Multi-modal Token Merging Block (MMB). The UMB enhances individual modality representation by adaptively consolidating spatially redundant tokens within and outside tumor-related regions, thereby refining token sequences for augmented representational capacity. Meanwhile, the MMB mitigates multi-modal feature fusion bias, exclusively leveraging tokens from present modalities and merging them into a unified multi-modal representation to accommodate varying modality combinations. Extensive experimental results on the BraTS 2018 and 2020 datasets demonstrate the superiority and efficacy of TMFormer compared to state-of-the-art methods when dealing with missing modalities.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cf8b336b-522d-4381-afc5-c458ffbcad6eCited by top-tier papers5
- SimMLM: A Simple Framework for Multi-Modal Learning with Missing ModalitySijie Li, Chen Chen, Jungong HanICCV 2025 · 14 citations
- Virtual Nodes Guided Dynamic Graph Neural Network for Brain Tumor Segmentation with Missing ModalitiesSha Tao, Jiao Pan, Yu Guo, Chao YaoCVPR 2026
- Saliency-Driven Token Merging for Vision TransformersWeiying Xie, Xiaoyu Chen, Xin Zhang, Chenhe Hao et al.CVPR 2026
- Incomplete Multi-modal Brain Tumor Segmentation via Learnable Sorting State Space ModelZheyu Zhang, Yayuan Lu, Feipeng Ma, Yueyi Zhang et al.CVPR 2025
- Bayesian Decomposition and Semantic Completion for Few-shot Semantic SegmentationGuangchen Shi, Yirui Wu, Wei Zhu, Tao Wang et al.CVPR 2026
Builds on9
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- DynamicViT: Efficient Vision Transformers with Dynamic Token SparsificationYongming Rao, Wenliang Zhao, Benlin Liu, Jiwen Lu et al.NeurIPS 2021 · 1,343 citations
- CSWin Transformer: A General Vision Transformer Backbone with Cross-Shaped WindowsXiaoyi Dong, Jianmin Bao, Dongdong Chen, Weiming Zhang et al.CVPR 2022 · 1,207 citations
- RFNet: Region-aware Fusion Network for Incomplete Multi-modal Brain Tumor SegmentationYuhang Ding, Xin Yu, Yi YangICCV 2021 · 160 citations
Related papers
- Uni-Encoder Meets Multi-Encoders: Representation Before Fusion for Brain Tumor Segmentation with Missing ModalitiesPeibo Song, Xiaotian Xue, Jinshuo Zhang, Zihao Wang et al.CVPR 2026
- Sequential Information Bottleneck Fusion: Towards Robust and Generalizable Multi-Modal Brain Tumor SegmentationTIANYI LIU, Xi Yang, Wei Wang, Anh Nguyen et al.ICLR 2026
- M3AE: Multimodal Representation Learning for Brain Tumor Segmentation with Missing ModalitiesHong Liu, Dong Wei, Donghuan Lu, Jinghan Sun et al.AAAI 2023 · 101 citations
- Scratch Each Other's Back: Incomplete Multi-modal Brain Tumor Segmentation Via Category Aware Group Self-Support LearningYansheng Qiu, Delin Chen, Hongdou Yao, Yongchao Xu et al.ICCV 2023 · 30 citations
- DTMFormer: Dynamic Token Merging for Boosting Transformer-Based Medical Image SegmentationZhehao Wang, Xian Lin, Nannan Wu, Li Yu et al.AAAI 2024 · 14 citations
