Rolling-Unet: Revitalizing MLP's Ability to Efficiently Extract Long-Distance Dependencies for Medical Image Segmentation
Yutong Liu, Haijiang Zhu, Mengting Liu, Huaiyuan Yu, Zihan Chen, Jie Gao
Abstract
Medical image segmentation methods based on deep learning network are mainly divided into CNN and Transformer. However, CNN struggles to capture long-distance dependencies, while Transformer suffers from high computational complexity and poor local feature learning. To efficiently extract and fuse local features and long-range dependencies, this paper proposes Rolling-Unet, which is a CNN model combined with MLP. Specifically, we propose the core R-MLP module, which is responsible for learning the long-distance dependency in a single direction of the whole image. By controlling and combining R-MLP modules in different directions, OR-MLP and DOR-MLP modules are formed to capture long-distance dependencies in multiple directions. Further, Lo2 block is proposed to encode both local context information and long-distance dependencies without excessive computational burden. Lo2 block has the same parameter size and computational complexity as a 3×3 convolution. The experimental results on four public datasets show that Rolling-Unet achieves superior performance compared to the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a2bef627-816c-4a12-919b-173c88e3d8b5Cited by top-tier papers12
- U-KAN Makes Strong Backbone for Medical Image Segmentation and GenerationChenxin Li, Xinyu Liu, Wuyang Li, Cheng Wang et al.AAAI 2025 · 452 citations
- AIF-SFDA: Autonomous Information Filter Driven Source-Free Domain Adaptation for Medical Image SegmentationHaojin Li, Heng Li, Jianyu Chen, Rihan Zhong et al.AAAI 2025 · 5 citations
- Aligning and Prompting Anything for Zero-Shot Generalized Anomaly DetectionJitao Ma, Weiying Xie, Hangyu Ye, Daixun Li et al.AAAI 2025 · 3 citations
- Neighbor Does Matter: Density-Aware Contrastive Learning for Medical Semi-supervised SegmentationFeilong Tang, Zhongxing Xu, Ming Hu, Wenxue Li et al.AAAI 2025 · 3 citations
- LoMix: Learnable Weighted Multi-Scale Logits Mixing for Medical Image SegmentationMd Mostafijur Rahman, Radu MarculescuNeurIPS 2025 · 2 citations
Builds on7
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer et al.NeurIPS 2021 · 3,862 citations
- UCTransNet: Rethinking the Skip Connections in U-Net from a Channel-Wise Perspective with TransformerHaonan Wang, Peng Cao, Jiaqi Wang, Osmar R. ZaïaneAAAI 2022 · 1,144 citations
- AS-MLP: An Axial Shifted MLP Architecture for VisionDongze Lian, Zehao Yu, Xing Sun, Shenghua GaoICLR 2022 · 217 citations
Related papers
- Correlation-aware Coarse-to-fine MLPs for Deformable Medical Image RegistrationMingyuan Meng, Dagan Feng, Lei Bi, Jinman KimCVPR 2024 · 47 citations
- Segmenting Medical MRI via Recurrent Decoding CellYing Wen, Kai Xie, Lianghua HeAAAI 2020 · 12 citations
- nnWNet: Rethinking the Use of Transformers in Biomedical Image Segmentation and Calling for a Unified Evaluation BenchmarkYanfeng Zhou, Lingrui Li, Le Lu, Minfeng XuCVPR 2025
- CDDFuse: Correlation-Driven Dual-Branch Feature Decomposition for Multi-Modality Image FusionZixiang Zhao, Haowen Bai, Jiangshe Zhang, Yulun Zhang et al.CVPR 2023
- Learning Contextual Transformer Network for Image InpaintingYe Deng, Siqi Hui, Sanping Zhou, Deyu Meng et al.ACM MM 2021 · 29 citations
