MAGIC: Rethinking Dynamic Convolution Design for Medical Image Segmentation
Shijie Li, Yunbin Tu, Qingyuan Xiang, Zheng Li
Abstract
Recently, dynamic convolution shows performance boost for the CNN-related networks in medical image segmentation. The core idea is to replace static convolutional kernel with a linear combination of multiple convolutional kernels, conditioned on input-dependent attention function. However, the existing dynamic convolution design suffers from two limitations: i) The convolutional kernels are weighted by enforcing a single-dimensional attention function upon the input maps, overlooking the synergy in multi-dimensional information. This results in sub-optimal computations of convolution kernels. ii) The linear kernel aggregation is inefficient, restricting the model's capacity to learn more intricate patterns. In this paper, we rethink the dynamic convolution design to address these limitations and propose multi-dimensional aggregation dynamic convolution (MAGIC). Specifically, our MAGIC introduce a dimensional-reciprocal fusion module to capture correlations among input maps across the spatial, channel, and global dimensions simultaneously for computing convolutional kernels. Furthermore, we design kernel recalculation module, which enhances the efficiency of aggregation through learning the interaction between kernels. As a drop-in replacement for regular convolution, our MAGIC can be flexibly integrated into prevalent pure CNN or hybrid CNN-Transformer backbones. The extensive experiments on four benchmarks demonstrate that our MAGIC outperforms regular convolution and existing dynamic convolution. Code is available at: https://github.com/Segment82/MAGIC
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 6379aef4-e354-4cbf-874d-d7b077ee971eCited by top-tier papers3
- BSAFusion: A Bidirectional Stepwise Feature Alignment Network for Unaligned Medical Image FusionHuafeng Li, Dayong Su, Qing Cai, Yafei ZhangAAAI 2025 · 43 citations
- Unsupervised Photometric-Consistent Depth Estimation from Endoscopic Monocular VideoShijie Li, Weijun Lin, Qingyuan Xiang, Yunbin Tu et al.AAAI 2025 · 4 citations
- CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ SegmentationXinlei Yu, Changmiao Wang, Hui Jin, Ahmed Elazab et al.ACM MM 2025 · 3 citations
Related papers
- Revisiting Dynamic Convolution via Matrix DecompositionYunsheng Li, Yinpeng Chen, Xiyang Dai, Mengchen Liu et al.ICLR 2021 · 82 citations
- Omni-Dimensional Dynamic ConvolutionChao Li, Aojun Zhou, Anbang YaoICLR 2022 · 408 citations
- Dynamic Convolution: Attention Over Convolution KernelsYinpeng Chen, Xiyang Dai, Mengchen Liu, Dongdong Chen et al.CVPR 2020
- Mutual-Guided Dynamic Network for Image FusionYuanshen Guan, Ruikang Xu, Mingde Yao, Lizhi Wang et al.ACM MM 2023 · 24 citations
- DTMFormer: Dynamic Token Merging for Boosting Transformer-Based Medical Image SegmentationZhehao Wang, Xian Lin, Nannan Wu, Li Yu et al.AAAI 2024 · 14 citations
