Efficient Folded Attention for Medical Image Reconstruction and Segmentation
Hang Zhang, Jinwei Zhang, Rongguang Wang, Qihao Zhang, Pascal Spincemaille, Thanh D. Nguyen, Yi Wang
Abstract
Recently, 3D medical image reconstruction (MIR) and segmentation (MIS) based on deep neural networks have been developed with promising results, and attention mechanism has been further designed to capture global contextual information for performance enhancement. However, the large size of 3D volume images poses a great computational challenge to traditional attention methods. In this paper, we propose a folded attention (FA) approach to improve the computational efficiency of traditional attention methods on 3D medical images. The main idea is that we apply tensor folding and unfolding operations with four permutations to build four small sub-affinity matrices to approximate the original affinity matrix. Through four consecutive sub-attention modules of FA, each element in the feature tensor can aggregate spatial-channel information from all other elements. Compared to traditional attention methods, with moderate improvement of accuracy, FA can substantially reduce the computational complexity and GPU memory consumption. We demonstrate the superiority of our method on two challenging tasks for 3D MIR and MIS, which are quantitative susceptibility mapping and multiple sclerosis lesion segmentation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3b134b1a-4834-4af0-b40b-285ff76d0cbdBuilds on1
Related papers
- EMCAD: Efficient Multi-Scale Convolutional Attention Decoding for Medical Image SegmentationMd Mostafijur Rahman, Mustafa Munir, Radu MarculescuCVPR 2024 · 352 citations
- MAGIC: Rethinking Dynamic Convolution Design for Medical Image SegmentationShijie Li, Yunbin Tu, Qingyuan Xiang, Zheng LiACM MM 2024 · 8 citations
- MR Image Super-Resolution With Squeeze and Excitation Reasoning Attention NetworkYulun Zhang, Kai Li, Kunpeng Li, Yun FuCVPR 2021
- EffiDec3D: An Optimized Decoder for High-Performance and Efficient 3D Medical Image SegmentationMd Mostafijur Rahman, Radu MarculescuCVPR 2025
- CARL: A Framework for Equivariant Image RegistrationThomas Hastings Greer, Lin Tian, François-Xavier Vialard, Roland Kwitt et al.CVPR 2025
