Affine Medical Image Registration with Coarse-to-Fine Vision Transformer
Tony C. W. Mok, Albert C. S. Chung
Abstract
Affine registration is indispensable in a comprehensive medical image registration pipeline. However, only a few studies focus on fast and robust affine registration algorithms. Most of these studies utilize convolutional neural networks (CNNs) to learn joint affine and non-parametric registration, while the standalone performance of the affine subnetwork is less explored. Moreover, existing CNN-based affine registration approaches focus either on the local mis-alignment or the global orientation and position of the input to predict the affine transformation matrix, which are sensitive to spatial initialization and exhibit limited generalizability apart from the training dataset. In this paper, we present a fast and robust learning-based algorithm, Coarse-to-Fine Vision Transformer (C2FViT), for 3D affine medical image registration. Our method naturally leverages the global connectivity and locality of the convolutional vision transformer and the multi-resolution strategy to learn the global affine registration. We evaluate our method on 3D brain atlas registration and template-matching normalization. Comprehensive results demonstrate that our method is superior to the existing CNNs-based affine registration methods in terms of registration accuracy, robustness and generalizability while preserving the runtime advantage of the learning-based methods. The source code is available at https://github.com/cwmok/C2FViT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e043c87d-6094-4b22-a13f-fd137fc264ebCited by top-tier papers6
- Deep Learning in Medical Image Registration: Magic or Mirage?Rohit Jena, Deeksha Sethi, Pratik Chaudhari, James C. GeeNeurIPS 2024 · 36 citations
- Modality-Agnostic Structural Image Representation Learning for Deformable Multi-Modality Medical Image RegistrationTony C. W. Mok, Zi Li, Yunhao Bai, Jianpeng Zhang et al.CVPR 2024 · 21 citations
- Planar Affine Rectification from Local Change of Scale and OrientationYuval Nissan, Marc Pollefeys, Daniel BarathICCV 2025
- H-ViT: A Hierarchical Vision Transformer for Deformable Image RegistrationMorteza Ghahremani, Mohammad Khateri, Bailiang Jian, Benedikt Wiestler et al.CVPR 2024
- SACB-Net: Spatial-awareness Convolutions for Medical Image RegistrationXinxing Cheng, Tianyang Zhang, Wenqi Lu, Qingjie Meng et al.CVPR 2025
Builds on8
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- CvT: Introducing Convolutions to Vision TransformersHaiping Wu, Bin Xiao, Noel Codella, Mengchen Liu et al.ICCV 2021 · 2,397 citations
- CoAtNet: Marrying Convolution and Attention for All Data SizesZihang Dai, Hanxiao Liu, Quoc V. Le, Mingxing TanNeurIPS 2021 · 1,747 citations
- Twins: Revisiting the Design of Spatial Attention in Vision TransformersXiangxiang Chu, Zhi Tian, Yuqing Wang, Bo Zhang et al.NeurIPS 2021 · 1,388 citations
Related papers
- Correlation-aware Coarse-to-fine MLPs for Deformable Medical Image RegistrationMingyuan Meng, Dagan Feng, Lei Bi, Jinman KimCVPR 2024 · 47 citations
- ShiftMorph: A Fast and Robust Convolutional Neural Network for 3D Deformable Medical Image RegistrationLijian Yang, Weisheng Li, Yucheng Shu, Jian-Xun Mi et al.ACM MM 2024 · 4 citations
- Aladdin: Joint Atlas Building and Diffeomorphic Registration Learning with Pairwise AlignmentZhipeng Ding, Marc NiethammerCVPR 2022 · 21 citations
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li et al.CVPR 2026 · 4 citations
- Fast Symmetric Diffeomorphic Image Registration with Convolutional Neural NetworksTony C. W. Mok, Albert C. S. ChungCVPR 2020
