Affine Medical Image Registration with Coarse-to-Fine Vision Transformer
Tony C. W. Mok, Albert C. S. Chung
摘要
Affine registration is indispensable in a comprehensive medical image registration pipeline. However, only a few studies focus on fast and robust affine registration algorithms. Most of these studies utilize convolutional neural networks (CNNs) to learn joint affine and non-parametric registration, while the standalone performance of the affine subnetwork is less explored. Moreover, existing CNN-based affine registration approaches focus either on the local mis-alignment or the global orientation and position of the input to predict the affine transformation matrix, which are sensitive to spatial initialization and exhibit limited generalizability apart from the training dataset. In this paper, we present a fast and robust learning-based algorithm, Coarse-to-Fine Vision Transformer (C2FViT), for 3D affine medical image registration. Our method naturally leverages the global connectivity and locality of the convolutional vision transformer and the multi-resolution strategy to learn the global affine registration. We evaluate our method on 3D brain atlas registration and template-matching normalization. Comprehensive results demonstrate that our method is superior to the existing CNNs-based affine registration methods in terms of registration accuracy, robustness and generalizability while preserving the runtime advantage of the learning-based methods. The source code is available at https://github.com/cwmok/C2FViT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Deep Learning in Medical Image Registration: Magic or Mirage?Rohit Jena, Deeksha Sethi, Pratik Chaudhari, James C. GeeNeurIPS 2024 · 被引用 36 次
- Modality-Agnostic Structural Image Representation Learning for Deformable Multi-Modality Medical Image RegistrationTony C. W. Mok, Zi Li, Yunhao Bai, Jianpeng Zhang 等CVPR 2024 · 被引用 21 次
- Planar Affine Rectification from Local Change of Scale and OrientationYuval Nissan, Marc Pollefeys, Daniel BarathICCV 2025
- H-ViT: A Hierarchical Vision Transformer for Deformable Image RegistrationMorteza Ghahremani, Mohammad Khateri, Bailiang Jian, Benedikt Wiestler 等CVPR 2024
- SACB-Net: Spatial-awareness Convolutions for Medical Image RegistrationXinxing Cheng, Tianyang Zhang, Wenqi Lu, Qingjie Meng 等CVPR 2025
它引用的顶会 Paper8
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- CvT: Introducing Convolutions to Vision TransformersHaiping Wu, Bin Xiao, Noel Codella, Mengchen Liu 等ICCV 2021 · 被引用 2,397 次
- CoAtNet: Marrying Convolution and Attention for All Data SizesZihang Dai, Hanxiao Liu, Quoc V. Le, Mingxing TanNeurIPS 2021 · 被引用 1,747 次
- Twins: Revisiting the Design of Spatial Attention in Vision TransformersXiangxiang Chu, Zhi Tian, Yuqing Wang, Bo Zhang 等NeurIPS 2021 · 被引用 1,388 次
相关 Paper
- Correlation-aware Coarse-to-fine MLPs for Deformable Medical Image RegistrationMingyuan Meng, Dagan Feng, Lei Bi, Jinman KimCVPR 2024 · 被引用 47 次
- ShiftMorph: A Fast and Robust Convolutional Neural Network for 3D Deformable Medical Image RegistrationLijian Yang, Weisheng Li, Yucheng Shu, Jian-Xun Mi 等ACM MM 2024 · 被引用 4 次
- Aladdin: Joint Atlas Building and Diffeomorphic Registration Learning with Pairwise AlignmentZhipeng Ding, Marc NiethammerCVPR 2022 · 被引用 21 次
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li 等CVPR 2026 · 被引用 4 次
- Fast Symmetric Diffeomorphic Image Registration with Convolutional Neural NetworksTony C. W. Mok, Albert C. S. ChungCVPR 2020
