Unsupervised Multi-Modal Image Registration via Geometry Preserving Image-to-Image Translation
Moab Arar, Yiftach Ginger, Dov Danon, Amit H. Bermano, Daniel Cohen-Or
Abstract
Many applications, such as autonomous driving, heavily rely on multi-modal data where spatial alignment between the modalities is required. Most multi-modal registration methods struggle computing the spatial correspondence between the images using prevalent cross-modality similarity measures. In this work, we bypass the difficulties of developing cross-modality similarity measures, by training an image-to-image translation network on the two input modalities. This learned translation allows training the registration network using simple and reliable mono-modality metrics. We perform multi-modal registration using two networks -a spatial transformation network and a translation network. We show that by encouraging our translation network to be geometry preserving, we manage to train an accurate spatial transformation network. Compared to state-of-the-art multi-modal methods our presented method is unsupervised, requiring no pairs of aligned modalities for training, and can be adapted to any pair of modalities. We evaluate our method quantitatively and qualitatively on commercial datasets, showing that it performs well on several modalities and achieves accurate alignment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 06bddb04-5372-44db-88b1-7f779db4e465Cited by top-tier papers18
- Breaking the Dilemma of Medical Image-to-image TranslationLingke Kong, Chenyu Lian, Detian Huang, Zhenjiang Li et al.NeurIPS 2021 · 234 citations
- RFNet: Unsupervised Network for Mutually Reinforcing Multi-modal Image Registration and FusionHan Xu, Jiayi Ma, Jiteng Yuan, Zhuliang Le et al.CVPR 2022 · 161 citations
- FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth EstimatorsHaiping Wang, Yuan Liu, Bing Wang, Yujing Sun et al.ICLR 2024 · 33 citations
- Promoting Single-Modal Optical Flow Network for Diverse Cross-Modal Flow EstimationShili Zhou, Weimin Tan, Bo YanAAAI 2022 · 33 citations
- Lesion-Inspired Denoising Network: Connecting Medical Image Denoising and Lesion DetectionKecheng Chen, Kun Long, Yazhou Ren, Jiayu Sun et al.ACM MM 2021 · 23 citations
Related papers
- The Spatially-Correlative Loss for Various Image Translation TasksChuanxia Zheng, Tat-Jen Cham, Jianfei CaiCVPR 2021
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li et al.CVPR 2026 · 4 citations
- DINO: A Conditional Energy-Based GAN for Domain TranslationKonstantinos Vougioukas, Stavros Petridis, Maja PanticICLR 2021 · 8 citations
- Attribute-Driven Spontaneous Motion in Unpaired Image TranslationRuizheng Wu, Xin Tao, Xiaodong Gu, Xiaoyong Shen et al.ICCV 2019 · 18 citations
- Modality-Agnostic Structural Image Representation Learning for Deformable Multi-Modality Medical Image RegistrationTony C. W. Mok, Zi Li, Yunhao Bai, Jianpeng Zhang et al.CVPR 2024 · 21 citations
