CoMIR: Contrastive Multimodal Image Representation for Registration
Nicolas Pielawski, Elisabeth Wetzer, Johan Öfverstedt, Jiahao Lu, Carolina Wählby, Joakim Lindblad, Natasa Sladoje
Abstract
We propose contrastive coding to learn shared, dense image representations, referred to as CoMIRs (Contrastive Multimodal Image Representations). CoMIRs enable the registration of multimodal images where existing registration methods often fail due to a lack of sufficiently similar image structures. CoMIRs reduce the multimodal registration problem to a monomodal one, in which general intensitybased, as well as feature-based, registration algorithms can be applied. The method involves training one neural network per modality on aligned images, using a contrastive loss based on noise-contrastive estimation (InfoNCE). Unlike other contrastive coding methods, used for, e.g., classification, our approach generates image-like representations that contain the information shared between modalities. We introduce a novel, hyperparameter-free modification to InfoNCE, to enforce rotational equivariance of the learnt representations, a property essential to the registration task. We assess the extent of achieved rotational equivariance and the stability of the representations with respect to weight initialization, training set, and hyperparameter settings, on a remote sensing dataset of RGB and near-infrared images. We evaluate the learnt representations through registration of a biomedical dataset of bright-field and second-harmonic generation microscopy images; two modalities with very little apparent correlation. The proposed approach based on CoMIRs significantly outperforms registration of representations created by GANbased image-to-image translation, as well as a state-of-the-art, application-specific method which takes additional knowledge about the data into account. Code is available at: https://github.com/MIDA-group/CoMIR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8e5d4b1e-17ba-408b-82bf-330419cd0cbfCited by top-tier papers12
- Multi-level Feature Learning for Contrastive Multi-view ClusteringJie Xu, Huayi Tang, Yazhou Ren, Liang Peng et al.CVPR 2022 · 335 citations
- Self-Supervised Equivariant Learning for Oriented Keypoint DetectionJongmin Lee, Byungjin Kim, Minsu ChoCVPR 2022 · 39 citations
- Infusing Lattice Symmetry Priors in Attention Mechanisms for Sample-Efficient Abstract Geometric ReasoningMattia Atzeni, Mrinmaya Sachan, Andreas LoukasICML 2023 · 6 citations
- AdaTS: Learning Adaptive Time Series Representations via Dynamic Soft ContrastsDenizhan Kara, Tomoyoshi Kimura, Jinyang Li, Bowen He et al.NeurIPS 2025 · 3 citations
- InfoMAE: Pair-Efficient Cross-Modal Alignment for Multimodal Time-Series Sensing SignalsTomoyoshi Kimura, Xinlin Li, Osama A. Hanna, Yatong Chen et al.WWW 2025 · 3 citations
Builds on6
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Data-Efficient Image Recognition with Contrastive Predictive CodingOlivier J. HénaffICML 2020 · 1,553 citations
- On Mutual Information Maximization for Representation LearningMichael Tschannen, Josip Djolonga, Paul K. Rubenstein, Sylvain Gelly et al.ICLR 2020 · 559 citations
- Contrastive Learning of Structured World ModelsThomas N. Kipf, Elise van der Pol, Max WellingICLR 2020 · 322 citations
- Self-Supervised Learning of Pretext-Invariant RepresentationsIshan Misra, Laurens van der MaatenCVPR 2020
Related papers
- Modality-Agnostic Structural Image Representation Learning for Deformable Multi-Modality Medical Image RegistrationTony C. W. Mok, Zi Li, Yunhao Bai, Jianpeng Zhang et al.CVPR 2024 · 21 citations
- Contrastive Multimodal Fusion with TupleInfoNCEYunze Liu, Qingnan Fan, Shanghang Zhang, Hao Dong et al.ICCV 2021 · 84 citations
- On Discriminative Probabilistic Modeling for Self-Supervised Representation LearningBokun Wang, Yunwen Lei, Yiming Ying, Tianbao YangICLR 2025
- Exploring Patch-wise Semantic Relation for Contrastive Learning in Image-to-Image Translation TasksChanyong Jung, Gihyun Kwon, Jong Chul YeCVPR 2022 · 103 citations
- Understanding Contrastive Learning via Gaussian Mixture ModelsParikshit Bansal, Ali Kavis, Sujay SanghaviNeurIPS 2025 · 6 citations
