Single Pair Cross-Modality Super Resolution
Guy Shacht, Dov Danon, Sharon Fogel, Daniel Cohen-Or
Abstract
Non-visual imaging sensors are widely used in the industry for different purposes. Those sensors are more expensive than visual (RGB) sensors, and usually produce images with lower resolution. To this end, Cross-Modality Super-Resolution methods were introduced, where an RGB image of a high-resolution assists in increasing the resolution of a low-resolution modality. However, fusing images from different modalities is not a trivial task, since each multimodal pair varies greatly in its internal correlations. For this reason, traditional state-of-the-arts which are trained on external datasets often struggle with yielding an artifactfree result that is still loyal to the target modality characteristics. We present CMSR, a single-pair approach for Cross-Modality Super-Resolution. The network is internally trained on the two input images only, in a self-supervised manner, learns their internal statistics and correlations, and applies them to up-sample the target modality. CMSR contains an internal transformer which is trained on-the-fly together with the up-sampling process itself and without supervision, to allow dealing with pairs that are only weakly aligned. We show that CMSR produces state-of-the-art super resolved images, yet without introducing artifacts or irrelevant details that originate from the RGB image only.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 59eb8864-0e48-4c60-9155-c27190dabb34Cited by top-tier papers3
- Learning Graph Regularisation for Guided Super-ResolutionRiccardo de Lutio, Alexander Becker, Stefano D'Aronco, Stefania Russo et al.CVPR 2022 · 40 citations
- Regularization-free Diffeomorphic Temporal Alignment NetsRon Shapira Weber, Oren FreifeldICML 2023 · 10 citations
- Metadata-Based RAW Reconstruction via Implicit Neural FunctionsLeyi Li, Huijie Qiao, Qi Ye, Qinmin YangCVPR 2023
Builds on1
Related papers
- Learning Cross-Spectral Prior for Image Super-ResolutionChenxi Ma, Weimin Tan, Shili Zhou, Bo YanACM MM 2024
- RGB-Multispectral Matching: Dataset, Learning Methodology, EvaluationFabio Tosi, Pierluigi Zama Ramirez, Matteo Poggi, Samuele Salti et al.CVPR 2022 · 5 citations
- Transformer-empowered Multi-scale Contextual Matching and Aggregation for Multi-contrast MRI Super-resolutionGuangyuan Li, Jun Lv, Yapeng Tian, Qi Dou et al.CVPR 2022 · 100 citations
- 3M-TI: High-Quality Mobile Thermal Imaging via Calibration-free Multi-Camera Cross-Modal DiffusionMinchong Chen, Xiaoyun Yuan, Junzhe Wan, Jianing Zhang et al.CVPR 2026 · 2 citations
- Learning Scene Structure Guidance via Cross-Task Knowledge Transfer for Single Depth Super-ResolutionBaoli Sun, Xinchen Ye, Baopu Li, Haojie Li et al.CVPR 2021
