Learning Multi-Modal Cross-Scale Deformable Transformer Network for Unregistered Hyperspectral Image Super-resolution
Wenqian Dong, Yang Xu, Jiahui Qu, Shaoxiong Hou
Abstract
Hyperspectral image super-resolution (HSI-SR) is a technology to improve the spatial resolution of HSI. Existing fusion-based SR methods have shown great performance, but still have some problems as follows: 1) existing methods assume that the auxiliary image providing spatial information is strictly registered with the HSI, but images are difficult to be registered finely due to the shooting platforms, shooting viewpoints and the influence of atmospheric turbulence; 2) most of the methods are based on convolutional neural networks (CNNs), which is effective for local features but cannot utilize the global features. To this end, we propose a multi-modal cross-scale deformable transformer network (M 2 DTN) to achieve unregistered HSI-SR. Specifically, we formulate a spectrum-preserving based spatialguided registration-SR unified model (SSRU) from the view of the realistic degradation scenarios. According to SSRU, we propose the multi-modal registration deformable module (MMRD) to align features between different modalities by deformation field. In order to efficiently utilize the unique information between different modals, we design the multiscale feature transformer (MSFT) to emphasize the spatialspectral features at different scales. In addition, we propose the cross-scale feature aggregation module (CSFA) to accurately reconstruct the HSI by aggregating feature information at different scales. Experiments show that M 2 DTN outperforms the-state-of-the-art HSI-SR methods. Code is obtainable at https://github.com/Jiahuiqu/M2DTN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 550f1026-7a8b-46b5-aff7-0ce20844ba69Cited by top-tier papers1
Ask how each one uses itBuilds on4
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Recursive Cascaded Networks for Unsupervised Medical Image RegistrationShengyu Zhao, Yue Dong, Eric I-Chao Chang, Yan XuICCV 2019 · 289 citations
- HyperTransformer: A Textural and Spectral Feature Fusion Transformer for PansharpeningWele Gedara Chaminda Bandara, Vishal M. PatelCVPR 2022 · 175 citations
- Learning Texture Transformer Network for Image Super-ResolutionFuzhi Yang, Huan Yang, Jianlong Fu, Hongtao Lu et al.CVPR 2020
Related papers
- MFTN: A Multi-scale Feature Transfer Network Based on IMatchFormer for Hyperspectral Image Super-ResolutionShuying Huang, Mingyang Ren, Yong Yang, Xiaozheng Wang et al.ICML 2024 · 2 citations
- Enhancing Unregistered Hyperspectral Image Super-Resolution via Unmixing-based Abundance Fusion LearningYingkai Zhang, Tao Zhang, Jing Nie, Ying FuCVPR 2026 · 6 citations
- Breaking the Spatial-Temporal Consistency Constraint: Towards Reference-Based Hyperspectral Image Super-ResolutionXuyao Liu, Jiahui Qu, Wenqian DongACM MM 2025
- Cycle-Consistent Mamba-Based Registration-Fusion Joint Network for Unregistered Hyperspectral Image Super-ResolutionQuangui He, Jiahui Qu, Wenqian Dong, Song Xiao et al.ACM MM 2025
- SCPSN: Spectral Clustering-based Pyramid Super-resolution Network for Hyperspectral ImagesYong Yang, Aoqi Zhao, Shuying Huang, Xiaozheng Wang et al.ACM MM 2024 · 5 citations
