SpecSolver: Solving Spatial-Spectral Fusion via Semantic Transformer
Wei Li, Junwei Zhu, Honghui Xu, Jiawei Jiang, Jianwei Zheng
摘要
By clustering pixels with locally similar values, superpixel-based approaches have shown great potential in processing hyperspectral images (HSI), thereby reducing the computational burden associated with large spatial dimensions. However, specific for spatial-spectral fusion (SSF), superpixel segmentation is inherently non-differentiable and irreversible; hence it is inapplicable. To address the issues, we propose a semantic transformer-based solver, namely SpecSolver, which is basically inspired by the benefits of superpixel-based approaches, yet with the inner mechanism completely improved. The core idea lies in learning the intrinsic semantic states of HSIs hidden behind discretized pixel representations. Specifically, we propose a new Semantic-Attention to adaptively split the image domain into a series of learnable slices of flexible shapes, where image pixels under similar semantic states will be ascribed to the same slice. By calculating attention to the Semantic-Superpixel tokens encoded from slices, SpecSolver can effectively capture intricate semantic correlations from the vast number of pixels, which also empowers the solver with an endogenous capacity for modeling different magnification scales and allows for efficient computation in linear complexity. On that basis, we elaborate a SpatialNet module, which extracts multiscale local spectral information, and a FreqNet module, which supplements global information, capturing subtle details and variations across different spectra. Experiments on two benchmark SSF datasets verify the state-of-the-art (SOTA) performance of the proposed method, both visually and quantitatively. Also, ablation studies validate the mentioned contributions.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Spatial-Spectral Transformer for Hyperspectral Image DenoisingMiaoyu Li, Ying Fu, Yulun ZhangAAAI 2023 · 被引用 115 次
- HyperTransformer: A Textural and Spectral Feature Fusion Transformer for PansharpeningWele Gedara Chaminda Bandara, Vishal M. PatelCVPR 2022 · 被引用 175 次
- SCPSN: Spectral Clustering-based Pyramid Super-resolution Network for Hyperspectral ImagesYong Yang, Aoqi Zhao, Shuying Huang, Xiaozheng Wang 等ACM MM 2024 · 被引用 5 次
- U2Net: A General Framework with Spatial-Spectral-Integrated Double U-Net for Image FusionSiran Peng, Chenhao Guo, Xiao Wu, Liang-Jian DengACM MM 2023 · 被引用 43 次
- Breaking the Spatial-Temporal Consistency Constraint: Towards Reference-Based Hyperspectral Image Super-ResolutionXuyao Liu, Jiahui Qu, Wenqian DongACM MM 2025
