Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction
Yuanhao Cai, Jing Lin, Xiaowan Hu, Haoqian Wang, Xin Yuan, Yulun Zhang, Radu Timofte, Luc Van Gool
Abstract
Hyperspectral image (HSI) reconstruction aims to recover the 3D spatial-spectral signal from a 2D measurement in the coded aperture snapshot spectral imaging (CASSI) system. The HSI representations are highly similar and correlated across the spectral dimension. Modeling the inter-spectra interactions is beneficial for HSI reconstruction. However, existing CNN-based methods show limitations in capturing spectral-wise similarity and long-range dependencies. Besides, the HSI information is modulated by a coded aperture (physical mask) in CASSI. Nonetheless, current algorithms have not fully explored the guidance effect of the mask for HSI restoration. In this paper, we propose a novel framework, Mask-guided Spectral-wise Transformer (MST), for HSI reconstruction. Specifically, we present a Spectral-wise Multi-head Self-Attention (S-MSA) that treats each spectral feature as a token and calculates self-attention along the spectral dimension. In addition, we customize a Mask-guided Mechanism (MM) that directs S- MSA to pay attention to spatial regions with high-fidelity spectral representations. Extensive experiments show that our MST significantly outperforms state-of-the-art (SOTA) methods on simulation and real HSI datasets while requiring dramatically cheaper computational and memory costs. https://github.com/caiyuanhao1998/MST/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers47
- Degradation-Aware Unfolding Half-Shuffle Transformer for Spectral Compressive ImagingYuanhao Cai, Jing Lin, Haoqian Wang, Xin Yuan et al.NeurIPS 2022 · 222 citations
- HDNet: High-resolution Dual-domain Learning for Spectral Compressive ImagingXiaowan Hu, Yuanhao Cai, Jing Lin, Haoqian Wang et al.CVPR 2022 · 193 citations
- Spatial-Spectral Transformer for Hyperspectral Image DenoisingMiaoyu Li, Ying Fu, Yulun ZhangAAAI 2023 · 115 citations
- Flow-Guided Sparse Transformer for Video DeblurringJing Lin, Yuanhao Cai, Xiaowan Hu, Haoqian Wang et al.ICML 2022 · 82 citations
- Learning to Generate Realistic Noisy Images via Pixel-level Noise-aware Adversarial TrainingYuanhao Cai, Xiaowan Hu, Haoqian Wang, Yulun Zhang et al.NeurIPS 2021 · 81 citations
Builds on18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun et al.ICCV 2021 · 2,947 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
Related papers
- Dual-Window Multiscale Transformer for Hyperspectral Snapshot Compressive ImagingFulin Luo, Xi Chen, Xiuwen Gong, Weiwen Wu et al.AAAI 2024 · 19 citations
- Improving Spectral Snapshot Reconstruction with Spectral-Spatial RectificationJiancheng Zhang, Haijin Zeng, Yongyong Chen, Dengxiu Yu et al.CVPR 2024 · 10 citations
- Joint Spectral Image Reconstruction and Semantic Segmentation with Cooperative UnfoldingZijun He, Ping Wang, Xiaodong Wang, Chang Chen et al.CVPR 2026
- Spectral Compressive Imaging via Chromaticity-Intensity DecompositionXiaodong Wang, Zijun He, Ping Wang, Lishun Wang et al.NeurIPS 2025 · 4 citations
- SGDE: Self-supervised Geometry Degradation Estimation Framework for Coded Aperture Compressive Spectral ImagingYuqiao He, Xiaoyan Liu, Jianxu Mao, Yaonan Wang et al.CVPR 2026
