VolFormer: Explore More Comprehensive Cube Interaction for Hyperspectral Image Restoration and Beyond
Dabing Yu, Zheng Gao
摘要
Capitalizing on the talent of self-attention in capturing non-local features, Transformer architectures have exhibited remarkable performance in single hyperspectral image restoration. For hyperspectral images, each pixel is located in the hyperspectral image cubes with a large spectral dimension and two spatial dimensions. Although uni-dimensional self-attention, like channel self-attention or spatial self-attention, builds long-range dependencies in spectral or spatial dimensions, they lack more comprehensive interactions across dimensions. To tackle the above drawback, we propose a VolFormer, a volumetric self-attention embedded Transformer network for single hyperspectral image restoration. Specifically, we propose volumetric self-attention (VolSA), which extends the interaction from 2D flat to 3D cube. VolSA can simultaneously model token interaction in the 3D cube, mining the potential correlations between the hyperspectral image cube. An attention decomposition form is proposed to reduce the computational burden of modeling volumetric information. In practical terms, VolSA adapts double similarity matrixes in spatial and channel dimensions to implicitly model 3D context information while transforming the complexity from cubic to quadratic. Additionally, we introduce the explicit spectral location prior to enhance the proposed self-attention. This property allows the target token to perceive global spectral information while simultaneously assigning different levels of attention to tokens at varying wavelength bands. Extensive experiments demonstrate that VolFormer achieves record-high performance on hyperspectral image super-resolution, denoise and classification benchmarks. The source code is available at https://github.com/yudadabing/VolFormer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Real Noise Decoupling for Hyperspectral Image DenoisingYingkai Zhang, Tao Zhang, Jing Nie, Ying FuAAAI 2026 · 被引用 3 次
- Degradation-Aware Metric Prompting for Hyperspectral Image RestorationBinfeng Wang, Di Wang, Haonan Guo, Ying Fu 等ICML 2026 · 被引用 2 次
它引用的顶会 Paper11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image ReconstructionYuanhao Cai, Jing Lin, Xiaowan Hu, Haoqian Wang 等CVPR 2022 · 被引用 310 次
- Spatial-Spectral Transformer for Hyperspectral Image DenoisingMiaoyu Li, Ying Fu, Yulun ZhangAAAI 2023 · 被引用 115 次
相关 Paper
- ESSAformer: Efficient Transformer for Hyperspectral Image Super-resolutionMingjin Zhang, Chi Zhang, Qiming Zhang, Jie Guo 等ICCV 2023 · 被引用 73 次
- Learning Spectral-wise Correlation for Spectral Super-Resolution: Where Similarity Meets ParticularityHongyuan Wang, Lizhi Wang, Chang Chen, Xue Hu 等ACM MM 2023 · 被引用 12 次
- Hybrid Spectral Denoising Transformer with Guided AttentionZeqiang Lai, Chenggang Yan, Ying FuICCV 2023 · 被引用 35 次
- Spectral Enhanced Rectangle Transformer for Hyperspectral Image DenoisingMiaoyu Li, Ji Liu, Ying Fu, Yulun Zhang 等CVPR 2023
- SCPSN: Spectral Clustering-based Pyramid Super-resolution Network for Hyperspectral ImagesYong Yang, Aoqi Zhao, Shuying Huang, Xiaozheng Wang 等ACM MM 2024 · 被引用 5 次
