DLGSANet: Lightweight Dynamic Local and Global Self-Attention Network for Image Super-Resolution
Xiang Li, Jiangxin Dong, Jinhui Tang, Jinshan Pan
Abstract
We propose an effective lightweight dynamic local and global self-attention network (DLGSANet) to solve image super-resolution. Our method explores the properties of Transformers while having low computational costs. Motivated by the network designs of Transformers, we develop a simple yet effective multi-head dynamic local self-attention (MHDLSA) module to extract local features efficiently. In addition, we note that existing Transformers usually explore all similarities of the tokens between the queries and keys for the feature aggregation. However, not all the tokens from the queries are relevant to those in keys, using all the similarities does not effectively facilitate the highresolution image reconstruction. To overcome this problem, we develop a sparse global self-attention (SparseGSA) module to select the most useful similarity values so that the most useful global features can be better utilized for the high-resolution image reconstruction. We develop a hybrid dynamic-Transformer block (HDTB) that integrates the MHDLSA and SparseGSA for both local and global feature exploration. To ease the network training, we formulate the HDTBs into a residual hybrid dynamic-Transformer group (RHDTG). By embedding the RHDTGs into an end-to-end trainable network, we show that our proposed method has fewer network parameters and lower computational costs while achieving competitive performance against state-ofthe-art ones in terms of accuracy. More information is available at https://neonleexiang.github.io/ DLGSANet/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2f3398be-db61-494d-9fcc-b47156d76a5cBuilds on8
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- LAPAR: Linearly-Assembled Pixel-Adaptive Regression Network for Single Image Super-resolution and BeyondWenbo Li, Kun Zhou, Lu Qi, Nianjuan Jiang et al.NeurIPS 2020 · 293 citations
Related papers
- Recursive Generalization Transformer for Image Super-ResolutionZheng Chen, Yulun Zhang, Jinjin Gu, Linghe Kong et al.ICLR 2024 · 81 citations
- Image Super-Resolution With Non-Local Sparse AttentionYiqun Mei, Yuchen Fan, Yuqian ZhouCVPR 2021
- From Coarse to Fine: Hierarchical Pixel Integration for Lightweight Image Super-resolutionJie Liu, Chao Chen, Jie Tang, Gangshan WuAAAI 2023 · 26 citations
- CATANet: Efficient Content-Aware Token Aggregation for Lightweight Image Super-ResolutionXin Liu, Jie Liu, Jie Tang, Gangshan WuCVPR 2025
- Compacter: A Lightweight Transformer for Image RestorationZhijian Wu, Jun Li, Yang Hu, Dingjiang HuangACM MM 2024
