Feature Modulation Transformer: Cross-Refinement of Global Representation via High-Frequency Prior for Image Super-Resolution
Ao Li, Le Zhang, Yun Liu, Ce Zhu
摘要
Transformer-based methods have exhibited remarkable potential in single image super-resolution (SISR) by effectively extracting long-range dependencies. However, most of the current research in this area has prioritized the design of transformer blocks to capture global information, while overlooking the importance of incorporating high-frequency priors, which we believe could be beneficial. In our study, we conducted a series of experiments and found that transformer structures are more adept at capturing low-frequency information, but have limited capacity in constructing high-frequency representations when compared to their convolutional counterparts. Our proposed solution, the cross-refinement adaptive feature modulation transformer (CRAFT), integrates the strengths of both convolutional and transformer structures. It comprises three key components: the high-frequency enhancement residual block (HFERB) for extracting high-frequency information, the shift rectangle window attention block (SRWAB) for capturing global information, and the hybrid fusion block (HFB) for refining the global representation. Our experiments on multiple datasets demonstrate that CRAFT outperforms state-of-the-art methods by up to 0.29dB while using fewer parameters. The source code will be made available at: https://github.com/AVC2-UESTC/CRAFT-SR.git.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Efficient Face Super-Resolution via Wavelet-based Feature Enhancement NetworkWenjie Li, Heng Guo, Xuannan Liu, Kongming Liang 等ACM MM 2024 · 被引用 77 次
- LoFormer: Local Frequency Transformer for Image DeblurringXintian Mao, Jiansheng Wang, Xingran Xie, Qingli Li 等ACM MM 2024 · 被引用 44 次
- MobileIE: An Extremely Lightweight and Effective ConvNet for Real-Time Image Enhancement on Mobile DevicesHailong Yan, Ao Li, Xiangtao Zhang, Zhe Liu 等ICCV 2025 · 被引用 12 次
- Boosting Flow-based Generative Super-Resolution Models via Learned PriorLi-Yuan Tsao, Yi-Chen Lo, Chia-Che Chang, Hao-Wei Chen 等CVPR 2024 · 被引用 10 次
- WiFi CSI Based Temporal Activity Detection via Dual Pyramid NetworkZhendong Liu, Le Zhang, Bing Li, Yingjie Zhou 等AAAI 2025 · 被引用 6 次
它引用的顶会 Paper14
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- CrossFormer: A Versatile Vision Transformer Hinging on Cross-scale AttentionWenxiao Wang, Lu Yao, Long Chen, Binbin Lin 等ICLR 2022 · 被引用 367 次
- LAPAR: Linearly-Assembled Pixel-Adaptive Regression Network for Single Image Super-resolution and BeyondWenbo Li, Kun Zhou, Lu Qi, Nianjuan Jiang 等NeurIPS 2020 · 被引用 293 次
- Cross Aggregation Transformer for Image RestorationZheng Chen, Yulun Zhang, Jinjin Gu, Yongbing Zhang 等NeurIPS 2022 · 被引用 274 次
相关 Paper
- Activating More Pixels in Image Super-Resolution TransformerXiangyu Chen, Xintao Wang, Jiantao Zhou, Yu Qiao 等CVPR 2023
- Recursive Generalization Transformer for Image Super-ResolutionZheng Chen, Yulun Zhang, Jinjin Gu, Linghe Kong 等ICLR 2024 · 被引用 81 次
- Cross-Modality High-Frequency Transformer for MR Image Super-ResolutionChaowei Fang, Dingwen Zhang, Liang Wang, Yulun Zhang 等ACM MM 2022 · 被引用 57 次
- Learning Texture Transformer Network for Image Super-ResolutionFuzhi Yang, Huan Yang, Jianlong Fu, Hongtao Lu 等CVPR 2020
- From Coarse to Fine: Hierarchical Pixel Integration for Lightweight Image Super-resolutionJie Liu, Chao Chen, Jie Tang, Gangshan WuAAAI 2023 · 被引用 26 次
