Omni Aggregation Networks for Lightweight Image Super-Resolution
Hang Wang, Xuanhong Chen, Bingbing Ni, Yutian Liu, Jinfan Liu
Abstract
While lightweight ViT framework has made tremendous progress in image super-resolution, its uni-dimensional self-attention modeling, as well as homogeneous aggregation scheme, limit its effective receptive field (ERF) to include more comprehensive interactions from both spatial and channel dimensions. To tackle these drawbacks, this work proposes two enhanced components under a new Omni-SR architecture. First, an Omni Self-Attention (OSA) block is proposed based on dense interaction principle, which can simultaneously model pixel-interaction from both spatial and channel dimensions, mining the potential correlations across omni-axis (i.e., spatial and channel). Coupling with mainstream window partitioning strategies, OSA can achieve superior performance with compelling computational budgets. Second, a multi-scale interaction scheme is proposed to mitigate sub-optimal ERF (i.e., premature saturation) in shallow models, which facilitates local propagation and meso-/global-scale interactions, rendering an omni-scale aggregation building block. Extensive experiments demonstrate that Omni-SR achieves recordhigh performance on lightweight super-resolution benchmarks (e.g., 26.95dB@Urban100 ×4 with only 792K parameters). Our code is available at https://github . com/Francis0625/Omni-SR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers26
- Transcending the Limit of Local Window: Advanced Super-Resolution Transformer with Adaptive Token DictionaryLeheng Zhang, Yawei Li, Xingyu Zhou, Xiaorui Zhao et al.CVPR 2024 · 73 citations
- I2EBench: A Comprehensive Benchmark for Instruction-based Image EditingYiwei Ma, Jiayi Ji, Ke Ye, Weihuang Lin et al.NeurIPS 2024 · 67 citations
- 4KAgent: Agentic Any Image to 4K Super-ResolutionYushen Zuo, Qi Zheng, Mingyang Wu, Xinrui Jiang et al.NeurIPS 2025 · 51 citations
- See More Details: Efficient Image Super-Resolution by Experts MiningEduard Zamfir, Zongwei Wu, Nancy Mehta, Yulun Zhang et al.ICML 2024 · 38 citations
- Image Processing GNN: Breaking Rigidity in Super-ResolutionYuchuan Tian, Hanting Chen, Chao Xu, Yunhe WangCVPR 2024 · 34 citations
Builds on17
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision TransformerSachin Mehta, Mohammad RastegariICLR 2022 · 2,162 citations
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani et al.ICCV 2019 · 1,149 citations
Related papers
- Dual Aggregation Transformer for Image Super-ResolutionZheng Chen, Yulun Zhang, Jinjin Gu, Linghe Kong et al.ICCV 2023 · 345 citations
- Compacter: A Lightweight Transformer for Image RestorationZhijian Wu, Jun Li, Yang Hu, Dingjiang HuangACM MM 2024
- Learning Omni-Frequency Region-adaptive Representations for Real Image Super-ResolutionXin Li, Xin Jin, Tao Yu, Simeng Sun et al.AAAI 2021 · 50 citations
- From Coarse to Fine: Hierarchical Pixel Integration for Lightweight Image Super-resolutionJie Liu, Chao Chen, Jie Tang, Gangshan WuAAAI 2023 · 26 citations
- Spatially-Adaptive Feature Modulation for Efficient Image Super-ResolutionLong Sun, Jiangxin Dong, Jinhui Tang, Jinshan PanICCV 2023 · 211 citations
