CAMixerSR: Only Details Need More "Attention"
Yan Wang, Yi Liu, Shijie Zhao, Junlin Li, Li Zhang
Abstract
To satisfy the rapidly increasing demands on the large image (2K-8K) super-resolution (SR), prevailing methods follow two independent tracks: 1) accelerate existing networks by content-aware routing, and 2) design better super-resolution networks via token mixer refining. Despite directness, they encounter unavoidable defects (e.g., inflexible route or non-discriminative processing) limiting further improvements of quality-complexity trade-off. To erase the drawbacks, we integrate these schemes by proposing a content-aware mixer (CAMixer), which assigns convolution for simple contexts and additional deformable window-attention for sparse textures. Specifically, the CAMixer uses a learnable predictor to generate multiple bootstraps, including offsets for windows warping, a mask for classifying windows, and convolutional attentions for endowing convolution with the dynamic property, which modulates attention to include more useful textures self-adaptively and improves the representation capability of convolution. We further introduce a global classification loss to improve the accuracy of predictors. By simply stacking CAMixers, we obtain CAMixerSR which achieves superior performance on large-image SR, lightweight SR, and omnidirectional-image SR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- DriveArena: A Closed-Loop Generative Simulation Platform for Autonomous DrivingXuemeng Yang, Licheng Wen, Tiantian Wei, Yukai Ma et al.ICCV 2025 · 13 citations
- HybridFlow: Infusing Continuity into Masked Codebook for Extreme Low-Bitrate Image CompressionLei Lu, Yanyue Xie, Wei Jiang, Wei Wang et al.ACM MM 2024 · 9 citations
- Rethinking Imbalance in Image Super-Resolution for Efficient InferenceWei Yu, Bowen Yang, Qinglin Liu, Jianing Li et al.NeurIPS 2024 · 7 citations
- PBR-SR: Mesh PBR Texture Super Resolution from 2D Image PriorsYujin Chen, Yinyu Nie, Benjamin Ummenhofer, Reiner Birkl et al.NeurIPS 2025 · 6 citations
- SRSR: Enhancing Semantic Accuracy in Real-World Image Super-Resolution with Spatially Re-Focused Text-ConditioningChen Chen, Majid Abdolshah, Violetta Shevchenko, Hongdong Li et al.NeurIPS 2025 · 3 citations
Builds on7
- DynamicViT: Efficient Vision Transformers with Dynamic Token SparsificationYongming Rao, Wenliang Zhao, Benlin Liu, Jiwen Lu et al.NeurIPS 2021 · 1,343 citations
- On the Integration of Self-Attention and ConvolutionXuran Pan, Chunjiang Ge, Rui Lu, Shiji Song et al.CVPR 2022 · 516 citations
- Feature Distillation Interaction Weighting Network for Lightweight Image Super-resolutionGuangwei Gao, Wenjie Li, Juncheng Li, Fei Wu et al.AAAI 2022 · 113 citations
- N-Gram in Swin Transformers for Efficient Lightweight Image Super-ResolutionHaram Choi, Jeongmin Lee, Jihoon YangCVPR 2023
- LAU-Net: Latitude Adaptive Upscaling Network for Omnidirectional Image Super-ResolutionXin Deng, Hao Wang, Mai Xu, Yichen Guo et al.CVPR 2021
Related papers
- UCAN: Unified Convolutional Attention Network for Expansive Receptive Fields in Lightweight Super-ResolutionCao Thien Tan, Phan Thi Thu Trang, Do Nghiem Duc, Ho Ngoc Anh et al.CVPR 2026 · 2 citations
- ShuffleMixer: An Efficient ConvNet for Image Super-ResolutionLong Sun, Jinshan Pan, Jinhui TangNeurIPS 2022 · 177 citations
- AdaFormer: Efficient Transformer with Adaptive Token Sparsification for Image Super-resolutionXiaotong Luo, Zekun Ai, Qiuyuan Liang, Ding Liu et al.AAAI 2024 · 16 citations
- Learning Texture Transformer Network for Image Super-ResolutionFuzhi Yang, Huan Yang, Jianlong Fu, Hongtao Lu et al.CVPR 2020
- Content-Aware Local GAN for Photo-Realistic Super-ResolutionJoonKyu Park, Sanghyun Son, Kyoung Mu LeeICCV 2023 · 73 citations
