Rubik's Cube: High-Order Channel Interactions with a Hierarchical Receptive Field
Naishan Zheng, Man Zhou, Chong Zhou, Chen Change Loy
Abstract
Image restoration techniques, spanning from the convolution to the transformer paradigm, have demonstrated robust spatial representation capabilities to deliver high-quality performance. Yet, many of these methods, such as convolution and the Feed Forward Network (FFN) structure of transformers, primarily leverage the basic first-order channel interactions and have not maximized the potential benefits of higher-order modeling. To address this limitation, our research dives into understanding relationships within the channel dimension and introduces a simple yet efficient, high-order channel-wise operator tailored for image restoration. Instead of merely mimicking high-order spatial interaction, our approach offers several added benefits: Efficiency: It adheres to the zero-FLOP and zero-parameter principle, using a spatial-shifting mechanism across channel-wise groups. Simplicity: It turns the favorable channel interaction and aggregation capabilities into element-wise multiplications and convolution units with 1 × 1 kernel. Our new formulation expands the first-order channel-wise interactions seen in previous works to arbitrary high orders, generating a hierarchical receptive field akin to a Rubik's cube through the combined action of shifting and interactions. Furthermore, our proposed Rubik's cube convolution is a flexible operator that can be incorporated into existing image restoration networks, serving as a drop-in replacement for the standard convolution unit with fewer parameters overhead. We conducted experiments across various low-level vision tasks, including image denoising, low-light image enhancement, guided image super-resolution, and image de-blurring. The results consistently demonstrate that our Rubik's cube operator enhances performance across all tasks. Code is publicly available at https://github.com/zheng980629/RubikCube.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 12c59a9f-1311-4d59-87f0-3fd4ee6e0448Builds on26
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- HorNet: Efficient High-Order Spatial Interactions with Recursive Gated ConvolutionsYongming Rao, Wenliang Zhao, Yansong Tang, Jie Zhou et al.NeurIPS 2022 · 422 citations
- Human-Aware Motion DeblurringZiyi Shen, Wenguan Wang, Xiankai Lu, Jianbing Shen et al.ICCV 2019 · 374 citations
Related papers
- Attention Cube Network for Image RestorationYucheng Hang, Qingmin Liao, Wenming Yang, Yupeng Chen et al.ACM MM 2020 · 22 citations
- Random Shuffle Transformer for Image RestorationJie Xiao, Xueyang Fu, Man Zhou, Hongjian Liu et al.ICML 2023 · 38 citations
- Scale-Wise Convolution for Image RestorationYuchen Fan, Jiahui Yu, Ding Liu, Thomas S. HuangAAAI 2020 · 45 citations
- Improving the Learning Capability of Small-size Image Restoration Network by Deep Fourier ShiftingMan ZhouNeurIPS 2024 · 1 citation
- Selective Frequency Network for Image RestorationYuning Cui, Yi Tao, Zhenshan Bing, Wenqi Ren et al.ICLR 2023
