Rubik's Cube: High-Order Channel Interactions with a Hierarchical Receptive Field
Naishan Zheng, Man Zhou, Chong Zhou, Chen Change Loy
摘要
Image restoration techniques, spanning from the convolution to the transformer paradigm, have demonstrated robust spatial representation capabilities to deliver high-quality performance. Yet, many of these methods, such as convolution and the Feed Forward Network (FFN) structure of transformers, primarily leverage the basic first-order channel interactions and have not maximized the potential benefits of higher-order modeling. To address this limitation, our research dives into understanding relationships within the channel dimension and introduces a simple yet efficient, high-order channel-wise operator tailored for image restoration. Instead of merely mimicking high-order spatial interaction, our approach offers several added benefits: Efficiency: It adheres to the zero-FLOP and zero-parameter principle, using a spatial-shifting mechanism across channel-wise groups. Simplicity: It turns the favorable channel interaction and aggregation capabilities into element-wise multiplications and convolution units with 1 × 1 kernel. Our new formulation expands the first-order channel-wise interactions seen in previous works to arbitrary high orders, generating a hierarchical receptive field akin to a Rubik's cube through the combined action of shifting and interactions. Furthermore, our proposed Rubik's cube convolution is a flexible operator that can be incorporated into existing image restoration networks, serving as a drop-in replacement for the standard convolution unit with fewer parameters overhead. We conducted experiments across various low-level vision tasks, including image denoising, low-light image enhancement, guided image super-resolution, and image de-blurring. The results consistently demonstrate that our Rubik's cube operator enhances performance across all tasks. Code is publicly available at https://github.com/zheng980629/RubikCube.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper26
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou 等CVPR 2022 · 被引用 1,970 次
- HorNet: Efficient High-Order Spatial Interactions with Recursive Gated ConvolutionsYongming Rao, Wenliang Zhao, Yansong Tang, Jie Zhou 等NeurIPS 2022 · 被引用 422 次
- Human-Aware Motion DeblurringZiyi Shen, Wenguan Wang, Xiankai Lu, Jianbing Shen 等ICCV 2019 · 被引用 374 次
相关 Paper
- Attention Cube Network for Image RestorationYucheng Hang, Qingmin Liao, Wenming Yang, Yupeng Chen 等ACM MM 2020 · 被引用 22 次
- Random Shuffle Transformer for Image RestorationJie Xiao, Xueyang Fu, Man Zhou, Hongjian Liu 等ICML 2023 · 被引用 38 次
- Scale-Wise Convolution for Image RestorationYuchen Fan, Jiahui Yu, Ding Liu, Thomas S. HuangAAAI 2020 · 被引用 45 次
- Improving the Learning Capability of Small-size Image Restoration Network by Deep Fourier ShiftingMan ZhouNeurIPS 2024 · 被引用 1 次
- Selective Frequency Network for Image RestorationYuning Cui, Yi Tao, Zhenshan Bing, Wenqi Ren 等ICLR 2023
