Omni-Dimensional Dynamic Convolution
Chao Li, Aojun Zhou, Anbang Yao
摘要
Learning a single static convolutional kernel 1 in each convolutional layer is the common training paradigm of modern Convolutional Neural Networks (CNNs). Instead, recent research in dynamic convolution shows that learning a linear combination of n convolutional kernels weighted with their input-dependent attentions can significantly improve the accuracy of light-weight CNNs, while maintaining efficient inference. However, we observe that existing works endow convolutional kernels with the dynamic property through one dimension (regarding the convolutional kernel number) of the kernel space, but the other three dimensions (regarding the spatial size, the input channel number and the output channel number for each convolutional kernel) are overlooked. Inspired by this, we present Omni-dimensional Dynamic Convolution (ODConv), a more generalized yet elegant dynamic convolution design, to advance this line of research. ODConv leverages a novel multi-dimensional attention mechanism with a parallel strategy to learn complementary attentions for convolutional kernels along all four dimensions of the kernel space at any convolutional layer. As a drop-in replacement of regular convolutions, ODConv can be plugged into many CNN architectures. Extensive experiments on the ImageNet and MS-COCO datasets show that OD-Conv brings solid accuracy boosts for various prevailing CNN backbones including both light-weight and large ones, e.g., 3.77%∼5.71%|1.86%∼3.72% absolute top-1 improvements to MobivleNetV2|ResNet family on the ImageNet dataset. Intriguingly, thanks to its improved feature learning ability, ODConv with even one single kernel can compete with or outperform existing dynamic convolution counterparts with multiple kernels, substantially reducing extra parameters. Furthermore, ODConv is also superior to other attention modules for modulating the output features or the convolutional weights. Code and models are available at https://github.com/OSVAI/ODConv . * This work was done when Chao Li was an intern at Intel Labs China, supervised by Anbang Yao who proposed the original idea and led the writing of the paper. † Corresponding author. 1 Here, we follow the definitions in (Yang et al., 2019; Chen et al., 2020) where a convolutional kernel refers to the filter set of a convolutional layer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Adaptive Dynamic Filtering Network for Image DenoisingHao Shen, Zhong-Qiu Zhao, Wandi ZhangAAAI 2023 · 被引用 69 次
- Reasoning Beyond Points: A Visual Introspective Approach for Few-Shot 3D SegmentationChangshuo Wang, Shuting He, Xiang Fang, Zhijian Hu 等NeurIPS 2025 · 被引用 28 次
- SAVSR: Arbitrary-Scale Video Super-Resolution via a Learned Scale-Adaptive NetworkZekun Li, Hongying Liu, Fanhua Shang, Yuanyuan Liu 等AAAI 2024 · 被引用 23 次
- KernelWarehouse: Rethinking the Design of Dynamic ConvolutionChao Li, Anbang YaoICML 2024 · 被引用 12 次
- Weak Distribution Detectors Lead to Stronger Generalizability of Vision-Language Prompt TuningKun Ding, Haojian Zhang, Qiang Yu, Ying Wang 等AAAI 2024 · 被引用 8 次
它引用的顶会 Paper6
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- SRM: A Style-Based Recalibration Module for Convolutional Neural NetworksHyunJae Lee, Hyo-Eun Kim, Hyeonseob NamICCV 2019 · 被引用 286 次
- DynamoNet: Dynamic Action and Motion NetworkAli Diba, Vivek Sharma, Luc Van Gool, Rainer StiefelhagenICCV 2019 · 被引用 123 次
- Revisiting Dynamic Convolution via Matrix DecompositionYunsheng Li, Yinpeng Chen, Xiyang Dai, Mengchen Liu 等ICLR 2021 · 被引用 82 次
- Dynamic Convolution: Attention Over Convolution KernelsYinpeng Chen, Xiyang Dai, Mengchen Liu, Dongdong Chen 等CVPR 2020
相关 Paper
- MAGIC: Rethinking Dynamic Convolution Design for Medical Image SegmentationShijie Li, Yunbin Tu, Qingyuan Xiang, Zheng LiACM MM 2024 · 被引用 8 次
- Dynamic Region-Aware ConvolutionJin Chen, Xijun Wang, Zichao Guo, Xiangyu Zhang 等CVPR 2021
- Decoupled Dynamic Filter NetworksJingkai Zhou, Varun Jampani, Zhixiong Pi, Qiong Liu 等CVPR 2021
- Convolutional Networks with Oriented 1D KernelsAlexandre Kirchmeyer, Jia DengICCV 2023 · 被引用 9 次
- PartialNet: Compute Less, Perform BetterHaiduo Huang, Tian Xia, Wenzhe Zhao, Pengju RenAAAI 2026
