Revisiting Dynamic Convolution via Matrix Decomposition
Yunsheng Li, Yinpeng Chen, Xiyang Dai, Mengchen Liu, Dongdong Chen, Ye Yu, Lu Yuan, Zicheng Liu, Mei Chen, Nuno Vasconcelos
Abstract
Recent research in dynamic convolution shows substantial performance boost for efficient CNNs, due to the adaptive aggregation of K static convolution kernels. It has two limitations: (a) it increases the number of convolutional weights by Ktimes, and (b) the joint optimization of dynamic attention and static convolution kernels is challenging. In this paper, we revisit it from a new perspective of matrix decomposition and reveal the key issue is that dynamic convolution applies dynamic attention over channel groups after projecting into a higher dimensional latent space. To address this issue, we propose dynamic channel fusion to replace dynamic attention over channel groups. Dynamic channel fusion not only enables significant dimension reduction of the latent space, but also mitigates the joint optimization difficulty. As a result, our method is easier to train and requires significantly fewer parameters without sacrificing accuracy. Source code is at https://github.com/liyunsheng13/dcd .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b83a6038-848e-4a82-bec8-3833d5349da8Cited by top-tier papers14
- Omni-Dimensional Dynamic ConvolutionChao Li, Aojun Zhou, Anbang YaoICLR 2022 · 408 citations
- Self-Sustaining Representation Expansion for Non-Exemplar Class-Incremental LearningKai Zhu, Wei Zhai, Yang Cao, Jiebo Luo et al.CVPR 2022 · 155 citations
- TAda! Temporally-Adaptive Convolutions for Video UnderstandingZiyuan Huang, Shiwei Zhang, Liang Pan, Zhiwu Qing et al.ICLR 2022 · 72 citations
- Multilinear Mixture of Experts: Scalable Expert Specialization through FactorizationJames Oldfield, Markos Georgopoulos, Grigorios Chrysos, Christos Tzelepis et al.NeurIPS 2024 · 41 citations
- Super-efficient Echocardiography Video Segmentation via Proxy- and Kernel-Based Semi-supervised LearningHuisi Wu, Jingyin Lin, Wende Xie, Jing QinAAAI 2023 · 16 citations
Builds on7
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- GhostNet: More Features From Cheap OperationsKai Han, Yunhe Wang, Qi Tian, Jianyuan Guo et al.CVPR 2020
- Dynamic Convolution: Attention Over Convolution KernelsYinpeng Chen, Xiyang Dai, Mengchen Liu, Dongdong Chen et al.CVPR 2020
- Dynamic Region-Aware ConvolutionJin Chen, Xijun Wang, Zichao Guo, Xiangyu Zhang et al.CVPR 2021
- Factorized Higher-Order CNNs With an Application to Spatio-Temporal Emotion EstimationJean Kossaifi, Antoine Toisoul, Adrian Bulat, Yannis Panagakis et al.CVPR 2020
Related papers
- MAGIC: Rethinking Dynamic Convolution Design for Medical Image SegmentationShijie Li, Yunbin Tu, Qingyuan Xiang, Zheng LiACM MM 2024 · 8 citations
- Decoupled Dynamic Filter NetworksJingkai Zhou, Varun Jampani, Zhixiong Pi, Qiong Liu et al.CVPR 2021
- Efficient Equivariant NetworkLingshen He, Yuxuan Chen, Zhengyang Shen, Yiming Dong et al.NeurIPS 2021 · 46 citations
- KernelWarehouse: Rethinking the Design of Dynamic ConvolutionChao Li, Anbang YaoICML 2024 · 12 citations
- PartialNet: Compute Less, Perform BetterHaiduo Huang, Tian Xia, Wenzhe Zhao, Pengju RenAAAI 2026
