Frequency Dynamic Convolution for Dense Image Prediction
Linwei Chen, Lin Gu, Liang Li, Chenggang Yan, Ying Fu
摘要
While Dynamic Convolution (DY-Conv) has shown promising performance by enabling adaptive weight selection through multiple parallel weights combined with an attention mechanism, the frequency response of these weights tends to exhibit high similarity, resulting in high parameter costs but limited adaptability. In this work, we introduce Frequency Dynamic Convolution (FDConv), a novel approach that mitigates these limitations by learning a fixed parameter budget in the Fourier domain. FDConv divides this budget into frequency-based groups with disjoint Fourier indices, enabling the construction of frequencydiverse weights without increasing the parameter cost. To further enhance adaptability, we propose Kernel Spatial Modulation (KSM) and Frequency Band Modulation (FBM). KSM dynamically adjusts the frequency response of each filter at the spatial level, while FBM decomposes weights into distinct frequency bands in the frequency domain and modulates them dynamically based on local content. Extensive experiments on object detection, segmentation, and classification validate the effectiveness of FD-Conv. We demonstrate that when applied to ResNet-50, FDConv achieves superior performance with a modest increase of +3.6M parameters, outperforming previous methods that require substantial increases in parameter budgets (e.g., CondConv +90M, KW +76.5M). Moreover, FD-Conv seamlessly integrates into a variety of architectures, including ConvNeXt, Swin-Transformer, offering a flexible and efficient solution for modern vision tasks. The code is made publicly available at https://github.com/Linwei- Chen/FDConv.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Frequency-Dynamic Attention Modulation for Dense PredictionLinwei Chen, Lin Gu, Ying FuICCV 2025 · 被引用 13 次
- Fourier Angle Alignment for Oriented Object Detection in Remote SensingChangyu Gu, Linwei Chen, Lin Gu, Ying FuCVPR 2026 · 被引用 9 次
- Statistical Characteristic-Guided Denoising for Rapid High-Resolution Transmission Electron Microscopy ImagingHesong Li, Ziqi Wu, Ruiwen Shao, Ying FuCVPR 2026 · 被引用 4 次
- MFEN: Multi-Frequency Expert Network for Visible-Infrared Person Re-IDXulin Li, Yan Lu, Bin Liu, Qinhong Yang 等CVPR 2026 · 被引用 2 次
- Revisiting Downsampling in Semantic Segmentation: Fighting Aliasing with Dynamic Gaussian and Gabor Frequency FiltersYuBing Luo, Nian Shi, Jia Qin, Zekai Ji 等AAAI 2026
它引用的顶会 Paper31
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
相关 Paper
- Frequency-Adaptive Dilated Convolution for Semantic SegmentationLinwei Chen, Lin Gu, Dezhi Zheng, Ying FuCVPR 2024
- KernelWarehouse: Rethinking the Design of Dynamic ConvolutionChao Li, Anbang YaoICML 2024 · 被引用 12 次
- Differentiable Learning-to-Group Channels via Groupable Convolutional Neural NetworksZhaoyang Zhang, Jingyu Li, Wenqi Shao, Zhanglin Peng 等ICCV 2019 · 被引用 39 次
- PeLK: Parameter-Efficient Large Kernel ConvNets with Peripheral ConvolutionHonghao Chen, Xiangxiang Chu, Yongjian Ren, Xin Zhao 等CVPR 2024
- FlexConv: Continuous Kernel Convolutions With Differentiable Kernel SizesDavid W. Romero, Robert-Jan Bruintjes, Jakub Mikolaj Tomczak, Erik J. Bekkers 等ICLR 2022 · 被引用 94 次
