Adaptive Convolutions with Per-pixel Dynamic Filter Atom
Ze Wang, Zichen Miao, Jun Hu, Qiang Qiu
Abstract
Applying feature dependent network weights have been proved to be effective in many fields. However, in practice, restricted by the enormous size of model parameters and memory footprints, scalable and versatile dynamic convolutions with per-pixel adapted filters are yet to be fully explored. In this paper, we address this challenge by de-composing filters, adapted to each spatial position, over dynamic filter atoms generated by a light-weight network from local features. Adaptive receptive fields can be supported by further representing each filter atom over sets of pre-fixed multi-scale bases. As plug-and-play replacements to convolutional layers, the introduced adaptive convolutions with per-pixel dynamic atoms enable explicit modeling of intra-image variance, while avoiding heavy computation, parameters, and memory cost. Our method preserves the appealing properties of conventional convolutions as being translation-equivariant and parametrically efficient. We present experiments to show that, the proposed method delivers comparable or even better performance across tasks, and are particularly effective on handling tasks with significant intra-image variance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9ec3696f-7cbe-4ae5-93a4-1342cfa41cfcCited by top-tier papers11
- TVConv: Efficient Translation Variant Convolution for Layout-aware Visual ProcessingJierun Chen, Tianlang He, Weipeng Zhuo, Li Ma et al.CVPR 2022 · 39 citations
- APG: Adaptive Parameter Generation Network for Click-Through Rate PredictionBencheng Yan, Pengjie Wang, Kai Zhang, Feng Li et al.NeurIPS 2022 · 36 citations
- BNUDC: A Two-Branched Deep Neural Network for Restoring Images from Under-Display CamerasJaihyun Koh, Jangho Lee, Sungroh YoonCVPR 2022 · 25 citations
- Improving Depth Completion via Depth Feature UpsamplingYufei Wang, Ge Zhang, Shaoqian Wang, Bo Li et al.CVPR 2024 · 15 citations
- Spatiotemporal Joint Filter Decomposition in 3D Convolutional Neural NetworksZichen Miao, Ze Wang, Xiuyuan Cheng, Qiang QiuNeurIPS 2021 · 12 citations
Builds on8
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Toward Real-World Single Image Super-Resolution: A New Benchmark and a New ModelJianrui Cai, Hui Zeng, Hongwei Yong, Zisheng Cao et al.ICCV 2019 · 713 citations
- Bayesian Loss for Crowd Count Estimation With Point SupervisionZhiheng Ma, Xing Wei, Xiaopeng Hong, Yihong GongICCV 2019 · 612 citations
- Adaptive Density Map Generation for Crowd CountingJia Wan, Antoni B. ChanICCV 2019 · 171 citations
- Learning Associative Inference Using Fast Weight MemoryImanol Schlag, Tsendsuren Munkhdalai, Jürgen SchmidhuberICLR 2021 · 64 citations
Related papers
- Dynamic Region-Aware ConvolutionJin Chen, Xijun Wang, Zichao Guo, Xiangyu Zhang et al.CVPR 2021
- Dynamic Convolution: Attention Over Convolution KernelsYinpeng Chen, Xiyang Dai, Mengchen Liu, Dongdong Chen et al.CVPR 2020
- Large Convolutional Model Tuning via Filter SubspaceWei Chen, Zichen Miao, Qiang QiuICLR 2025
- Decoupled Dynamic Filter NetworksJingkai Zhou, Varun Jampani, Zhixiong Pi, Qiong Liu et al.CVPR 2021
- Dynamic Multi-Scale Filters for Semantic SegmentationJunjun He, Zhongying Deng, Yu QiaoICCV 2019 · 287 citations
