Gated Channel Transformation for Visual Recognition
Zongxin Yang, Linchao Zhu, Yu Wu, Yi Yang
摘要
In this work, we propose a generally applicable transformation unit for visual recognition with deep convolutional neural networks. This transformation explicitly models channel relationships with explainable control variables. These variables determine the neuron behaviors of competition or cooperation, and they are jointly optimized with convolutional weights towards more accurate recognition. In Squeeze-and-Excitation (SE) Networks, the channel relationships are implicitly learned by fully connected layers, and the SE block is integrated at the block-level. We instead introduce a channel normalization layer to reduce the number of parameters and computational complexity. This lightweight layer incorporates a simple l 2 normalization, enabling our transformation unit applicable to operator-level without much increase of additional parameters. Extensive experiments demonstrate the effectiveness of our unit with clear margins on many vision tasks, i.e., image classification on ImageNet, object detection and instance segmentation on COCO, video classification on Kinetics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- SimAM: A Simple, Parameter-Free Attention Module for Convolutional Neural NetworksLingxiao Yang, Ru-Yuan Zhang, Lida Li, Xiaohua XieICML 2021 · 被引用 1,593 次
- Deep Digging into the Generalization of Self-Supervised Monocular Depth EstimationJinwoo Bae, Sungho Moon, Sunghoon ImAAAI 2023 · 被引用 127 次
- Reliable Propagation-Correction Modulation for Video Object SegmentationXiaohao Xu, Jinglu Wang, Xiao Li, Yan LuAAAI 2022 · 被引用 74 次
- From Contexts to Locality: Ultra-high Resolution Image Segmentation via Locality-aware Contextual CorrelationQi Li, Weixiang Yang, Wenxi Liu, Yuanlong Yu 等ICCV 2021 · 被引用 55 次
- MIGC: Multi-Instance Generation Controller for Text-to-Image SynthesisDewei Zhou, You Li, Fan Ma, Xiaoting Zhang 等CVPR 2024 · 被引用 52 次
相关 Paper
- Linear Context Transform BlockDongsheng Ruan, Jun Wen, Nenggan Zheng, Min ZhengAAAI 2020 · 被引用 26 次
- Channel Equilibrium Networks for Learning Deep RepresentationWenqi Shao, Shitao Tang, Xingang Pan, Ping Tan 等ICML 2020 · 被引用 17 次
- SRM: A Style-Based Recalibration Module for Convolutional Neural NetworksHyunJae Lee, Hyo-Eun Kim, Hyeonseob NamICCV 2019 · 被引用 286 次
- Gaussian Context TransformerDongsheng Ruan, Daiyin Wang, Yuan Zheng, Nenggan Zheng 等CVPR 2021
- Improving Convolutional Networks With Self-Calibrated ConvolutionsJiang-Jiang Liu, Qibin Hou, Ming-Ming Cheng, Changhu Wang 等CVPR 2020
