Channel Equilibrium Networks for Learning Deep Representation
Wenqi Shao, Shitao Tang, Xingang Pan, Ping Tan, Xiaogang Wang, Ping Luo
Abstract
Convolutional Neural Networks (CNNs) are typically constructed by stacking multiple building blocks, each of which contains a normalization layer such as batch normalization (BN) and a rectified linear function such as ReLU. However, this work shows that the combination of normalization and rectified linear function leads to inhibited channels, which have small magnitude and contribute little to the learned feature representation, impeding the generalization ability of CNNs. Unlike prior arts that simply removed the inhibited channels, we propose to "wake them up" during training by designing a novel neural building block, termed Channel Equilibrium (CE) block, which enables channels at the same layer to contribute equally to the learned representation. We show that CE is able to prevent inhibited channels both empirically and theoretically. CE has several appealing benefits. (1) It can be integrated into many advanced CNN architectures such as ResNet and MobileNet, outperforming their original networks. (2) CE has an interesting connection with the Nash Equilibrium, a well-known solution of a non-cooperative game. (3) Extensive experiments show that CE achieves state-of-the-art performance on various challenging benchmarks such as ImageNet and COCO.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9281852e-7337-4471-ab39-7dbdeaeb69c9Cited by top-tier papers5
- Deep Multimodal Fusion by Channel ExchangingYikai Wang, Wenbing Huang, Fuchun Sun, Tingyang Xu et al.NeurIPS 2020 · 321 citations
- Quadtree Attention for Vision TransformersShitao Tang, Jiahui Zhang, Siyu Zhu, Ping TanICLR 2022 · 194 citations
- OGP-Net: Optical Guidance Meets Pixel-Level Contrastive Distillation for Robust Multi-Modal and Missing Modality SegmentationAniruddh Sikdar, Jayant Teotia, Suresh SundaramAAAI 2025 · 8 citations
- Channel Regeneration: Improving Channel Utilization for Compact DNNsAnkit Kumar Sharma, Hassan ForooshAAAI 2023
- Group Whitening: Balancing Learning Efficiency and Representational CapacityLei Huang, Yi Zhou, Li Liu, Fan Zhu et al.CVPR 2021
Builds on1
Related papers
- Gated Channel Transformation for Visual RecognitionZongxin Yang, Linchao Zhu, Yu Wu, Yi YangCVPR 2020
- Filter Response Normalization Layer: Eliminating Batch Dependence in the Training of Deep Neural NetworksSaurabh Singh, Shankar KrishnanCVPR 2020
- Delving into the Estimation Shift of Batch Normalization in a NetworkLei Huang, Yi Zhou, Tian Wang, Jie Luo et al.CVPR 2022 · 25 citations
- PatchUp: A Feature-Space Block-Level Regularization Technique for Convolutional Neural NetworksMojtaba Faramarzi, Mohammad Amini, Akilesh Badrinaaraayanan, Vikas Verma et al.AAAI 2022 · 40 citations
- Tied Block Convolution: Leaner and Better CNNs with Shared Thinner FiltersXudong Wang, Stella X. YuAAAI 2021 · 50 citations
