Network Deconvolution
Chengxi Ye, Matthew Evanusa, Hua He, Anton Mitrokhin, Tom Goldstein, James A. Yorke, Cornelia Fermüller, Yiannis Aloimonos
Abstract
Convolution is a central operation in Convolutional Neural Networks (CNNs), which applies a kernel to overlapping regions shifted across the image. However, because of the strong correlations in real-world image data, convolutional kernels are in effect re-learning redundant data. In this work, we show that this redundancy has made neural network training challenging, and propose network deconvolution, a procedure which optimally removes pixel-wise and channel-wise correlations before the data is fed into each layer. Network deconvolution can be efficiently calculated at a fraction of the computational cost of a convolution layer. We also show that the deconvolution filters in the first layer of the network resemble the center-surround structure found in biological neurons in the visual regions of the brain. Filtering with such kernels results in a sparse representation, a desired property that has been missing in the training of neural networks. Learning from the sparse representation promotes faster convergence and superior results without the use of batch normalization. We apply our network deconvolution operation to 10 modern neural network models by replacing batch normalization within each. Extensive experiments show that the network deconvolution operation is able to deliver performance improvement in all cases on the CIFAR-10, CIFAR-100, MNIST, Fashion-MNIST, Cityscapes, and ImageNet datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e1e2c670-9816-4e89-bcb4-2a057df7cf5dCited by top-tier papers3
- Convolutional Normalization: Improving Deep Convolutional Network Robustness and TrainingSheng Liu, Xiao Li, Yuexiang Zhai, Chong You et al.NeurIPS 2021 · 30 citations
- Exploiting Invariance in Training Deep Neural NetworksChengxi Ye, Xiong Zhou, Tristan McKinney, Yanfeng Liu et al.AAAI 2022 · 4 citations
- Controllable Feature Whitening for Hyperparameter-Free Bias MitigationYooshin Cho, Hanbyel Cho, Janghyeon Lee, Hyeong Gwon Hong et al.ICCV 2025 · 2 citations
Related papers
- Improving Generalization of Batch Whitening by Convolutional Unit OptimizationYooshin Cho, Hanbyel Cho, Youngsoo Kim, Junmo KimICCV 2021 · 3 citations
- Is normalization indispensable for training deep neural network?Jie Shao, Kai Hu, Changhu Wang, Xiangyang Xue et al.NeurIPS 2020 · 70 citations
- Deep Isometric Learning for Visual RecognitionHaozhi Qi, Chong You, Xiaolong Wang, Yi Ma et al.ICML 2020 · 57 citations
- Towards an Effective Orthogonal Dictionary Convolution StrategyYishi Li, Kunran Xu, Rui Lai, Lin GuAAAI 2022 · 4 citations
- Delving into Variance Transmission and Normalization: Shift of Average Gradient Makes the Network CollapseYuxiang Liu, Jidong Ge, Chuanyi Li, Jie GuiAAAI 2021 · 2 citations
