Brain-inspired Lp-Convolution benefits large kernels and aligns better with visual cortex
Jea Kwon, Sungjun Lim, Kyungwoo Song, C. Justin Lee
Abstract
Convolutional Neural Networks (CNNs) have profoundly influenced the field of computer vision, drawing significant inspiration from the visual processing mechanisms inherent in the brain. Despite sharing fundamental structural and representational similarities with the biological visual system, differences in local connectivity patterns within CNNs open up an interesting area to explore. In this work, we explore whether integrating biologically observed connectivity patterns can enhance model performance and foster alignment with brain representations. We introduce a novel methodology, termed L p -convolution, which employs the multivariate p-generalized normal distribution (MPND). We took advantage of MPND's conformational flexibility to carefully bridge disparities between artificial and biological connectivity patterns by designing an adaptable L p -masks. L pmasks finds the optimal conformation through task-dependent adaptation such as distortion, scale, and rotation. This allows L p -convolution to perform well in tasks that require flexible input field shapes, including not only square-shape but also horizontal and vertical ones. Furthermore, we demonstrate that L p -convolution with biological constraint which we call Gaussian structured sparsity significantly enhances the performance of historically successful CNNs with large kernels. Lastly, we present that neural representations of CNNs aligns better with the visual cortex when the conformation of L p -masks is close to a Gaussian distribution, a biologicially closer condition. * This work was conducted at the Institute for Basic Science. † Co-corresponding authors. where Γ is gamma function, and det(•) denotes the determinant. 2 When we optimize for MPND, C is initialized with 1/σ 0 0 1/σ where σ determines the scale.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8bd2c9c4-e167-4351-a6c0-d675e184af28Builds on13
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
Related papers
- Deep Continuous NetworksNergis Tomen, Silvia-Laura Pintea, Jan van GemertICML 2021 · 15 citations
- Towards Biologically Plausible Convolutional NetworksRoman Pogodin, Yash Mehta, Timothy P. Lillicrap, Peter E. LathamNeurIPS 2021 · 30 citations
- Network DeconvolutionChengxi Ye, Matthew Evanusa, Hua He, Anton Mitrokhin et al.ICLR 2020
- Log-Polar Space Convolution LayersBing Su, Ji-Rong WenNeurIPS 2022 · 3 citations
- Emergence of Shape Bias in Convolutional Neural Networks through Activation SparsityTianqin Li, Ziqi Wen, Yangfan Li, Tai Sing LeeNeurIPS 2023 · 24 citations
