Convolution Goes Higher-Order: A Biologically Inspired Mechanism Empowers Image Classification
Simone Azeglio, Olivier Marre, Peter Neri, Ulisse Ferrari
Abstract
We propose a novel approach to image classification inspired by complex nonlinear biological visual processing, whereby classical convolutional neural networks (CNNs) are equipped with learnable higher-order convolutions. Our model incorporates a Volterra-like expansion of the convolution operator, capturing multiplicative interactions akin to those observed in early and advanced stages of biological visual processing. We evaluated this approach on synthetic datasets by measuring sensitivity to testing higher-order correlations and performance in standard benchmarks (MNIST, FashionMNIST, CIFAR10, CIFAR100 and Imagenette). Our architecture outperforms traditional CNN baselines, and achieves optimal performance with expansions up to 3rd/4th order, aligning remarkably well with the distribution of pixel intensities in natural images. Through systematic perturbation analysis, we validate this alignment by isolating the contributions of specific image statistics to model performance, demonstrating how different orders of convolution process distinct aspects of visual information. Furthermore, Representational Similarity Analysis reveals distinct geometries across network layers, indicating qualitatively different modes of visual information processing. Our work bridges neuroscience and deep learning, offering a path towards more effective, biologically inspired computer vision models. It provides insights into visual information processing and lays the groundwork for neural networks that better capture complex visual patterns, particularly in resource-constrained scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3d682bf7-1dee-4752-a4d8-60dd0916e9d5Cited by top-tier papers1
Ask how each one uses itBuilds on4
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Simulating a Primary Visual Cortex at the Front of CNNs Improves Robustness to Image PerturbationsJoel Dapello, Tiago Marques, Martin Schrimpf, Franziska Geiger et al.NeurIPS 2020 · 250 citations
- Grounding Representation Similarity Through Statistical TestingFrances Ding, Jean-Stanislas Denain, Jacob SteinhardtNeurIPS 2021 · 88 citations
Related papers
- Conquering the CNN Over-Parameterization Dilemma: A Volterra Filtering Approach for Action RecognitionSiddharth Roheda, Hamid KrimAAAI 2020 · 12 citations
- MR-VNet: Media Restoration using Volterra NetworksSiddharth Roheda, Amit Satish Unde, Loay RashidCVPR 2024 · 3 citations
- Prune and distill: similar reformatting of image information along rat visual cortex and deep neural networksPaolo Muratore, Sina Tafazoli, Eugenio Piasini, Alessandro Laio et al.NeurIPS 2022 · 11 citations
- Deep Spiking Neural Networks with High Representation Similarity Model Visual Pathways of Macaque and MouseLiwei Huang, Zhengyu Ma, Liutao Yu, Huihui Zhou et al.AAAI 2023 · 15 citations
- Neural Regression, Representational Similarity, Model Zoology & Neural Taskonomy at Scale in Rodent Visual CortexColin Conwell, David Mayo, Andrei Barbu, Michael A. Buice et al.NeurIPS 2021 · 31 citations
