Towards robust vision by multi-task learning on monkey visual cortex
Shahd Safarani, Arne Nix, Konstantin Willeke, Santiago A. Cadena, Kelli Restivo, George H. Denfield, Andreas S. Tolias, Fabian H. Sinz
Abstract
Deep neural networks set the state-of-the-art across many tasks in computer vision, but their generalization ability to simple image distortions is surprisingly fragile. In contrast, the mammalian visual system is robust to a wide range of perturbations. Recent work suggests that this generalization ability can be explained by useful inductive biases encoded in the representations of visual stimuli throughout the visual cortex. Here, we successfully leveraged these inductive biases with a multitask learning approach: we jointly trained a deep network to perform image classification and to predict neural activity in macaque primary visual cortex (V1) in response to the same natural stimuli. We measured the out-of-distribution generalization abilities of our resulting network by testing its robustness to common image distortions. We found that co-training on monkey V1 data indeed leads to increased robustness despite the absence of those distortions during training. Additionally, we showed that our network's robustness is often very close to that of an Oracle network where parts of the architecture are directly trained on noisy images. Our results also demonstrated that the network's representations become more brain-like as their robustness improves. Using a novel constrained reconstruction analysis, we investigated what makes our brain-regularized network more robust. We found that our monkey co-trained network is more sensitive to content than noise when compared to a Baseline network that we trained for image classification alone. Using DeepGaze-predicted saliency maps for ImageNet images, we found that the monkey co-trained network tends to be more sensitive to salient regions in a scene, reminiscent of existing theories on the role of V1 in the detection of object borders and bottom-up saliency. Overall, our work expands the promising research avenue of transferring inductive biases from biological to artificial neural networks on the representational level, and provides a novel analysis of the effects of our transfer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- Robust Heterogeneous Federated Learning under Data CorruptionXiuwen Fang, Mang Ye, Xiyuan YangICCV 2023 · 44 citations
- Aligning Model and Macaque Inferior Temporal Cortex Representations Improves Model-to-Human Behavioral Alignment and Adversarial RobustnessJoel Dapello, Kohitij Kar, Martin Schrimpf, Robert Baldwin Geary et al.ICLR 2023 · 27 citations
- LCANets: Lateral Competition Improves Robustness Against Corruption and AttackMichael A. Teti, Garrett T. Kenyon, Ben Migliori, Juston MooreICML 2022 · 22 citations
- Explaining V1 Properties with a Biologically Constrained Deep Learning ArchitectureGalen Pogoncheff, Jacob Granley, Michael BeyelerNeurIPS 2023 · 17 citations
- Characterizing the Ventral Visual Stream with Response-Optimized Neural Encoding ModelsMeenakshi Khosla, Keith Jamison, Amy Kuceyeski, Mert R. SabuncuNeurIPS 2022 · 16 citations
Builds on4
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Simulating a Primary Visual Cortex at the Front of CNNs Improves Robustness to Image PerturbationsJoel Dapello, Tiago Marques, Martin Schrimpf, Franziska Geiger et al.NeurIPS 2020 · 250 citations
- Beyond accuracy: quantifying trial-by-trial behaviour of CNNs and humans by measuring error consistencyRobert Geirhos, Kristof Meding, Felix A. WichmannNeurIPS 2020 · 154 citations
- Generalization in data-driven models of primary visual cortexKonstantin-Klemens Lurz, Mohammad Bashiri, Konstantin Willeke, Akshay Kumar Jagadish et al.ICLR 2021 · 71 citations
Related papers
- Brain-like representational straightening of natural movies in robust feedforward neural networksTahereh Toosi, Elias B. IssaICLR 2023 · 4 citations
- Modeling Biological Immunity to Adversarial ExamplesEdward Kim, Jocelyn Rego, Yijing Watkins, Garrett T. KenyonCVPR 2020
- Adversarially trained neural representations are already as robust as biological neural representationsChong Guo, Michael J. Lee, Guillaume Leclerc, Joel Dapello et al.ICML 2022 · 31 citations
- Anatomically inspired digital twins capture hierarchical object representations in visual cortexEmanuele Luconi, Dario Liscai, Carlo Baldassi, Alessandro Marin Vargas et al.NeurIPS 2025 · 1 citation
- Counterfactual Generative NetworksAxel Sauer, Andreas GeigerICLR 2021 · 145 citations
