Flexible Context-Driven Sensory Processing in Dynamical Vision Models
Lakshmi Narasimhan Govindarajan, Abhiram Iyer, Valmiki Kothare, Ila Fiete
摘要
Visual representations become progressively more abstract along the cortical hierarchy. These abstract representations define notions like objects and shapes, but at the cost of spatial specificity. By contrast, low-level regions represent spatially local but simple input features. How do spatially non-specific representations of abstract concepts in high-level areas flexibly modulate the low-level sensory representations in appropriate ways to guide context-driven and goal-directed behaviors across a range of tasks? We build a biologically motivated and trainable neural network model of dynamics in the visual pathway, incorporating local, lateral, and feedforward synaptic connections, excitatory and inhibitory neurons, and long-range top-down inputs conceptualized as low-rank modulations of the input-driven sensory responses by high-level areas. We study this D ynamical C ortical net work ( DCnet ) in a visual cue-delay-search task and show that the model uses its own cue representations to adaptively modulate its perceptual responses to solve the task, outperforming state-of-the-art DNN vision and LLM models. The model’s population states over time shed light on the nature of contextual modulatory dynamics, generating predictions for experiments. We fine-tune the same model on classic psychophysics attention tasks, and find that the model closely replicates known reaction time results. This work represents a promising new foundation for understanding and making predictions about perturbations to visual processing in the brain.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Organizing recurrent network dynamics by task-computation to enable continual learningLea Duncker, Laura Driscoll, Krishna V. Shenoy, Maneesh Sahani 等NeurIPS 2020 · 被引用 108 次
- Disentangling neural mechanisms for perceptual groupingJunkyung Kim, Drew Linsley, Kalpit Thakkar, Thomas SerreICLR 2020 · 被引用 61 次
- Stable and expressive recurrent vision modelsDrew Linsley, Alekh Karkada Ashok, Lakshmi Narasimhan Govindarajan, Rex G. Liu 等NeurIPS 2020 · 被引用 56 次
相关 Paper
- Attention over Learned Object Embeddings Enables Complex Visual ReasoningDavid Ding, Felix Hill, Adam Santoro, Malcolm Reynolds 等NeurIPS 2021 · 被引用 87 次
- Cognitive Steering in Deep Neural Networks via Long-Range Modulatory Feedback ConnectionsTalia Konkle, George A. AlvarezNeurIPS 2023 · 被引用 22 次
- Long-Range Feedback Spiking Network Captures Dynamic and Static Representations of the Visual Cortex under Movie StimuliLiwei Huang, Zhengyu Ma, Liutao Yu, Huihui Zhou 等NeurIPS 2024 · 被引用 5 次
- CogReact: A Reinforced Framework to Model Human Cognitive Reaction Modulated by Dynamic InterventionSonglin Xu, Xinyu ZhangICML 2025
- DynaVieW: Schema-Guided World Modeling for Understanding Hierarchical Visual DynamicsSilin Gao, Hao Zhao, Zeming Chen, Sepideh Mamooler 等ICML 2026
