Disentangling neural mechanisms for perceptual grouping
Junkyung Kim, Drew Linsley, Kalpit Thakkar, Thomas Serre
摘要
Forming perceptual groups and individuating objects in visual scenes is an essential step towards visual intelligence. This ability is thought to arise in the brain from computations implemented by bottom-up, horizontal, and top-down connections between neurons. However, the relative contributions of these connections to perceptual grouping are poorly understood. We address this question by systematically evaluating neural network architectures featuring combinations of these connections on two synthetic visual tasks, which stress low-level `gestalt' vs. high-level object cues for perceptual grouping. We show that increasing the difficulty of either task strains learning for networks that rely solely on bottom-up processing. Horizontal connections resolve this limitation on tasks with gestalt cues by supporting incremental spatial propagation of activities, whereas top-down connections rescue learning on tasks featuring object cues by propagating coarse predictions about the position of the target object. Our findings disassociate the computational roles of bottom-up, horizontal and top-down connectivity, and demonstrate how a model featuring all of these interactions can more flexibly learn to form perceptual groups.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Long Range Arena : A Benchmark for Efficient TransformersYi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen 等ICLR 2021 · 被引用 881 次
- Harmonizing the object recognition strategies of deep neural networks with humansThomas Fel, Ivan F. Rodriguez Rodriguez, Drew Linsley, Thomas SerreNeurIPS 2022 · 被引用 111 次
- Learning Physical Graph Representations from Visual ScenesDaniel Bear, Chaofei Fan, Damian Mrowca, Yunzhu Li 等NeurIPS 2020 · 被引用 88 次
- Stable and expressive recurrent vision modelsDrew Linsley, Alekh Karkada Ashok, Lakshmi Narasimhan Govindarajan, Rex G. Liu 等NeurIPS 2020 · 被引用 56 次
- Never Train from Scratch: Fair Comparison of Long-Sequence Models Requires Data-Driven PriorsIdo Amos, Jonathan Berant, Ankit GuptaICLR 2024 · 被引用 39 次
它引用的顶会 Paper1
相关 Paper
- Flexible Context-Driven Sensory Processing in Dynamical Vision ModelsLakshmi Narasimhan Govindarajan, Abhiram Iyer, Valmiki Kothare, Ila FieteNeurIPS 2024 · 被引用 1 次
- Psychologically-inspired, unsupervised inference of perceptual groups of GUI widgets from GUI imagesMulong Xie, Zhenchang Xing, Sidong Feng, Xiwei Xu 等FSE 2022 · 被引用 30 次
- GUST: Combinatorial Generalization by Unsupervised Grouping with Neuronal CoherenceHao Zheng, Hui Lin, Rong ZhaoNeurIPS 2023 · 被引用 3 次
- Brain-like Flexible Visual Inference by Harnessing Feedback Feedforward AlignmentTahereh Toosi, Elias B. IssaNeurIPS 2023 · 被引用 5 次
- Bidirectional Predictive CodingGaspard Oliviers, Mufeng Tang, Rafal BogaczICLR 2026 · 被引用 7 次
