Stable and expressive recurrent vision models
Drew Linsley, Alekh Karkada Ashok, Lakshmi Narasimhan Govindarajan, Rex G. Liu, Thomas Serre
Abstract
Primate vision depends on recurrent processing for reliable perception (Gilbert & Li, 2013). At the same time, there is a growing body of literature demonstrating that recurrent connections improve the learning efficiency and generalization of vision models on classic computer vision challenges. Why then, are current large-scale challenges dominated by feedforward networks? We posit that the effectiveness of recurrent vision models is bottlenecked by the widespread algorithm used for training them, "back-propagation through time" (BPTT), which has O(N) memory-complexity for training an N step model. Thus, recurrent vision model design is bounded by memory constraints, forcing a choice between rivaling the enormous capacity of leading feedforward models or trying to compensate for this deficit through granular and complex dynamics. Here, we develop a new learning algorithm, "contractor recurrent back-propagation" (C-RBP), which alleviates these issues by achieving constant O(1) memory-complexity with steps of recurrent processing. We demonstrate that recurrent vision models trained with C-RBP can detect long-range spatial dependencies in a synthetic contour tracing task that BPTT-trained models cannot. We further demonstrate that recurrent vision models trained with C-RBP to solve the large-scale Panoptic Segmentation MS-COCO challenge outperform the leading feedforward approach. C-RBP is a general-purpose learning algorithm for any application that can benefit from expansive recurrent dynamics. Code and data are available at this https URL.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f1b6ae96-220f-4915-99c4-917ce7d8b9f5Cited by top-tier papers14
- Monotone operator equilibrium networksEzra Winston, J. Zico KolterNeurIPS 2020 · 177 citations
- Learning Physical Graph Representations from Visual ScenesDaniel Bear, Chaofei Fan, Damian Mrowca, Yunzhu Li et al.NeurIPS 2020 · 88 citations
- Stabilizing Equilibrium Models by Jacobian RegularizationShaojie Bai, Vladlen Koltun, J. Zico KolterICML 2021 · 80 citations
- Predify: Augmenting deep neural networks with brain-inspired predictive coding dynamicsBhavin Choksi, Milad Mozafari, Callum Biggs O'May, Benjamin Ador et al.NeurIPS 2021 · 48 citations
- The least-control principle for local learning at equilibriumAlexander Meulemans, Nicolas Zucchet, Seijin Kobayashi, Johannes von Oswald et al.NeurIPS 2022 · 32 citations
Builds on2
Related papers
- Training Recurrent Neural Networks via Forward Propagation Through TimeAnil Kag, Venkatesh SaligramaICML 2021 · 48 citations
- Training Recurrent Neural Networks Online by Learning Explicit State VariablesSomjit Nath, Vincent Liu, Alan Chan, Xin Li et al.ICLR 2020 · 9 citations
- Global Credit Assignment via Dynamical CriticalityWentao Wang, Keren Gao, Guozhang ChenICML 2026
- Orderless Recurrent Models for Multi-Label ClassificationVacit Oguz Yazici, Abel Gonzalez-Garcia, Arnau Ramisa, Bartlomiej Twardowski et al.CVPR 2020
- RTify: Aligning Deep Neural Networks with Human Behavioral DecisionsYu-Ang Cheng, Ivan F. Rodriguez Rodriguez, Sixuan Chen, Kohitij Kar et al.NeurIPS 2024 · 11 citations
