Learning low-dimensional generalizable natural features from retina using a U-net
Siwei Wang, Benjamin Hoshal, Elizabeth de Laittre, Thierry Mora, Michael Berry, Stephanie E. Palmer
Abstract
Much of sensory neuroscience focuses on presenting stimuli that are chosen by the experimenter because they are parametric and easy to sample and are thought to be behaviorally relevant to the organism. However, it is not generally known what these relevant features are in complex, natural scenes. This work focuses on using the retinal encoding of natural movies to determine the presumably behaviorally-relevant features that the brain represents. It is prohibitive to parameterize a natural movie and its respective retinal encoding fully. We use time within a natural movie as a proxy for the whole suite of features evolving across the scene. We then use a task-agnostic deep architecture, an encoder-decoder, to model the retinal encoding process and characterize its representation of “time in the natural scene” in a compressed latent space. In our end-to-end training, an encoder learns a compressed latent representation from a large population of salamander retinal ganglion cells responding to natural movies, while a decoder samples from this compressed latent space to generate the appropriate future movie frame. By comparing latent representations of retinal activity from three movies, we find that the retina has a generalizable encoding for time in the natural scene: the precise, low-dimensional representation of time learned from one movie can be used to represent time in a different movie, with up to 17 ms resolution. We then show that static textures and velocity features of a natural movie are synergistic. The retina simultaneously encodes both to establishes a generalizable, low-dimensional representation of time in the natural scene.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0bffa3e6-d6f0-434a-89f6-7c7e599c5550Cited by top-tier papers1
Ask how each one uses itBuilds on4
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- Learning identifiable and interpretable latent models of high-dimensional neural activity using pi-VAEDing Zhou, Xue-Xin WeiNeurIPS 2020 · 110 citations
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 59 citations
- Drop, Swap, and Generate: A Self-Supervised Approach for Generating Neural ActivityRan Liu, Mehdi Azabou, Max Dabagia, Chi-Heng Lin et al.NeurIPS 2021 · 49 citations
Related papers
- Information Geometry of the Retinal Representation ManifoldXuehao Ding, Dongsoo Lee, Joshua Melander, George Sivulka et al.NeurIPS 2023 · 17 citations
- Long-Range Feedback Spiking Network Captures Dynamic and Static Representations of the Visual Cortex under Movie StimuliLiwei Huang, Zhengyu Ma, Liutao Yu, Huihui Zhou et al.NeurIPS 2024 · 5 citations
- System Identification with Biophysical Constraints: A Circuit Model of the Inner RetinaCornelius Schröder, David A. Klindt, Sarah Strauß, Katrin Franke et al.NeurIPS 2020 · 15 citations
- Maximum a posteriori natural scene reconstruction from retinal ganglion cells with deep denoiser priorsEric Wu, Nora Brackbill, Alexander Sher, Alan M. Litke et al.NeurIPS 2022 · 9 citations
- A polar prediction model for learning to represent visual transformationsPierre-Étienne H. Fiquet, Eero P. SimoncelliNeurIPS 2023 · 9 citations
