Learning low-dimensional generalizable natural features from retina using a U-net
Siwei Wang, Benjamin Hoshal, Elizabeth de Laittre, Thierry Mora, Michael Berry, Stephanie E. Palmer
摘要
Much of sensory neuroscience focuses on presenting stimuli that are chosen by the experimenter because they are parametric and easy to sample and are thought to be behaviorally relevant to the organism. However, it is not generally known what these relevant features are in complex, natural scenes. This work focuses on using the retinal encoding of natural movies to determine the presumably behaviorally-relevant features that the brain represents. It is prohibitive to parameterize a natural movie and its respective retinal encoding fully. We use time within a natural movie as a proxy for the whole suite of features evolving across the scene. We then use a task-agnostic deep architecture, an encoder-decoder, to model the retinal encoding process and characterize its representation of “time in the natural scene” in a compressed latent space. In our end-to-end training, an encoder learns a compressed latent representation from a large population of salamander retinal ganglion cells responding to natural movies, while a decoder samples from this compressed latent space to generate the appropriate future movie frame. By comparing latent representations of retinal activity from three movies, we find that the retina has a generalizable encoding for time in the natural scene: the precise, low-dimensional representation of time learned from one movie can be used to represent time in a different movie, with up to 17 ms resolution. We then show that static textures and velocity features of a natural movie are synergistic. The retina simultaneously encodes both to establishes a generalizable, low-dimensional representation of time in the natural scene.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum 等ICLR 2021 · 被引用 381 次
- Learning identifiable and interpretable latent models of high-dimensional neural activity using pi-VAEDing Zhou, Xue-Xin WeiNeurIPS 2020 · 被引用 110 次
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 被引用 59 次
- Drop, Swap, and Generate: A Self-Supervised Approach for Generating Neural ActivityRan Liu, Mehdi Azabou, Max Dabagia, Chi-Heng Lin 等NeurIPS 2021 · 被引用 49 次
相关 Paper
- Information Geometry of the Retinal Representation ManifoldXuehao Ding, Dongsoo Lee, Joshua Melander, George Sivulka 等NeurIPS 2023 · 被引用 17 次
- Long-Range Feedback Spiking Network Captures Dynamic and Static Representations of the Visual Cortex under Movie StimuliLiwei Huang, Zhengyu Ma, Liutao Yu, Huihui Zhou 等NeurIPS 2024 · 被引用 5 次
- System Identification with Biophysical Constraints: A Circuit Model of the Inner RetinaCornelius Schröder, David A. Klindt, Sarah Strauß, Katrin Franke 等NeurIPS 2020 · 被引用 15 次
- Maximum a posteriori natural scene reconstruction from retinal ganglion cells with deep denoiser priorsEric Wu, Nora Brackbill, Alexander Sher, Alan M. Litke 等NeurIPS 2022 · 被引用 9 次
- A polar prediction model for learning to represent visual transformationsPierre-Étienne H. Fiquet, Eero P. SimoncelliNeurIPS 2023 · 被引用 9 次
