Learning V1 Simple Cells with Vector Representation of Local Content and Matrix Representation of Local Motion
Ruiqi Gao, Jianwen Xie, Siyuan Huang, Yufan Ren, Song-Chun Zhu, Ying Nian Wu
摘要
This paper proposes a representational model for image pairs such as consecutive video frames that are related by local pixel displacements, in the hope that the model may shed light on motion perception in primary visual cortex (V1). The model couples the following two components: (1) the vector representations of local contents of images and (2) the matrix representations of local pixel displacements caused by the relative motions between the agent and the objects in the 3D scene. When the image frame undergoes changes due to local pixel displacements, the vectors are multiplied by the matrices that represent the local displacements. Thus the vector representation is equivariant as it varies according to the local displacements. Our experiments show that our model can learn Gabor-like filter pairs of quadrature phases. The profiles of the learned filters match those of simple cells in Macaque V1. Moreover, we demonstrate that the model can learn to infer local motions in either a supervised or unsupervised manner. With such a simple model, we achieve competitive results on optical flow estimation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Learning from Pattern Completion: Self-supervised Controllable GenerationZhiqiang Chen, Guofan Fan, Jinying Gao, Lei Ma 等NeurIPS 2024 · 被引用 1 次
- Learning Neural Representation of Camera Pose with Matrix Representation of Pose Shift via View SynthesisYaxuan Zhu, Ruiqi Gao, Siyuan Huang, Song-Chun Zhu 等CVPR 2021
它引用的顶会 Paper2
- On Path Integration of Grid Cells: Group Representation and Isotropic ScalingRuiqi Gao, Jianwen Xie, Xue-Xin Wei, Song-Chun Zhu 等NeurIPS 2021 · 被引用 23 次
- Learning Neural Representation of Camera Pose with Matrix Representation of Pose Shift via View SynthesisYaxuan Zhu, Ruiqi Gao, Siyuan Huang, Song-Chun Zhu 等CVPR 2021
相关 Paper
- A polar prediction model for learning to represent visual transformationsPierre-Étienne H. Fiquet, Eero P. SimoncelliNeurIPS 2023 · 被引用 9 次
- Learning Fine-Grained Features for Pixel-wise Video CorrespondencesRui Li, Shenglong Zhou, Dong LiuICCV 2023 · 被引用 7 次
- Featurising Pixels from Dynamic 3D Scenes with Linear In-Context LearnersNikita Araslanov, Martin Sundermeyer, Hidenobu Matsuki, David Joseph Tan 等CVPR 2026
- CroCo: Self-Supervised Pre-training for 3D Vision Tasks by Cross-View CompletionPhilippe Weinzaepfel, Vincent Leroy, Thomas Lucas, Romain Brégier 等NeurIPS 2022 · 被引用 189 次
- Object Concepts Emerge from MotionHaoqian Liang, Xiaohui Wang, Zhichao Li, Ya Yang 等NeurIPS 2025
