A polar prediction model for learning to represent visual transformations
Pierre-Étienne H. Fiquet, Eero P. Simoncelli
摘要
All organisms make temporal predictions, and their evolutionary fitness level depends on the accuracy of these predictions. In the context of visual perception, the motions of both the observer and objects in the scene structure the dynamics of sensory signals, allowing for partial prediction of future signals based on past ones. Here, we propose a self-supervised representation-learning framework that extracts and exploits the regularities of natural videos to compute accurate predictions. We motivate the polar architecture by appealing to the Fourier shift theorem and its group-theoretic generalization, and we optimize its parameters on next-frame prediction. Through controlled experiments, we demonstrate that this approach can discover the representation of simple transformation groups acting in data. When trained on natural video datasets, our framework achieves better prediction performance than traditional motion compensation and rivals conventional deep networks, while maintaining interpretability and speed. Furthermore, the polar computations can be restructured into components resembling normalized simple and direction-selective complex cell models of primate V1 neurons. Thus, polar prediction offers a principled framework for understanding how the visual system represents sensory inputs in a form that simplifies temporal prediction.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Pre-trained Large Language Models Use Fourier Features to Compute AdditionTianyi Zhou, Deqing Fu, Vatsal Sharan, Robin JiaNeurIPS 2024 · 被引用 48 次
- FoNE: Precise Single-Token Number Embeddings via Fourier FeaturesTianyi Zhou, Deqing Fu, Mahdi Soltanolkotabi, Robin Jia 等ICLR 2026 · 被引用 24 次
- Poisson Variational AutoencoderHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2024 · 被引用 18 次
- Learning predictable and robust neural representations by straightening image sequencesXueyan Niu, Cristina Savin, Eero P. SimoncelliNeurIPS 2024 · 被引用 13 次
- Brain-like Variational InferenceHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2025 · 被引用 7 次
它引用的顶会 Paper3
- Forecasting Sequential Data Using Consistent Koopman AutoencodersOmri Azencot, N. Benjamin Erichson, Vanessa Lin, Michael W. MahoneyICML 2020 · 被引用 203 次
- Bispectral Neural NetworksSophia Sanborn, Christian Shewmake, Bruno A. Olshausen, Christopher J. HillarICLR 2023 · 被引用 78 次
- Biological Learning of Irreducible Representations of Commuting TransformationsAlexander Genkin, David Lipshutz, Siavash Golkar, Tiberiu Tesileanu 等NeurIPS 2022 · 被引用 5 次
相关 Paper
- Learning V1 Simple Cells with Vector Representation of Local Content and Matrix Representation of Local MotionRuiqi Gao, Jianwen Xie, Siyuan Huang, Yufan Ren 等AAAI 2022 · 被引用 2 次
- VCT: A Video Compression TransformerFabian Mentzer, George Toderici, David Minnen, Sergi Caelles 等NeurIPS 2022 · 被引用 155 次
- Video Playback Rate Perception for Self-Supervised Spatio-Temporal Representation LearningYuan Yao, Chang Liu, Dezhao Luo, Yu Zhou 等CVPR 2020
- Learning the Predictability of the FutureDidac Suris, Ruoshi Liu, Carl VondrickCVPR 2021
- Self-Supervised Representation Learning from Flow EquivarianceYuwen Xiong, Mengye Ren, Wenyuan Zeng, Raquel Urtasun WaabiICCV 2021 · 被引用 32 次
