Hierarchical VAEs provide a normative account of motion processing in the primate brain
Hadi Vafaii, Jacob L. Yates, Daniel Butts
Abstract
The relationship between perception and inference, as postulated by Helmholtz in the 19th century, is paralleled in modern machine learning by generative models like Variational Autoencoders (VAEs) and their hierarchical variants. Here, we evaluate the role of hierarchical inference and its alignment with brain function in the domain of motion perception. We first introduce a novel synthetic data framework, Retinal Optic Flow Learning (ROFL), which enables control over motion statistics and their causes. We then present a new hierarchical VAE and test it against alternative models on two downstream tasks: (i) predicting ground truth causes of retinal optic flow (e.g., self-motion); and (ii) predicting the responses of neurons in the motion processing pathway of primates. We manipulate the model architectures (hierarchical versus non-hierarchical), loss functions, and the causal structure of the motion stimuli. We find that hierarchical latent structure in the model leads to several improvements. First, it improves the linear decodability of ground truth factors and does so in a sparse and disentangled manner. Second, our hierarchical VAE outperforms previous state-of-the-art models in predicting neuronal responses and exhibits sparse latent-to-neuron relationships. These results depend on the causal structure of the world, indicating that alignment between brains and artificial neural networks depends not only on architecture but also on matching ecologically relevant stimulus statistics. Taken together, our results suggest that hierarchical Bayesian inference underlines the brain’s understanding of the world, and hierarchical VAEs can effectively model this understanding.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9c0f55b5-a0f7-411b-8952-239c33d458f6Cited by top-tier papers3
- Poisson Variational AutoencoderHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2024 · 18 citations
- Brain-like Variational InferenceHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2025 · 7 citations
- Negative Binomial Variational Autoencoders for Overdispersed Latent ModelingYixuan Zhang, Jinhao Sheng, Wenxin Zhang, Quyu Kong et al.CVPR 2026 · 3 citations
Builds on13
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Generalized Shape Metrics on Neural RepresentationsAlex H. Williams, Erin Kunz, Simon Kornblith, Scott W. LindermanNeurIPS 2021 · 182 citations
- Your head is there to move you around: Goal-driven models of the primate dorsal pathwayPatrick J. Mineault, Shahab Bakhtiari, Blake A. Richards, Christopher C. PackNeurIPS 2021 · 61 citations
- Neural Networks with Recurrent Generative FeedbackYujia Huang, James Gornet, Sihui Dai, Zhiding Yu et al.NeurIPS 2020 · 48 citations
- Synergies between Disentanglement and Sparsity: Generalization and Identifiability in Multi-Task LearningSébastien Lachapelle, Tristan Deleu, Divyat Mahajan, Ioannis Mitliagkas et al.ICML 2023 · 46 citations
Related papers
- Variational Predictive Routing with Nested Subjective TimescalesAlexey Zakharov, Qinghai Guo, Zafeirios FountasICLR 2022 · 12 citations
- Deep Hierarchical Video CompressionMing Lu, Zhihao Duan, Fengqing Zhu, Zhan MaAAAI 2024 · 19 citations
- Learning Temporally Causal Latent Processes from General Temporal DataWeiran Yao, Yuewen Sun, Alex Ho, Changyin Sun et al.ICLR 2022 · 108 citations
- Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational AutoencodersHien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat HoICLR 2024 · 8 citations
- EVOKE: Efficient and High-Fidelity EEG-to-Video Reconstruction via Decoupling Implicit Neural RepresentationHaodong Jing, Panqi Yang, Dongyao Jiang, Zhipeng Liu et al.AAAI 2026 · 1 citation
