Predictive Coding beyond Gaussian Distributions
Luca Pinchetti, Tommaso Salvatori, Yordan Yordanov, Beren Millidge, Yuhang Song, Thomas Lukasiewicz
摘要
A large amount of recent research has the far-reaching goal of finding training methods for deep neural networks that can serve as alternatives to backpropagation (BP). A prominent example is predictive coding (PC), which is a neuroscience-inspired method that performs inference on hierarchical Gaussian generative models. These methods, however, fail to keep up with modern neural networks, as they are unable to replicate the dynamics of complex layers and activation functions. In this work, we solve this problem by generalizing PC to arbitrary probability distributions, enabling the training of architectures, such as transformers, that are hard to approximate with only Gaussian assumptions. We perform three experimental analyses. First, we study the gap between our method and the standard formulation of PC on multiple toy examples. Second, we test the reconstruction quality on variational autoencoders, where our method reaches the same reconstruction quality as BP. Third, we show that our method allows us to train transformer networks and achieve a performance comparable with BP on conditional language models. More broadly, this method allows neuroscience-inspired learning to be applied to multiple domains, since the internal distributions can be flexibly adapted to the data, tasks, and architectures used.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?Francesco Innocenti, El Mehdi Achour, Ryan Singh, Christopher L. BuckleyNeurIPS 2024 · 被引用 10 次
- Divide-and-Conquer Predictive Coding: a structured Bayesian inference algorithmEli Sennesh, Hao Wu, Tommaso SalvatoriNeurIPS 2024 · 被引用 7 次
- Predictive Coding beyond CorrelationsTommaso Salvatori, Luca Pinchetti, Amine M'Charrak, Beren Millidge 等ICML 2024 · 被引用 6 次
- Neural Sampling in Hierarchical Exponential-family Energy-based ModelsXingsi Dong, Si WuNeurIPS 2023 · 被引用 5 次
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Can the Brain Do Backpropagation? - Exact Implementation of Backpropagation in Predictive Coding NetworksYuhang Song, Thomas Lukasiewicz, Zhenghua Xu, Rafal BogaczNeurIPS 2020 · 被引用 117 次
- A Theoretical Framework for Target PropagationAlexander Meulemans, Francesco S. Carzaniga, Johan A. K. Suykens, João Sacramento 等NeurIPS 2020 · 被引用 110 次
- Associative Memories via Predictive CodingTommaso Salvatori, Yuhang Song, Yujian Hong, Lei Sha 等NeurIPS 2021 · 被引用 84 次
相关 Paper
- On the Infinite Width and Depth Limits of Predictive Coding NetworksFrancesco Innocenti, El Mehdi Achour, Rafal BogaczICML 2026
- A Theoretical Framework for Inference and Learning in Predictive Coding NetworksBeren Millidge, Yuhang Song, Tommaso Salvatori, Thomas Lukasiewicz 等ICLR 2023 · 被引用 8 次
- Learning on Arbitrary Graph Topologies via Predictive CodingTommaso Salvatori, Luca Pinchetti, Beren Millidge, Yuhang Song 等NeurIPS 2022 · 被引用 56 次
- A Stable, Fast, and Fully Automatic Learning Algorithm for Predictive Coding NetworksTommaso Salvatori, Yuhang Song, Yordan Yordanov, Beren Millidge 等ICLR 2024 · 被引用 22 次
- ePC: Fast and Deep Predictive Coding in Digital SimulationCédric Goemaere, Gaspard Oliviers, Rafal Bogacz, Thomas DemeesterICML 2026 · 被引用 3 次
