μPC: Scaling Predictive Coding to 100+ Layer Networks
Francesco Innocenti, El Mehdi Achour, Christopher L. Buckley
摘要
The biological implausibility of backpropagation (BP) has motivated many alternative, brain-inspired algorithms that attempt to rely only on local information, such as predictive coding (PC) and equilibrium propagation. However, these algorithms have notoriously struggled to train very deep networks, preventing them from competing with BP in large-scale settings. Indeed, scaling PC networks (PCNs) has recently been posed as a challenge for the community (Pinchetti et al., 2024). Here, we show that 100+ layer PCNs can be trained reliably using a Depth-P parameterisation (Yang et al., 2023; Bordelon et al., 2023) which we call"PC". By analysing the scaling behaviour of PCNs, we reveal several pathologies that make standard PCNs difficult to train at large depths. We then show that, despite addressing only some of these instabilities, PC allows stable training of very deep (up to 128-layer) residual networks on simple classification tasks with competitive performance and little tuning compared to current benchmarks. Moreover, PC enables zero-shot transfer of both weight and activity learning rates across widths and depths. Our results serve as a first step towards scaling PC to more complex architectures and have implications for other local algorithms. Code for PC is made available as part of a JAX library for PCNs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Towards the Training of Deeper Predictive Coding Neural NetworksChang Qi, Matteo Forasassi, Thomas Lukasiewicz, Tommaso SalvatoriICML 2026 · 被引用 6 次
- Difference Predictive Coding for Training Spiking Neural NetworksVille Karlsson, Nicklas Fianda, Joni-Kristian KämäräinenICLR 2026
- Local Reinforcement Learning with Action-Conditioned Root Mean Squared Q-FunctionsZequan Wu, Mengye RenICLR 2026
- Stable and Scalable Deep Predictive Coding Networks with Meta-Prediction ErrorsMyoung Hoon Ha, Hyunjun Kim, Yoondo Sung, Youngha Jo 等ICLR 2026
它引用的顶会 Paper21
- Tensor Programs IV: Feature Learning in Infinite-Width Neural NetworksGreg Yang, Edward J. HuICML 2021 · 被引用 242 次
- Tuning Large Neural Networks via Zero-Shot Hyperparameter TransferGe Yang, Edward J. Hu, Igor Babuschkin, Szymon Sidor 等NeurIPS 2021 · 被引用 208 次
- Can the Brain Do Backpropagation? - Exact Implementation of Backpropagation in Predictive Coding NetworksYuhang Song, Thomas Lukasiewicz, Zhenghua Xu, Rafal BogaczNeurIPS 2020 · 被引用 117 次
- Error-driven Input Modulation: Solving the Credit Assignment Problem without a Backward PassGiorgia Dellaferrera, Gabriel KreimanICML 2022 · 被引用 80 次
- Don't be lazy: CompleteP enables compute-efficient deep transformersNolan Dey, Bin Claire Zhang, Lorenzo Noci, Mufan Bill Li 等NeurIPS 2025 · 被引用 77 次
相关 Paper
- On the Infinite Width and Depth Limits of Predictive Coding NetworksFrancesco Innocenti, El Mehdi Achour, Rafal BogaczICML 2026
- A Theoretical Framework for Inference and Learning in Predictive Coding NetworksBeren Millidge, Yuhang Song, Tommaso Salvatori, Thomas Lukasiewicz 等ICLR 2023 · 被引用 8 次
- Local Loss Optimization in the Infinite Width: Stable Parameterization of Predictive Coding Networks and Target PropagationSatoki Ishikawa, Rio Yokota, Ryo KarakidaICLR 2025
- Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?Francesco Innocenti, El Mehdi Achour, Ryan Singh, Christopher L. BuckleyNeurIPS 2024 · 被引用 10 次
- Reverse Differentiation via Predictive CodingTommaso Salvatori, Yuhang Song, Zhenghua Xu, Thomas Lukasiewicz 等AAAI 2022 · 被引用 37 次
