μPC: Scaling Predictive Coding to 100+ Layer Networks
Francesco Innocenti, El Mehdi Achour, Christopher L. Buckley
Abstract
The biological implausibility of backpropagation (BP) has motivated many alternative, brain-inspired algorithms that attempt to rely only on local information, such as predictive coding (PC) and equilibrium propagation. However, these algorithms have notoriously struggled to train very deep networks, preventing them from competing with BP in large-scale settings. Indeed, scaling PC networks (PCNs) has recently been posed as a challenge for the community (Pinchetti et al., 2024). Here, we show that 100+ layer PCNs can be trained reliably using a Depth-P parameterisation (Yang et al., 2023; Bordelon et al., 2023) which we call"PC". By analysing the scaling behaviour of PCNs, we reveal several pathologies that make standard PCNs difficult to train at large depths. We then show that, despite addressing only some of these instabilities, PC allows stable training of very deep (up to 128-layer) residual networks on simple classification tasks with competitive performance and little tuning compared to current benchmarks. Moreover, PC enables zero-shot transfer of both weight and activity learning rates across widths and depths. Our results serve as a first step towards scaling PC to more complex architectures and have implications for other local algorithms. Code for PC is made available as part of a JAX library for PCNs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30e596f7-dc97-4b58-ad09-bd09e46e20f1Cited by top-tier papers4
- Towards the Training of Deeper Predictive Coding Neural NetworksChang Qi, Matteo Forasassi, Thomas Lukasiewicz, Tommaso SalvatoriICML 2026 · 6 citations
- Difference Predictive Coding for Training Spiking Neural NetworksVille Karlsson, Nicklas Fianda, Joni-Kristian KämäräinenICLR 2026
- Local Reinforcement Learning with Action-Conditioned Root Mean Squared Q-FunctionsZequan Wu, Mengye RenICLR 2026
- Stable and Scalable Deep Predictive Coding Networks with Meta-Prediction ErrorsMyoung Hoon Ha, Hyunjun Kim, Yoondo Sung, Youngha Jo et al.ICLR 2026
Builds on21
- Tensor Programs IV: Feature Learning in Infinite-Width Neural NetworksGreg Yang, Edward J. HuICML 2021 · 242 citations
- Tuning Large Neural Networks via Zero-Shot Hyperparameter TransferGe Yang, Edward J. Hu, Igor Babuschkin, Szymon Sidor et al.NeurIPS 2021 · 208 citations
- Can the Brain Do Backpropagation? - Exact Implementation of Backpropagation in Predictive Coding NetworksYuhang Song, Thomas Lukasiewicz, Zhenghua Xu, Rafal BogaczNeurIPS 2020 · 117 citations
- Error-driven Input Modulation: Solving the Credit Assignment Problem without a Backward PassGiorgia Dellaferrera, Gabriel KreimanICML 2022 · 80 citations
- Don't be lazy: CompleteP enables compute-efficient deep transformersNolan Dey, Bin Claire Zhang, Lorenzo Noci, Mufan Bill Li et al.NeurIPS 2025 · 77 citations
Related papers
- On the Infinite Width and Depth Limits of Predictive Coding NetworksFrancesco Innocenti, El Mehdi Achour, Rafal BogaczICML 2026
- A Theoretical Framework for Inference and Learning in Predictive Coding NetworksBeren Millidge, Yuhang Song, Tommaso Salvatori, Thomas Lukasiewicz et al.ICLR 2023 · 8 citations
- Local Loss Optimization in the Infinite Width: Stable Parameterization of Predictive Coding Networks and Target PropagationSatoki Ishikawa, Rio Yokota, Ryo KarakidaICLR 2025
- Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?Francesco Innocenti, El Mehdi Achour, Ryan Singh, Christopher L. BuckleyNeurIPS 2024 · 10 citations
- Reverse Differentiation via Predictive CodingTommaso Salvatori, Yuhang Song, Zhenghua Xu, Thomas Lukasiewicz et al.AAAI 2022 · 37 citations
