Towards training digitally-tied analog blocks via hybrid gradient computation
Timothy Nest, Maxence Ernoult
Abstract
Power efficiency is plateauing in the standard digital electronics realm such that novel hardware, models, and algorithms are needed to reduce the costs of AI training. The combination of energy-based analog circuits and the Equilibrium Propagation (EP) algorithm constitutes one compelling alternative compute paradigm for gradient-based optimization of neural nets. Existing analog hardware accelerators, however, typically incorporate digital circuitry to sustain auxiliary non-weight-stationary operations, mitigate analog device imperfections, and leverage existing digital accelerators.This heterogeneous hardware approach calls for a new theoretical model building block. In this work, we introduce Feedforward-tied Energy-based Models (ff-EBMs), a hybrid model comprising feedforward and energy-based blocks accounting for digital and analog circuits. We derive a novel algorithm to compute gradients end-to-end in ff-EBMs by backpropagating and"eq-propagating"through feedforward and energy-based parts respectively, enabling EP to be applied to much more flexible and realistic architectures. We experimentally demonstrate the effectiveness of the proposed approach on ff-EBMs where Deep Hopfield Networks (DHNs) are used as energy-based blocks. We first show that a standard DHN can be arbitrarily split into any uniform size while maintaining performance. We then train ff-EBMs on ImageNet32 where we establish new SOTA performance in the EP literature (46 top-1 %). Our approach offers a principled, scalable, and incremental roadmap to gradually integrate self-trainable analog computational primitives into existing digital accelerators.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a338269f-2f07-4810-bc32-abdc261bc097Cited by top-tier papers2
- Learning long range dependencies through time reversal symmetry breakingGuillaume Pourcel, Maxence ErnoultNeurIPS 2025 · 9 citations
- Towards the Training of Deeper Predictive Coding Neural NetworksChang Qi, Matteo Forasassi, Thomas Lukasiewicz, Tommaso SalvatoriICML 2026 · 6 citations
Builds on15
- Fine-Tuning Language Models with Just Forward PassesSadhika Malladi, Tianyu Gao, Eshaan Nichani, Alex Damian et al.NeurIPS 2023 · 495 citations
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig et al.NeurIPS 2022 · 386 citations
- DeepZero: Scaling Up Zeroth-Order Optimization for Deep Model TrainingAochuan Chen, Yimeng Zhang, Jinghan Jia, James Diffenderfer et al.ICLR 2024 · 88 citations
- Error-driven Input Modulation: Solving the Credit Assignment Problem without a Backward PassGiorgia Dellaferrera, Gabriel KreimanICML 2022 · 80 citations
- Holomorphic Equilibrium Propagation Computes Exact Gradients Through Finite Size OscillationsAxel Laborieux, Friedemann ZenkeNeurIPS 2022 · 65 citations
Related papers
- Energy-based learning algorithms for analog computing: a comparative studyBenjamin Scellier, Maxence Ernoult, Jack D. Kendall, Suhas KumarNeurIPS 2023 · 54 citations
- Dual Propagation: Accelerating Contrastive Hebbian Learning with Dyadic NeuronsRasmus Kjær Høier, D. Staudt, Christopher ZachICML 2023 · 14 citations
- Towards Exact Gradient-based Training on Analog In-memory ComputingZhaoxian Wu, Tayfun Gokmen, Malte J. Rasch, Tianyi ChenNeurIPS 2024 · 11 citations
- ePC: Fast and Deep Predictive Coding in Digital SimulationCédric Goemaere, Gaspard Oliviers, Rafal Bogacz, Thomas DemeesterICML 2026 · 3 citations
- A fast algorithm to simulate nonlinear resistive networksBenjamin ScellierICML 2024 · 8 citations
