Mind the Gap: Removing the Discretization Gap in Differentiable Logic Gate Networks
Shakir Yousefi, Andreas Plesner, Till Aczel, Roger Wattenhofer
Abstract
Modern neural networks exhibit state-of-the-art performance on many existing benchmarks, but their high computational requirements and energy usage cause researchers to explore more efficient solutions for real-world deployment. Differentiable logic gate networks (DLGNs) learns a large network of logic gates for efficient image classification. However, learning a network that can solve simple problems like CIFAR-10 or CIFAR-100 can take days to weeks to train. Even then, almost half of the neurons remains unused, causing a discretization gap. This discretization gap hinders real-world deployment of DLGNs, as the performance drop between training and inference negatively impacts accuracy. We inject Gumbel noise with a straight-through estimator during training to significantly speed up training, improve neuron utilization, and decrease the discretization gap. We theoretically show that this results from implicit Hessian regularization, which improves the convergence properties of DLGNs. We train networks 4.5× faster in wall-clock time, reduce the discretization gap by 98%, and reduce the number of unused gates by 100%. * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Light Differentiable Logic Gate NetworksLukas Rüttgers, Till Aczel, Andreas Plesner, Roger WattenhoferICLR 2026 · 11 citations
- Align Forward, Adapt Backward: Closing the Discretization Gap in Logic Gate NetworksYoungsung KimICML 2026 · 1 citation
- Differentiable Weightless Controllers: Learning Logic Circuits for Continuous ControlFabian Kresse, Christoph LampertICML 2026 · 1 citation
- Two-Stage Unit Tying for Simplifying Differentiable Logic Gate NetworksSeungheon Lee, Jeongmin Sun, Jaeyong ChungICML 2026
Builds on20
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 1,861 citations
- SparseGPT: Massive Language Models Can be Accurately Pruned in One-ShotElias Frantar, Dan AlistarhICML 2023 · 1,240 citations
- A Simple and Effective Pruning Approach for Large Language ModelsMingjie Sun, Zhuang Liu, Anna Bair, J. Zico KolterICLR 2024 · 794 citations
- Fantastic Generalization Measures and Where to Find ThemYiding Jiang, Behnam Neyshabur, Hossein Mobahi, Dilip Krishnan et al.ICLR 2020 · 705 citations
- Stabilizing Differentiable Architecture Search via Perturbation-based RegularizationXiangning Chen, Cho-Jui HsiehICML 2020 · 235 citations
Related papers
- Deep Differentiable Logic Gate NetworksFelix Petersen, Christian Borgelt, Hilde Kuehne, Oliver DeussenNeurIPS 2022 · 117 citations
- Convolutional Differentiable Logic Gate NetworksFelix Petersen, Hilde Kuehne, Christian Borgelt, Julian Welzel et al.NeurIPS 2024 · 58 citations
- Training Discrete Deep Generative Models via Gapped Straight-Through EstimatorTing-Han Fan, Ta-Chung Chi, Alexander I. Rudnicky, Peter J. RamadgeICML 2022 · 9 citations
- Discrete Model Compression With Resource Constraint for Deep Neural NetworksShangqian Gao, Feihu Huang, Jian Pei, Heng HuangCVPR 2020
- Injecting Logical Constraints into Neural Networks via Straight-Through EstimatorsZhun Yang, Joohyung Lee, Chiyoun ParkICML 2022 · 26 citations
