Lune

AAAI2026Top-tier venue

On the Impact of Weight Quantization on Deep Neural Network Uncertainty

Shuang Liang, Xun Lu, Zi-Ang Liu, Ming-Liang Wang, Yan Lyu, Shao-Qun Zhang

2026Year

Abstract

Weight Quantization (WQ) is a key technique for lightweight Deep Neural Network (DNN) computations. While existing algorithms often pursue memory compression and inference acceleration with accuracy comparable to full-precision models, the effect of WQ on DNN uncertainty remains largely unexplored. In this paper, we quantify the impact of WQ on DNN uncertainty through the novel Exact Moment Propagation (EMP) uncertainty estimator. It is observed that WQ significantly increases DNN uncertainty. Based on the EMP estimator, we propose the MOMent Alignment (MOMA) to reduce WQ-induced uncertainty and preserve the accuracy of weight-quantized DNNs. Empirical results across various DNN architectures and datasets validate the effectiveness of both EMP and MOMA methods.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext f5010313-ba11-4b3e-bdc8-6717aa088bfa

Builds on5

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines