Enhancing Post-Training Quantization Calibration Through Contrastive Learning
Yuzhang Shang, Gaowen Liu, Ramana Rao Kompella, Yan Yan
摘要
Post-training quantization (PTQ) converts a pre-trained full-precision (FP) model into a quantized model in a training-free manner. Determining suitable quantization parameters, such as scaling factors and zero points, is the primary strategy for mitigating the impact of quantization noise (calibration) and restoring the performance of the quantized models. However, the existing activation calibration methods have never considered information degradation between pre- (FP) and post-quantized activations. In this study, we introduce a well-defined distributional metric from information theory, mutual information, into PTQ calibration. We aim to calibrate the quantized activations by maximizing the mutual information between the pre- and post-quantized activations. To realize this goal, we establish a contrastive learning (CL) framework for the calibration, where the quantization parameters are optimized through a self-supervised proxy task. Specifically, by leveraging CL during the PTQ calibration, we can benefit from pulling the positive pairs of quantized and FP activations collected from the same input samples, while pushing negative pairs from different samples. Thanks to the ingeniously designed critic function, we avoid the unwanted but of tenencountered collision solution in CL, especially in calibration scenarios where the amount of calibration data is limited. Additionally, we provide a theoretical guarantee that minimizing our designed loss is equivalent to maximizing the desired mutual information. Consequently, the quantized activations retain more information, which ultimately enhances the performance of the quantized network. Experimental results show that our method can effectively serve as an add-on module to existing SoTA PTQ methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Preserving LLM Capabilities through Calibration Data Curation: From Analysis to OptimizationBowei He, Lihao Yin, Hui-Ling Zhen, Shuqi Liu 等NeurIPS 2025 · 被引用 9 次
- FP=XINT: Representing Neural Networks via Low-Bit Series Basis FunctionsBoyang Zhang, Daning Cheng, Yunquan Zhang, Jiake Tian 等AAAI 2026 · 被引用 4 次
- SHARP-Q: Spectral Hessian Alignment and Rectification for Post-training QuantizationMenghao Lv, Huiqiong Wang, Li Sun, Mingli SongICML 2026
- Rethinking Asymmetric Quantization: Hidden Symmetry in Vision Model WeightsMasafumi Mori, Shinya Gongyo, Mitsuru AmbaiCVPR 2026
- QT-DoG: Quantization-Aware Training for Domain GeneralizationSaqib Javed, Hieu Le, Mathieu SalzmannICML 2025
它引用的顶会 Paper16
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Contrastive Representation DistillationYonglong Tian, Dilip Krishnan, Phillip IsolaICLR 2020 · 被引用 1,305 次
- Learned Step Size quantizationSteven K. Esser, Jeffrey L. McKinstry, Deepika Bablani, Rathinakumar Appuswamy 等ICLR 2020 · 被引用 1,037 次
- Up or Down? Adaptive Rounding for Post-Training QuantizationMarkus Nagel, Rana Ali Amjad, Mart van Baalen, Christos Louizos 等ICML 2020 · 被引用 816 次
- Q-BERT: Hessian Based Ultra Low Precision Quantization of BERTSheng Shen, Zhen Dong, Jiayu Ye, Linjian Ma 等AAAI 2020 · 被引用 656 次
相关 Paper
- PD-Quant: Post-Training Quantization Based on Prediction Difference MetricJiawei Liu, Lin Niu, Zhihang Yuan, Dawei Yang 等CVPR 2023
- Toward Accurate Post-Training Quantization for Image Super ResolutionZhijun Tu, Jie Hu, Hanting Chen, Yunhe WangCVPR 2023
- QDrop: Randomly Dropping Quantization for Extremely Low-bit Post-Training QuantizationXiuying Wei, Ruihao Gong, Yuhang Li, Xianglong Liu 等ICLR 2022 · 被引用 248 次
- Beyond Uniformity: Sample and Frequency Meta Weighting for Post-Training Quantization of Diffusion ModelsVan Cuong Pham, Anh Hoang, Cuong Nguyen, Trung Le 等ICLR 2026
- LRQuant: Learnable and Robust Post-Training Quantization for Large Language ModelsJiaqi Zhao, Miao Zhang, Chao Zeng, Ming Wang 等ACL 2024
