Reg-PTQ: Regression-specialized Post-training Quantization for Fully Quantized Object Detector
Yifu Ding, Weilun Feng, Chuyan Chen, Jinyang Guo, Xianglong Liu
摘要
Although deep learning based object detection is of great significance for various applications, it faces challenges when deployed on edge devices due to the computation and energy limitations. Post-training quantization (PTQ) can improve inference efficiency through integer computing. However, they suffer from severe performance degradation when performing full quantization due to overlooking the unique characteristics of regression tasks in object detection. In this paper, we are the first to explore regression-friendly quantization and conduct full quantization on various detectors. We reveal the intrinsic reason behind the difficulty of quantizing regressors with empirical and theoretical justifications, and introduce a novel Regression-specialized Post-Training Quantization (Reg-PTQ) scheme. It includes Filtered Global Loss Integration Calibration to combine the global loss with a two-step filtering mechanism, mitigating the adverse impact of false positive bounding boxes, and Learnable Logarithmic-Affine Quantizer tailored for the non-uniform distributed parameters in regression structures. Extensive experiments on prevalent detectors showcase the effectiveness of the welldesigned Reg-PTQ. Notably, our Reg-PTQ achieves 7.6× and 5.4× reduction in computation and storage consumption under INT4 with little performance degradation, which indicates the immense potential of fully quantized detectors in real-world object detection applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion ModelsWeilun Feng, Haotong Qin, Chuanguang Yang, Zhulin An 等AAAI 2025 · 被引用 19 次
- QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention SparsificationWeilun Feng, Chuanguang Yang, Haotong Qin, Mingqiang Wu 等ICLR 2026 · 被引用 8 次
- S2Q-VDiT: Accurate Quantized Video Diffusion Transformer with Salient Data and Sparse Token DistillationWeilun Feng, Haotong Qin, Chuanguang Yang, Xiangqi Li 等NeurIPS 2025 · 被引用 1 次
- Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion TransformersWeilun Feng, Chuanguang Yang, Haotong Qin, Xiangqi Li 等ICML 2025
- FIMA-Q: Post-Training Quantization for Vision Transformers by Fisher Information Matrix ApproximationZhuguanyu Wu, Shihe Wang, Jiayi Zhang, Jiaxin Chen 等CVPR 2025
它引用的顶会 Paper25
- MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision TransformerSachin Mehta, Mohammad RastegariICLR 2022 · 被引用 2,162 次
- LeViT: a Vision Transformer in ConvNet's Clothing for Faster InferenceBenjamin Graham, Alaaeldin El-Nouby, Hugo Touvron, Pierre Stock 等ICCV 2021 · 被引用 1,009 次
- Up or Down? Adaptive Rounding for Post-Training QuantizationMarkus Nagel, Rana Ali Amjad, Mart van Baalen, Christos Louizos 等ICML 2020 · 被引用 816 次
- BRECQ: Pushing the Limit of Post-Training Quantization by Block ReconstructionYuhang Li, Ruihao Gong, Xu Tan, Yang Yang 等ICLR 2021 · 被引用 619 次
- Post-Training Quantization for Vision TransformerZhenhua Liu, Yunhe Wang, Kai Han, Wei Zhang 等NeurIPS 2021 · 被引用 528 次
相关 Paper
- Point4Bit: Post Training 4-bit Quantization for Point Cloud 3D DetectionJianyu Wang, Yu Wang, Shengjie Zhao, Sifan ZhouNeurIPS 2025 · 被引用 2 次
- LiDAR-PTQ: Post-Training Quantization for Point Cloud 3D Object DetectionSifan Zhou, Liang Li, Xinyu Zhang, Bo Zhang 等ICLR 2024 · 被引用 40 次
- AQD: Towards Accurate Quantized Object DetectionPeng Chen, Jing Liu, Bohan Zhuang, Mingkui Tan 等CVPR 2021
- PD-Quant: Post-Training Quantization Based on Prediction Difference MetricJiawei Liu, Lin Niu, Zhihang Yuan, Dawei Yang 等CVPR 2023
- Towards Accurate Post-training Network Quantization via Bit-Split and StitchingPeisong Wang, Qiang Chen, Xiangyu He, Jian ChengICML 2020 · 被引用 159 次
