Zero-Shot Adversarial Quantization
Yuang Liu, Wei Zhang, Jun Wang
Abstract
Model quantization is a promising approach to compress deep neural networks and accelerate inference, making it possible to be deployed on mobile and edge devices. To retain the high performance of full-precision models, most existing quantization methods focus on fine-tuning quantized model by assuming training datasets are accessible. However, this assumption sometimes is not satisfied in real situations due to data privacy and security issues, thereby making these quantization methods not applicable. To achieve zero-short model quantization without accessing training data, a tiny number of quantization methods adopt either post-training quantization or batch normalization statisticsguided data generation for fine-tuning. However, both of them inevitably suffer from low performance, since the former is a little too empirical and lacks training support for ultra-low precision quantization, while the latter could not fully restore the peculiarities of original data and is often low efficient for diverse data generation. To address the above issues, we propose a zero-shot adversarial quantization (ZAQ) framework, facilitating effective discrepancy estimation and knowledge transfer from a full-precision model to its quantized model. This is achieved by a novel two-level discrepancy modeling to drive a generator to synthesize informative and diverse data examples to optimize the quantized model in an adversarial learning fashion. We conduct extensive experiments on three fundamental vision tasks, demonstrating the superiority of ZAQ over the strong zero-shot baselines and validating the effectiveness of its main components. Code is available at https://git.io/Jqc0y .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2bfc4721-5c18-4ed2-ac4f-a04dbc5351cbCited by top-tier papers25
- SQuant: On-the-Fly Data-Free Quantization via Diagonal Hessian ApproximationCong Guo, Yuxian Qiu, Jingwen Leng, Xiaotian Gao et al.ICLR 2022 · 92 citations
- Qimera: Data-free Quantization with Synthetic Boundary Supporting SamplesKanghyun Choi, Deokki Hong, Noseong Park, Youngsok Kim et al.NeurIPS 2021 · 87 citations
- IntraQ: Learning Synthetic Images with Intra-Class Heterogeneity for Zero-Shot Network QuantizationYunshan Zhong, Mingbao Lin, Gongrui Nan, Jianzhuang Liu et al.CVPR 2022 · 79 citations
- Wavelet Feature Maps Compression for Image-to-Image CNNsShahaf E. Finder, Yair Zohav, Maor Ashkenazi, Eran TreisterNeurIPS 2022 · 63 citations
- It's All In the Teacher: Zero-Shot Quantization Brought Closer to the TeacherKanghyun Choi, Hyeyoon Lee, Deokki Hong, Joonsang Yu et al.CVPR 2022 · 33 citations
Builds on7
- Q-BERT: Hessian Based Ultra Low Precision Quantization of BERTSheng Shen, Zhen Dong, Jiayu Ye, Linjian Ma et al.AAAI 2020 · 656 citations
- Data-Free Quantization Through Weight Equalization and Bias CorrectionMarkus Nagel, Mart van Baalen, Tijmen Blankevoort, Max WellingICCV 2019 · 622 citations
- Data-Free Learning of Student NetworksHanting Chen, Yunhe Wang, Chang Xu, Zhaohui Yang et al.ICCV 2019 · 427 citations
- ZeroQ: A Novel Zero Shot Quantization FrameworkYaohui Cai, Zhewei Yao, Zhen Dong, Amir Gholami et al.CVPR 2020
- Data-Free Knowledge Amalgamation via Group-Stack Dual-GANJingwen Ye, Yixin Ji, Xinchao Wang, Xin Gao et al.CVPR 2020
Related papers
- Genie: Show Me the Data for QuantizationYongkweon Jeon, Chungman Lee, Ho-Young KimCVPR 2023
- SynQ: Accurate Zero-shot Quantization by Synthesis-aware Fine-tuningMinjun Kim, Jongjin Kim, U KangICLR 2025
- Sharpness-Aware Data Generation for Zero-shot QuantizationHoang Anh Dung, Cuong Pham, Trung Le, Jianfei Cai et al.ICML 2024 · 8 citations
- Quantize Sequential Recommenders Without Private DataLingfeng Shi, Yuang Liu, Jun Wang, Wei ZhangWWW 2023 · 3 citations
- Zero-shot Sharpness-Aware Quantization for Pre-trained Language ModelsMiaoxi Zhu, Qihuang Zhong, Li Shen, Liang Ding et al.EMNLP 2023 · 2 citations
