Implicit Feature Decoupling with Depthwise Quantization
Iordanis Fostiropoulos, Barry W. Boehm
摘要
Quantization has been applied to multiple domains in Deep Neural Networks (DNNs). We propose Depthwise Quantization (DQ) where quantization is applied to a de-composed sub-tensor along the feature axis of weak statis-tical dependence. The feature decomposition leads to an exponential increase in representation capacity with a linear increase in memory and parameter cost. In addition, DQ can be directly applied to existing encoder-decoder frame-works without modification of the DNN architecture. We use DQ in the context of Hierarchical Auto-Encoders and train end-to-end on an image feature representation. We provide an analysis of the cross-correlation between spatial and channel features and propose a decomposition of the image feature representation along the channel axis. The improved performance of the depthwise operator is due to the increased representation capacity from implicit feature decoupling. We evaluate DQ on the likelihood estimation task, where it outperforms the previous state-of-the-art on CIFAR-10, ImageNet-32 and ImageNet-64. We progressively train with increasing image size a single hierarchical model that uses 69% fewer parameters and has faster convergence than the previous work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- And the Bit Goes Down: Revisiting the Quantization of Neural NetworksPierre Stock, Armand Joulin, Rémi Gribonval, Benjamin Graham 等ICLR 2020 · 被引用 157 次
- Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on ImagesRewon ChildICLR 2021 · 被引用 45 次
- Rethinking Depthwise Separable Convolutions: How Intra-Kernel Correlations Lead to Improved MobileNetsDaniel Haase, Manuel AmthorCVPR 2020
相关 Paper
- Divide and Conquer: Leveraging Intermediate Feature Representations for Quantized Training of Neural NetworksAhmed Taha Elthakeb, Prannoy Pilligundla, Fatemeh Mireshghallah, Alexander Cloninger 等ICML 2020 · 被引用 9 次
- Term quantization: furthering quantization at run timeHsiang-Tsung Kung, Bradley McDanel, Sai Qian ZhangSC 2020 · 被引用 10 次
- Network Quantization With Element-Wise Gradient ScalingJunghyup Lee, Dohyung Kim, Bumsub HamCVPR 2021
- Learned Step Size quantizationSteven K. Esser, Jeffrey L. McKinstry, Deepika Bablani, Rathinakumar Appuswamy 等ICLR 2020 · 被引用 1,037 次
- JPEG Inspired Deep LearningAhmed H. Salamah, Kaixiang Zheng, Yiwen Liu, En-Hui YangICLR 2025
