Kernel Quantization for Efficient Network Compression

Zhongzhi Yu, Yemin Shi, Tiejun Huang, Yizhou Yu

Abstract

This paper presents a novel network compression framework Kernel Quantization (KQ), targeting to efficiently convert any pre-trained full-precision convolutional neural network (CNN) model into a low-precision version without significant performance loss. Unlike existing methods struggling with weight bit-length, KQ has the potential in improving the compression ratio by considering the convolution kernel as the quantization unit. Inspired by the evolution from weight pruning to filter pruning, we propose to quantize in both kernel and weight level. Instead of representing each weight parameter with a low-bit index, we learn a kernel codebook and replace all kernels in the convolution layer with corresponding low-bit indexes. Thus, KQ can represent the weight tensor in the convolution layer with low-bit indexes and a kernel codebook with limited size, which enables KQ to achieve significant compression ratio. Then, we conduct a 6-bit parameter quantization on the kernel codebook to further reduce redundancy. Extensive experiments on the ImageNet classification task prove that KQ needs 1.05 and 1.62 bits on average in VGG and ResNet18, respectively, to represent each parameter in the convolution layer and achieves the state-of-the-art compression ratio with little accuracy loss.

Keywords

Artificial Intelligence & Data Science

📄 Full Paper Available as PDF

This paper is available as a downloadable PDF.

📄 Download PDF

Comments (0)

No comments yet. Be the first to comment.

Paper Details

Authors Zhongzhi Yu ,
Yemin Shi ,
Tiejun Huang ,
Yizhou Yu
Published 2020-03-11
Category Artificial Intelligence And Data Science
Status Non-peer-reviewed Preprint
Language English
Word Count 187

Kernel Quantization for Efficient Network Compression

Abstract

Keywords

✨ AI Plain-English Summary

Comments (0)

Related Papers

Introducing New AdaBoost Features for Real-Time Vehicle Detection

Visual object categorization with new keypoint-based adaBoost features

Proceedings 6th International Workshop on Local Search Techniques in ...

Computer-Assisted Decision Support System in Pulmonary Cancer detection and...