Kernel Quantization for Efficient Network Compression
Published in Greater South Information System • Jan 1, 2022
NobleIDNI5P27W27R89S39
Authors:,
Zhongzhi Yu
Yemin Shi
Abstract
This paper presents a novel network compression framework, Kernel Quantization (KQ), targeting to efficiently convert any pre-trained full-precision convolutional neural network (CNN) model into a low-precision version without significant performance loss.Unlike existing methods struggling with weig...
Finding related papers...
Discussions
(0)No comments yet
Be the first to share your thoughts!