arrow
返回

Quantization Robust Pruning With Knowledge Distillation

delete2023-01-01
delete6
delete
OA
AI
J
Jangho Kim *
DOI:10.1109/ACCESS.2023.3257864delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
To resolve the problem that deep neural networks (DNN) require a large number of network parameters, many researchers have sought to compress the network. Network pruning, quantization and knowledge distillation have been studied for this purpose. Considering realistic scenarios such as deploying DNN on the resource constraint device where the network uploaded in the device performs wells in various bit-widths without re-training and the network with reasonable performance, we propose quantization robust pruning with knowledge distillation (QRPK) method. In QRPK, model weights are divided into essential weigths and inessential weights based on their magnitude value. Then, QRPK trains the quantization robustness model with a high pruning ratio by making the distribution of essential weights as a quantization friendly distribution. We conducted experiments on CIFAR-10 and CIFAR-100 to verify the effectiveness of QRPK and a QRPK trained model performs well in various bit-width, as it designed by pruning, quantization robustness and knowledge distillation.
Keyword:
Quantization (signal)
Computational modeling
Convolutional neural networks
Knowledge engineering
Robustness
Performance evaluation
Neural networks
network quantization
knowledge distillation
network pruning

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

K
kookmin university
学者数:
3.0K
论文数: 3.3K
被引数: 2
引用论文

引用论文

err分享
err收藏
err分享
err收藏
Carbonyl-isocyanide mono-substitution in [Fe2Cp2(CO)4]: A re-visitation
err2021-03-01
err0
errOAAI
errLorenzo Biancalana; Gianluca Ciancaleoni; Stefano Zacchini; Guido Pampaloni; Fabio Marchetti
err分享
err收藏
Model Compression via Position-Based Scaled Gradient
err2022-01-01
err0
errOAAI
errKim, Jangho; Yoo, Kiyoon; Kwak, Nojun
err分享
err收藏
没有更多内容