arrow
返回

Quantization-Aware Training With Dynamic and Static Pruning

delete2025-01-01
delete0
delete
OA
AI
S
Sangho An
S
Shin, Jongyun
J
Jangho Kim *
DOI:10.1109/ACCESS.2025.3556629delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The evolution of deep neural networks (DNNs) naturally leads to an increase in model size. This necessitates various model compression techniques, such as pruning and quantization, to reduce memory usage and power consumption. In particular, combining these compression techniques can achieve significant cost savings. However, we found that methods using dynamic pruning and quantization suffer from instability in training and poor generalization performance due to the effects of the two Straight Through Estimators (STE). To address this problem, we propose a Quantization-aware training with Dynamic and Static pruning (QADS) method that takes advantage of both pruning and quantization by performing STE operations only during quantization from a certain point in time. In our experiments, the proposed method exhibits more stable training compared to existing techniques and achieves performance improvements on the CIFAR-10/100, ImageNet, and Google Speech Command datasets. The code is provided at https://github.com/Ahnho/Quantization-aware-training-with-Dynamic-and-Static-Pruning.
Keyword:
Quantization (signal)
Training
Computational modeling
Artificial neural networks
Accuracy
Stability analysis
Backpropagation
Iterative methods
Degradation
Computational efficiency
Convolutional neural networks
network quantization
network pruning

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

K
kookmin university
学者数:
3.0K
论文数: 3.3K
被引数: 2
引用论文

引用论文

暂无论文信息