arrow
返回

Adaptive Filter Pruning via Sensitivity Feedback

delete2024-08-01
delete15
PRE
AI
Y
Yuyao Zhang
N
Nikolaos M. Freris *
DOI:10.1109/TNNLS.2023.3246263delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Filter pruning is advocated for accelerating deep neural networks without dedicated hardware or libraries, while maintaining high prediction accuracy. Several works have cast pruning as a variant of l(1)-regularized training, which entails two challenges: 1) the l(1)-norm is not scaling-invariant (i.e., the regularization penalty depends on weight values) and 2) there is no rule for selecting the penalty coefficient to trade off high pruning ratio for low accuracy drop. To address these issues, we propose a lightweight pruning method termed adaptive sensitivity-based pruning (ASTER) which: 1) achieves scaling-invariance by refraining from modifying unpruned filter weights and 2) dynamically adjusts the pruning threshold concurrently with the training process. ASTER computes the sensitivity of the loss to the threshold on the fly (without retraining); this is carried efficiently by an application of L-BFGS solely on the batch normalization (BN) layers. It then proceeds to adapt the threshold so as to maintain a fine balance between pruning ratio and model capacity. We have conducted extensive experiments on a number of state-of-the-art CNN models on benchmark datasets to illustrate the merits of our approach in terms of both FLOPs reduction and accuracy. For example, on ILSVRC-2012 our method reduces more than 76% FLOPs for ResNet-50 with only 2.0% Top-1 accuracy degradation, while for the MobileNet v2 model it achieves 46.6% FLOPs Drop with a Top-1 Acc. Drop of only 2.77%. Even for a very lightweight classification model like MobileNet v3-small, ASTER saves 16.1% FLOPs with a negligible Top-1 accuracy drop of 0.03%.
Keyword:
Training
Adaptation models
Sensitivity
Computational modeling
Adaptive filters
Libraries
Filtering algorithms
Adaptive pruning
deep neural networks
model compression
structured model pruning

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

U
university of science & technology of china, cas
学者数:
3.2W
论文数: 2.7W
被引数: 74
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

Pruning Networks With Cross-Layer Ranking & k-Reciprocal Nearest Filters
err2023-11-01
err31
PREAI
errLin, Mingbao; Cao, Liujuan; Zhang, Yuxin; Shao, Ling; Lin, Chia-Wen; Ji, Rongrong
err分享
err收藏
Memory cells specific for myelin oligodendrocyte glycoprotein (MOG) govern the transfer of experimental autoimmune encephalomyelitis
err2011-05-01
err0
errOAAI
errJessica L. Williams; Aaron P. Kithcart; Kristen M. Smith; Todd Shawler; Gina M. Cox; Caroline C. Whitacre
err分享
err收藏
Calcium-modulating cyclophilin ligand regulates membrane trafficking of postsynaptic GABAA receptors
err2008-06-01
err0
errOAAI
errXu Yuan; Jun Yao; David Norris; David D. Tran; Richard J. Bram; Gong Chen; Bernhard Luscher
err分享
err收藏
Recent advances in convolutional neural networks卷积神经网络的最新进展
err2018-05-01
err3.8K
errOAAI
errGu, Jiuxiang; Wang, Zhenhua; Kuen, Jason; Ma, Lianyang; Shahroudy, Amir; Shuai, Bing; Liu, Ting; Wang, Xingxing; Wang, Gang; Cai, Jianfei; Chen, Tsuhan
err分享
err收藏
err分享
err收藏
Jun expression is found in neurons located in the vicinity of subacute plaques in patients with multiple sclerosis
err1996-07-01
err0
PREAI
errG. Martín; J. Seguí; P. Díaz-Villoslada; X. Montalbán; A.M. Planas; I. Ferrer
err分享
err收藏
学者 查看更多内容