arrow
Return

Attention-based adaptive structured continuous sparse network pruning

delete2024-07-01
delete0
PRE
AI
J
Jiaxin Liu
W
Wei Liu *
李永明 cover
李永明 (Yongming Li)
J
Jun Hu
S
Shuai Cheng
W
Wenxing Yang
DOI:10.1016/j.neucom.2024.127698delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Deep neural network models, especially CNNs, have a wide range of applications in many fields, but their high computational power requirements limit the deployment applications in many resource-constrained embedded devices. Pruning techniques reduce the computational power requirements of models by removing redundant structures from CNNs. Most existing static pruning methods use a global uniform pruning rate to prune pretrained models and require finetuning to recover accuracy after pruning, resulting in high training costs, and the global uniform pruning rate is sub-optimal. While dynamic pruning methods perform pruning during training and use auxiliary modules to calculate the saliency scores of channels, but do not exploit its function of assisting network training. We propose an adaptive structured continuous sparse network pruning method based on the attention mechanism that prunes the original network during training. The attention -based channel similarity calculation module calculates the channel saliency scores while refining features to assist network training, and the adaptive continuous sparse control module gradually discretizes the channel saliency scores and assigns the pruning rate of each layer according to the preset pruning rate target. The pruned model is output after training and no additional fine -tuning is required. We validate the proposed method on CIFAR-10 and the large-scale dataset Imagenet using networks with different structures, and our method outperforms the comparative pruning methods at different pruning rates. In CIFAR-10, our method can reduce VGG-16 by 34.4% FLOPs while top -1 accuracy increased by 0.19%. In Imagenet, we can reduce ResNet-34 by 51.5% FLOPs, while the top -1 accuracy decreases by only 0.89%.
Keywords:
Convolutional neural networks
Model compression
Pruning
Attention

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

L
liaoning university of technology
Scholars:
2.3K
Papers: 1.6K
Citations: 1
N
northeastern university - china
Scholars:
3.1W
Papers: 2.7W
Citations: 37
Cited Papers

Cited Papers

errShare
errSave
Dynamic hard pruning of Neural Networks at the edge of the internet
err2022-04-01
err4
errOAAI
errValerio, Lorenzo; Nardini, Franco Maria; Passarella, Andrea; Perego, Raffaele
errShare
errSave
Dynamical Channel Pruning by Conditional Accuracy Change for Deep Neural Networks
err2021-02-01
err52
PREAI
errChen, Zhiqiang; Xu, Ting-Bing; Du, Changde; Liu, Cheng-Lin; He, Huiguang
errShare
errSave
Major Adverse Limb Events and Mortality in Patients With Peripheral Artery Disease
err2018-05-01
err0
errOAAI
errSonia S. Anand; Francois Caron; John W. Eikelboom; Jackie Bosch; Leanne Dyal; Victor Aboyans; Maria Teresa Abola; Kelley R.H. Branch; Katalin Keltai; Deepak L. Bhatt; Peter Verhamme; Keith A.A. Fox; Nancy Cook-Bruns; Vivian Lanius; Stuart J. Connolly; Salim Yusuf
errShare
errSave
Filter Sketch for Network Pruning
err2022-12-01
err55
errOAAI
errLin, Mingbao; Cao, Liujuan; Li, Shaojie; Ye, Qixiang; Tian, Yonghong; Liu, Jianzhuang; Tian, Qi; Ji, Rongrong
errShare
errSave
ImageNet Large Scale Visual Recognition Challenge
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
errShare
errSave
Asymptotic Soft Filter Pruning for Deep Convolutional Neural Networks
err2020-08-01
err135
errOAAI
errHe, Yang; Dong, Xuanyi; Kang, Guoliang; Fu, Yanwei; Yan, Chenggang; Yang, Yi
errShare
errSave
researcher View more