arrow
返回

Attention-based adaptive structured continuous sparse network pruning

delete2024-07-01
delete0
PRE
AI
J
Jiaxin Liu
W
Wei Liu *
李永明 封面图
李永明 (Yongming Li)
J
Jun Hu
S
Shuai Cheng
W
Wenxing Yang
DOI:10.1016/j.neucom.2024.127698delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep neural network models, especially CNNs, have a wide range of applications in many fields, but their high computational power requirements limit the deployment applications in many resource-constrained embedded devices. Pruning techniques reduce the computational power requirements of models by removing redundant structures from CNNs. Most existing static pruning methods use a global uniform pruning rate to prune pretrained models and require finetuning to recover accuracy after pruning, resulting in high training costs, and the global uniform pruning rate is sub-optimal. While dynamic pruning methods perform pruning during training and use auxiliary modules to calculate the saliency scores of channels, but do not exploit its function of assisting network training. We propose an adaptive structured continuous sparse network pruning method based on the attention mechanism that prunes the original network during training. The attention -based channel similarity calculation module calculates the channel saliency scores while refining features to assist network training, and the adaptive continuous sparse control module gradually discretizes the channel saliency scores and assigns the pruning rate of each layer according to the preset pruning rate target. The pruned model is output after training and no additional fine -tuning is required. We validate the proposed method on CIFAR-10 and the large-scale dataset Imagenet using networks with different structures, and our method outperforms the comparative pruning methods at different pruning rates. In CIFAR-10, our method can reduce VGG-16 by 34.4% FLOPs while top -1 accuracy increased by 0.19%. In Imagenet, we can reduce ResNet-34 by 51.5% FLOPs, while the top -1 accuracy decreases by only 0.89%.
Keyword:
Convolutional neural networks
Model compression
Pruning
Attention

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

L
liaoning university of technology
学者数:
2.3K
论文数: 1.6K
被引数: 1
N
northeastern university - china
学者数:
3.1W
论文数: 2.7W
被引数: 37
引用论文

引用论文

Dynamic hard pruning of Neural Networks at the edge of the internet
err2022-04-01
err4
errOAAI
errValerio, Lorenzo; Nardini, Franco Maria; Passarella, Andrea; Perego, Raffaele
err分享
err收藏
Dynamical Channel Pruning by Conditional Accuracy Change for Deep Neural Networks
err2021-02-01
err52
PREAI
errChen, Zhiqiang; Xu, Ting-Bing; Du, Changde; Liu, Cheng-Lin; He, Huiguang
err分享
err收藏
Major Adverse Limb Events and Mortality in Patients With Peripheral Artery Disease
err2018-05-01
err0
errOAAI
errSonia S. Anand; Francois Caron; John W. Eikelboom; Jackie Bosch; Leanne Dyal; Victor Aboyans; Maria Teresa Abola; Kelley R.H. Branch; Katalin Keltai; Deepak L. Bhatt; Peter Verhamme; Keith A.A. Fox; Nancy Cook-Bruns; Vivian Lanius; Stuart J. Connolly; Salim Yusuf
err分享
err收藏
Filter Sketch for Network Pruning用于网络修剪的过滤器草图
err2022-12-01
err55
errOAAI
errLin, Mingbao; Cao, Liujuan; Li, Shaojie; Ye, Qixiang; Tian, Yonghong; Liu, Jianzhuang; Tian, Qi; Ji, Rongrong
err分享
err收藏
ImageNet Large Scale Visual Recognition ChallengeImageNet大规模视觉识别挑战
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
err分享
err收藏
Asymptotic Soft Filter Pruning for Deep Convolutional Neural Networks
err2020-08-01
err135
errOAAI
errHe, Yang; Dong, Xuanyi; Kang, Guoliang; Fu, Yanwei; Yan, Chenggang; Yang, Yi
err分享
err收藏
学者 查看更多内容