arrow
Return

GAP: A group-based automatic pruning algorithm via convolution kernel fusion

delete2024-12-01
delete0
PRE
AI
D
Dingfu Chen
K
Kangwei Lin
Q
Qingxu Deng *
DOI:10.1016/j.neucom.2024.128488delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In recent years, the deployment and operation of convolution neural networks on edge devices with limited computing capabilities have become increasingly challenging due to the large network structure and computational cost. Currently, the mainstream structured pruning algorithms mainly compress the network at the filter or layer level. However, these methods introduce too much human intervention with large granularities, which may lead to unpredictable performance after compression. In this paper, we propose a group-based automatic pruning algorithm(GAP) via kernel fusion to automatically search for the optimal pruning structure in a more finegrained manner. Specifically, we first adopt a novel nonlinear dimensionality reduction clustering algorithm to divide the filters of each convolution layer into groups of equal size. Afterwards, we encode the mutual distribution similarity of the kernels within each group, and its KL divergence is employed as an importance indicator to determine the retained kernel groups through weighted fusion. Subsequently, we introduce an intelligent searching module that automatically explore and optimize the pruned structure of each layer. Finally, the pruned filters are permutated to form a dense group convolution and fine-tuned. Sufficient experiments show that, on two image classification datasets, for five advanced CNN models, our GAP algorithm outperforms most extant SOTA schemes, reduces artificial intervention, and enables efficient end-to-end training of compact models.
Keywords:
Convolutional neural network
Model compressing
Automatic group pruning
Filter clustering
Kernel fusion

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

N
northeastern university - china
Scholars:
3.1W
Papers: 2.7W
Citations: 37
Cited Papers

Cited Papers

Soft Hybrid Knowledge Distillation against deep neural networks
err2024-02-01
err8
PREAI
errZhang, Jian; Tao, Ze; Zhang, Shichao; Qiao, Zike; Guo, Kehua
errShare
errSave
Sensitivity pruner: Filter-Level compression algorithm for deep neural networks
err2023-08-01
err9
PREAI
errGuo, Suhan; Lai, Bilan; Yang, Suorong; Zhao, Jian; Shen, Furao
errShare
errSave
Pruning filters with L1-norm and capped L1-norm for CNN compression
err2020-09-17
err126
PREAI
errKumar, Aakash; Shaikh, Ali Muhammad; Li, Yun; Bilal, Hazrat; Yin, Baoqun
errShare
errSave
ImageNet Large Scale Visual Recognition Challenge
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
errShare
errSave
Methods for Pruning Deep Neural Networks
err2022-01-01
err91
errOAAI
errVadera, Sunil; Ameen, Salem
errShare
errSave
Asymptotic Soft Filter Pruning for Deep Convolutional Neural Networks
err2020-08-01
err135
errOAAI
errHe, Yang; Dong, Xuanyi; Kang, Guoliang; Fu, Yanwei; Yan, Chenggang; Yang, Yi
errShare
errSave
errShare
errSave
A comprehensive survey on model compression and acceleration
err2020-02-08
err260
PREAI
errChoudhary, Tejalal; Mishra, Vipul; Goswami, Anurag; Sarangapani, Jagannathan
errShare
errSave
researcher View more