arrow
Return

LSA-MEP: layer-wise sparsity allocation multi-metric evaluation pruning

delete2025-08-27
delete0
PRE
AI
梁鸿 (Liang Hong)
Q
Quanyi Guo *
邵明文 cover
邵明文 (Mingwen Shao)
Q
Qian Zhang
DOI:10.1007/s11227-025-07786-7delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
As deep neural networks continue to grow in scale, their rising computational demands increasingly require high-performance computing (HPC) resources such as distributed clusters. This highlights the need for efficient, HPC-compatible models. Network pruning is a key technique for reducing model complexity, but most existing methods apply uniform sparsity and rely on single-criterion importance metrics, overlooking structural heterogeneity and the multi-dimensional nature of parameter significance. To address these limitations, we propose LSA-MEP: a pruning-at-initialization framework that integrates layer-wise sparsity allocation with multi-metric parameter evaluation. Pruning ratios are determined by each layer’s information content, while parameter importance is assessed through a multi-dimensional evaluation that incorporates static structural characteristics, dynamic optimization signals, and information propagation capacity, yielding a comprehensive and robust importance metric. The process is fully parallelizable, making it well-suited for large-scale models and distributed environments. Experiments across multiple datasets demonstrate that LSA-MEP consistently outperforms existing methods in accuracy preservation while achieving efficient execution through parallel computing.
Keywords:
DNNs
Network pruning
Sparsity
Entropy

Journal

Journal of Supercomputing cover
Journal of Supercomputing
IF:
2.7
Papers:
1.0K
Citations:
1.0W

Organization

C
College of Computer Science and Technology
Scholars:
840
Papers: 292
Citations: 0