arrow
返回

Convolutional Neural Network Compression via Dynamic Parameter Rank Pruning

delete2025-01-01
delete0
delete
OA
AI
M
Manish Sharma
J
Jamison Heard
E
Eli Saber
P
Panos P. Markopoulos *
DOI:10.1109/ACCESS.2025.3533419delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
While Convolutional Neural Networks (CNNs) excel at learning complex latent-space representations, their over-parameterization can lead to overfitting and reduced performance, particularly with limited data. This, alongside their high computational and memory demands, limits the applicability of CNNs for edge deployment and applications where computational resources are constrained. Low-rank matrix approximation has emerged as a promising approach to reduce CNN parameters, but existing methods often require pre-determined ranks or involve complex post-training adjustments, leading to challenges in rank selection, performance loss, and limited practicality in resource-constrained environments. This underscores the need for an adaptive compression method that integrates into the training process, dynamically adjusting model complexity based on data and task requirements. To address this, we propose an efficient training method for CNN compression via dynamic parameter rank pruning. Our approach integrates efficient matrix factorization and novel regularization techniques, forming a robust framework for dynamic rank pruning and model compression. By using Singular Value Decomposition (SVD) to model low-rank convolutional filters and dense weight matrices, and training the SVD factors with back-propagation in an end-to-end manner, we achieve model compression. We evaluate our method on modern CNNs, including ResNet-18, ResNet-20, and ResNet-32, using datasets like CIFAR-10, CIFAR-100, and ImageNet (2012). Our experiments demonstrate that the proposed method can reduce model parameters by up to 50% and improve classification accuracy by up to 2% over baseline models, making CNNs more feasible for practical applications.
Keyword:
Training
Computational modeling
Convolutional neural networks
Adaptation models
Filters
Tensors
Quantization (signal)
Data models
Vectors
Matrix decomposition
Convolutional neural network
dynamic rank selection
image classification
low-rank factorization
model compression
model pruning

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

R
Rochester Institute of Technology
学者数:
3.8K
论文数: 3.3K
被引数: 45
U
university of texas system
学者数:
18.5W
论文数: 15.6W
被引数: 210
引用论文

引用论文

err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
err分享
err收藏
Representation and compression of Residual Neural Networks through a multilayer network based approach
err2023-04-01
err22
PREAI
errAmelio, Alessia; Bonifazi, Gianluca; Cauteruccio, Francesco; Corradini, Enrico; Marchetti, Michele; Ursino, Domenico; Virgili, Luca
err分享
err收藏
Calcium-modulating cyclophilin ligand regulates membrane trafficking of postsynaptic GABAA receptors
err2008-06-01
err0
errOAAI
errXu Yuan; Jun Yao; David Norris; David D. Tran; Richard J. Bram; Gong Chen; Bernhard Luscher
err分享
err收藏
err分享
err收藏
Axillary arch: Potential cause of neurovascular compression syndrome
err2003-10-10
err0
PREAI
errJ.R. Mérida‐Velasco; J.F. Rodríguez Vázquez; J.A. Mérida Velasco; J. Sobrado Pérez; J. Jiménez Collado
err分享
err收藏
err分享
err收藏
学者 查看更多内容