arrow
返回

Sparse low rank factorization for deep neural network compression

delete2020-07-01
delete87
PRE
AI
S
Sridhar Swaminathan *
D
Deepak Garg
R
Rajkumar Kannan
F
Frédéric Andrès
DOI:10.1016/j.neucom.2020.02.035delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Storing and processing millions of parameters in deep neural networks is highly challenging during the deployment of model in real-time application on resource constrained devices. Popular low-rank approximation approach singular value decomposition (SVD) is generally applied to the weights of fully connected layers where compact storage is achieved by keeping only the most prominent components of the decomposed matrices. Years of research on pruning-based neural network model compression revealed that the relative importance or contribution of each neuron in a layer highly vary among each other. Recently, synapses pruning has also demonstrated that having sparse matrices in network architecture achieve lower space and faster computation during inference time. We extend these arguments by proposing that the low-rank decomposition of weight matrices should also consider significance of both input as well as output neurons of a layer. Combining the ideas of sparsity and existence of unequal contributions of neurons towards achieving the target, we propose sparse low rank (SLR) method which sparsifies SVD matrices to achieve better compression rate by keeping lower rank for unimportant neurons. We demonstrate the effectiveness of our method in compressing famous convolutional neural networks based image recognition frameworks which are trained on popular datasets. Experimental results show that the proposed approach SLR outperforms vanilla truncated SVD and a pruning baseline, achieving better compression rates with minimal or no loss in the accuracy. Code of the proposed approach is avaialble at https://github.com/sridarah/slr. (C) 2020 Elsevier B.V. All rights reserved.
Keyword:
Low-rank approximation
Singular value decomposition
Sparse matrix
Deep neural networks
Convolutional neural networks
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

N
national institute of informatics (nii) - japan
学者数:
453
论文数: 420
被引数: 0
R
research organization of information & systems (rois)
学者数:
2.8K
论文数: 3.2K
被引数: 2
引用论文

引用论文

err分享
err收藏
Recent advances in convolutional neural network acceleration
err2019-01-01
err245
errOAAI
errZhang, Qianru; Zhang, Meng; Chen, Tinghuan; Sun, Zhifei; Ma, Yuzhe; Yu, Bei
err分享
err收藏
Structured Pruning of Convolutional Neural Networks via L1 Regularization基于L1正则化的卷积神经网络结构修剪
err2019-01-01
err21
errOAAI
errYang, Chen; Yang, Zhenghong; Khattak, Abdul Mateen; Yang, Liu; Zhang, Wenxin; Gao, Wanlin; Wang, Minjuan
err分享
err收藏
Generating Highly Accurate Predictions for Missing QoS Data via Aggregating Nonnegative Latent Factor Models
err2016-03-01
err233
PREAI
errLuo, Xin; Zhou, MengChu; Xia, Yunni; Zhu, Qingsheng; Ammari, Ahmed Chiheb; Alabdulwahab, Ahmed
err分享
err收藏
Targeted Deletion of the Cytosolic Domain of Tissue Factor in Mice Does Not Affect Development
err2001-08-01
err0
PREAI
errEls Melis; Lieve Moons; Maria De Mol; Jean-Marc Herbert; Nigel Mackman; Désiré Collen; Peter Carmeliet; Mieke Dewerchin
err分享
err收藏
学者 查看更多内容