arrow
Return

Nonlinear tensor train format for deep neural network compression

delete2021-12-01
delete21
PRE
AI
D
Dingheng Wang
G
Guangshe Zhao *
H
Hengnu Chen
Z
Zhexian Liu
邓
邓磊 (Lei Deng)
Guoqi Li cover
Guoqi Li (Guoqi Li) *
DOI:10.1016/j.neunet.2021.08.028delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Deep neural network (DNN) compression has become a hot topic in the research of deep learning since the scale of modern DNNs turns into too huge to implement on practical resource constrained platforms such as embedded devices. Among variant compression methods, tensor decomposition appears to be a relatively simple and efficient strategy owing to its solid mathematical foundations and regular data structure. Generally, tensorizing neural weights into higher-order tensors for better decomposition, and directly mapping efficient tensor structure to neural architecture with nonlinear activation functions, are the two most common ways. However, the considerable accuracy loss is still a fly in the ointment for the tensorizing way especially for convolutional neural networks (CNNs), while the number of studies in the mapping way is comparatively limited and corresponding compression ratio appears to be not considerable. Therefore, in this work, by researching multiple types of tensor decompositions, we realize that tensor train (TT), which has specific and efficient sequenced contractions, is potential to take into account both of tensorizing and mapping ways. Then we propose a novel nonlinear tensor train (NTT) format, which contains extra nonlinear activation functions embedded in sequenced contractions and convolutions on the top of the normal TT decomposition and the proposed TT format connected by convolutions, to compensate the accuracy loss that normal TT cannot give. Further than just shrinking the space complexity of original weight matrices and convolutional kernels, we prove that NTT can afford an efficient inference time as well. Extensive experiments and discussions demonstrate that the compressed DNNs in our NTT format can almost maintain the accuracy at least on MNIST, UCF11 and CIFAR-10 datasets, and the accuracy loss caused by normal TT could be compensated significantly on large-scale datasets such as ImageNet. (C) 2021 Published by Elsevier Ltd.
Keywords:
Tensor train decomposition
Nonlinear tensor train
Sequenced contractions
Sequenced convolutions
Neural network compression
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neural Networks cover
Neural Networks
IF:
6.3
Papers:
8.2K
Citations:
3.0W

Organization

X
xi'an jiaotong university
Scholars:
9.3W
Papers: 6.7W
Citations: 75
T
tsinghua university
Scholars:
11.9W
Papers: 10.0W
Citations: 137
Cited Papers

Cited Papers

Body Pose Prediction Based on Motion Sensor Data and Recurrent Neural Network
err2021-03-01
err51
PREAI
errWozniak, Marcin; Wieczorek, Michal; Silka, Jakub; Polap, Dawid
errShare
errSave
Compressing 3DCNNs based on tensor train decomposition
err2020-11-01
err22
errOAAI
errWang, Dingheng; Zhao, Guangshe; Li, Guoqi; Deng, Lei; Wu, Yang
errShare
errSave
Functional Diversity of Microbial Communities in Soils in the Vicinity of Wanda Glacier, Antarctic Peninsula
err2012-01-01
err0
errOAAI
errIgor Stelmach Pessi; Susana de Oliveira Elias; Felipe Lorenz Simões; Jefferson Cardia Simões; Alexandre José Macedo
errShare
errSave
Recurrent Neural Network Model for IoT and Networking Malware Threat Detection
err2021-08-01
err73
PREAI
errWozniak, Marcin; Silka, Jakub; Wieczorek, Michal; Alrashoud, Mubarak
errShare
errSave
DecomVQANet: Decomposing visual question answering deep network via tensor decomposition and regression
err2021-02-01
err58
PREAI
errBai, Zongwen; Li, Ying; Wozniak, Marcin; Zhou, Meili; Li, Di
errShare
errSave
researcher View more