arrow
Return

Block-term tensor neural networks

delete2020-10-01
delete23
delete
OA
AI
J
Jinmian Ye
G
Guangxi Li
D
Di Chen
H
Haiqin Yang
S
Shandian Zhe
Z
Zenglin Xu *
DOI:10.1016/j.neunet.2020.05.034delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Deep neural networks (DNNs) have achieved outstanding performance in a wide range of applications, e.g., image classification, natural language processing, etc. Despite the good performance, the huge number of parameters in DNNs brings challenges to efficient training of DNNs and also their deployment in low-end devices with limited computing resources. In this paper, we explore the correlations in the weight matrices, and approximate the weight matrices with the low-rank block term tensors. We name the new corresponding structure as block-term tensor layers (BT-layers), which can be easily adapted to neural network models, such as CNNs and RNNs. In particular, the inputs and the outputs in BT-layers are reshaped into low-dimensional high-order tensors with a similar or improved representation power. Sufficient experiments have demonstrated that BT-layers in CNNs and RNNs can achieve a very large compression ratio on the number of parameters while preserving or improving the representation power of the original DNNs. (C) 2020 Elsevier Ltd. All rights reserved.
Keywords:
Tensor networks
Network compression
Neural networks
Deep learning
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neural Networks cover
Neural Networks
IF:
6.3
Papers:
8.2K
Citations:
3.0W

Organization

U
University of Utah
Scholars:
3.0W
Papers: 2.2W
Citations: 4.6W
U
university of technology sydney
Scholars:
1.6W
Papers: 2.0W
Citations: 25
U
Utah System of Higher Education
Scholars:
4.6W
Papers: 4.0W
Citations: 161
researcher View more organizations
Cited Papers

Cited Papers

Deep learning in neural networks: An overview
err2015-01-01
err1.3W
errOAAI
errSchmidhuber, Juergen
errShare
errSave
Accelerating deep neural network training with inconsistent stochastic gradient descent
err2017-09-01
err73
PREAI
errWang, Linnan; Yang, Yi; Min, Renqiang; Chakradhar, Srimat
errShare
errSave
Gradient-based learning applied to document recognition
err1998-01-01
err3.8W
PREAI
errLecun, Y; Bottou, L; Bengio, Y; Haffner, P
errShare
errSave
TensorD: A tensor decomposition library in TensorFlow
err2018-11-01
err21
PREAI
errHao, Liyang; Liang, Siqi; Ye, Jinmian; Xu, Zenglin
errShare
errSave
researcher View more