arrow
Return

Tensor Train Recurrent Network Language Model Prediction

delete2025-10-30
delete0
delete
OA
AI
A
Alejandro Murua *
R
Ramchalam Kinattinkara Ramakrishnan
X
Xinlin Li
R
Rui Heng Yang
V
Vahid Partovi Nia
DOI:10.1002/sta4.70116delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Recurrent neural networks (RNN) such as long-short-term memory (LSTM) networks are essential in a multitude of daily tasks such as speech, language, video and multimodal learning. The shift from cloud to edge computation intensifies the need to contain the growth in size of RNNs. Current research on RNN shows that despite the performance obtained on convolutional neural networks (CNN), keeping a good performance in compressed RNNs is still a challenge. This paper shows that by incorporating informative matrix-normal priors on the tensor weights, tensor-compressed LSTM networks can achieve comparable performance to LSTM networks. Most literature on compression focuses on CNNs using matrix product (MPO) operator tensor trains. However, matrix product state (MPS) tensor trains have more attractive features in terms of storage reduction and computing time for prediction. The present work shows that MPS tensor trains should be at the forefront of LSTM network compression through a theoretical analysis and practical experiments on natural language processing (NLP) tasks.
Keywords:
network compression
probabilistic language models
tensor decomposition
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

S
STAT
IF:
0.8
Papers:
58
Citations:
655

Organization

U
universite de montreal
Scholars:
4.6W
Papers: 3.8W
Citations: 46