arrow
返回

Sequential routing framework: Fully capsule network-based speech recognition

delete2021-11-01
delete5
delete
OA
AI
K
Kyungmin Lee
H
Hyunwhan Joe
H
Hyeontaek Lim
K
Kwangyoun Kim
S
Sung-Soo Kim
C
Chang Woo Han
H
Hong‐Gee Kim *
DOI:10.1016/j.csl.2021.101228delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Capsule networks (CapsNets) have recently gotten attention as a novel neural architecture. This paper presents the sequential routing framework which we believe is the first method to adapt a CapsNet-only structure to sequence-to-sequence recognition. Input sequences are capsulized then sliced by a window size. Each slice is classified to a label at the corresponding time through iterative routing mechanisms. Afterwards, losses are computed by connectionist temporal classification (CTC). During routing, the required number of parameters can be controlled by the window size regardless of the length of sequences by sharing learnable weights across the slices. We additionally propose a sequential dynamic routing algorithm to replace traditional dynamic routing. The proposed technique can minimize decoding speed degradation caused by the routing iterations since it can operate in a non-iterative manner without dropping accuracy. The method achieves a 1.1% lower word error rate at 16.9% on the Wall Street Journal corpus compared to bidirectional long short-term memory based CTC networks. On the TIMIT corpus, it attains a 0.7% lower phone error rate at 17.5% compared to convolutional neural network-based CTC networks (Zhang et al., 2016). (c) 2021 Elsevier Ltd. All rights reserved.
Keyword:
Capsule network
Automatic speech recognition
Sequence-to-sequence
Connectionist temporal classification
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

C
Computer Speech and Language
IF:
3.4
论文数:
1.5K
被引数:
2.6K

机构

S
samsung
学者数:
8.6K
论文数: 6.4K
被引数: 8
S
seoul national university (snu)
学者数:
7.2W
论文数: 6.6W
被引数: 86
引用论文

引用论文

Attention guided capsule networks for chemical-protein interaction extraction
err2020-03-01
err15
PREAI
errSun, Cong; Yang, Zhihao; Wang, Lei; Zhang, Yin; Lin, Hongfei; Wang, Jian
err分享
err收藏
Polyphonic Sound Event Detection by Using Capsule Neural Networks
err2019-05-01
err39
errOAAI
errVesperini, Fabio; Gabrielli, Leonardo; Principi, Emanuele; Squartini, Stefano
err分享
err收藏
err分享
err收藏
Architecture of the Organic.Edunet Web Portal
err2009-01-01
err0
errOAAI
errNikos Manouselis; Kostas Kastrantas; Salvador Sanchez-Alonso; Jesus Caceres; Hannes Ebner; Matthais Palmer
err分享
err收藏
Deep Neural Networks for Acoustic Modeling in Speech Recognition深度神经网络在语音识别声学建模中的应用
err2012-11-01
err8.2K
PREAI
errHinton, Geoffrey; Deng, Li; Yu, Dong; Dahl, George E.; Mohamed, Abdel-rahman; Jaitly, Navdeep; Senior, Andrew; Vanhoucke, Vincent; Patrick Nguyen; Sainath, Tara N.; Kingsbury, Brian
err分享
err收藏
err分享
err收藏
没有更多内容