返回
Sequential routing framework: Fully capsule network-based speech recognition
DOI:10.1016/j.csl.2021.101228.png)
摘要
En 中文
Capsule networks (CapsNets) have recently gotten attention as a novel neural architecture. This paper presents the sequential routing framework which we believe is the first method to adapt a CapsNet-only structure to sequence-to-sequence recognition. Input sequences are capsulized then sliced by a window size. Each slice is classified to a label at the corresponding time through iterative routing mechanisms. Afterwards, losses are computed by connectionist temporal classification (CTC). During routing, the required number of parameters can be controlled by the window size regardless of the length of sequences by sharing learnable weights across the slices. We additionally propose a sequential dynamic routing algorithm to replace traditional dynamic routing. The proposed technique can minimize decoding speed degradation caused by the routing iterations since it can operate in a non-iterative manner without dropping accuracy. The method achieves a 1.1% lower word error rate at 16.9% on the Wall Street Journal corpus compared to bidirectional long short-term memory based CTC networks. On the TIMIT corpus, it attains a 0.7% lower phone error rate at 17.5% compared to convolutional neural network-based CTC networks (Zhang et al., 2016). (c) 2021 Elsevier Ltd. All rights reserved.
Keyword:
Capsule network
Automatic speech recognition
Sequence-to-sequence
Connectionist temporal classification
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
C
IF:
3.4
论文数:
1.5K
被引数:
2.6K
机构
引用论文
Value of the Student Pharmacist to Experiential Practice Sites: A Review of the Literature实习药师对实践基地的价值:文献综述
没有更多内容

