arrow
返回

Suppression by Selecting Wavelets for Feature Compression in Distributed Speech Recognition

delete2018-03-01
delete10
PRE
AI
S
Syu‐Siang Wang
P
Payton Lin
Y
Yu Tsao *
J
Jeih-weih Hung
B
Borching Su
DOI:10.1109/TASLP.2017.2779787delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Distributed speech recognition (DSR) splits the processing of data between amobile device and a network server. In the front-end, features are extracted and compressed to transmit over a wireless channel to a back-end server, where the incoming stream is received and reconstructed for recognition tasks. In this paper, we propose a feature compression algorithm termed suppression by selecting wavelets (SSW) to achieve the two main goals of DSR: Minimizingmemory and device requirements while also maintaining or even improving the recognition performance. The SSW approach first applies the discrete wavelet transform (DWT) to filter the incoming speech feature sequence into two temporal subsequences at the client terminal. Feature compression is achieved by keeping the low (modulation) frequency subsequence while discarding the high frequency counterpart. The low-frequency subsequence is then transmitted across the remote network for specific feature statistics normalization. Wavelets are favorable for resolving the temporal properties of the feature sequence, and the down-sampling process in DWT achieves data compression by reducing the amount of data at the terminal prior to transmission across the network. Once the compressed features have arrived at the server, the feature sequence can be enhanced by statistics normalization, reconstructed with inverse DWT, and compensated with a simple post filter to alleviate any over-smoothing effects from the compression stage. Results on a standard robustness task (Aurora-4) and on a Mandarin Chinese news corpus showed SSW outperforms conventional noise-robustness techniques while also providing nearly a 50% compression rate during the transmission stage of DSR systems.
Keyword:
Discrete wavelet transform
feature compression
distributed speech recognition
data transmission efficiency
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

I
IEEE-ACM Transactions on Audio Speech and Language Processing
IF:
5.1
论文数:
2.6K
被引数:
1.1W

机构

A
academia sinica - taiwan
学者数:
1.9W
论文数: 1.6W
被引数: 17
N
National Taiwan University
学者数:
4.7W
论文数: 4.2W
被引数: 3.6W
N
National Chi Nan University
学者数:
1.4K
论文数: 1.3K
被引数: 647
学者 查看更多机构
引用论文

引用论文

An Overview of Noise-Robust Automatic Speech Recognition
err2014-04-01
err421
PREAI
errLi, Jinyu; Deng, Li; Gong, Yifan; Haeb-Umbach, Reinhold
err分享
err收藏
err分享
err收藏
Mechanisms of Hemostimulating Effect of Aconitum baicalense Diterpene Alkaloids川贝乌头二萜生物碱促血机制研究
err2013-07-19
err0
PREAI
errG. N. Zyuz’kov; V. V. Zhdanov; L. A. Miroshnichenko; E. V. Udut; E. V. Simanina; L. A. Stavrova; V. I. Agafonov; A. V. Chaikovskiy; M. Yu. Minakova; T. N. Povet’eva; N. I. Suslov; A. V. Krapivin; Yu. V. Nesterova; A. A. Semenov; D. V. Reykhart; A. M. Dygai
err分享
err收藏
err分享
err收藏
err分享
err收藏
Improved modulation spectrum enhancement methods for robust speech recognition
err2012-11-01
err16
PREAI
errHung, Jeih-weih; Tu, Wen-hsiang; Lai, Chien-chou
err分享
err收藏
学者 查看更多内容