arrow
返回

Speech feature analysis using variational Bayesian PCA

delete2003-05-01
delete9
PRE
AI
O
Oh‐Wook Kwon
C
Chan, KL
T
Te‐Won Lee
DOI:10.1109/LSP.2003.810017delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In most hidden Markov model-based automatic speech recognition systems, one of the fundamental questions is to determine the intrinsic speech feature dimensionality and the number of clusters used in the Gaussian mixture model. We analyzed mel-frequency band energies using a variational Bayesian principal, component analysis method to estimate the feature dimensionality as well as the number of Gaussian mixtures by learning a maximum lower, bound of the evidence instead of maximizing the likelihood function as used in conventional speech recognition systems. In analyzing the Texas Instruments/Massachusetts Institute of Technology (TIMIT) speech database, our method revealed the intrinsic structures of vowels and consonants. The usefulness of this method is demonstrated in the superior classification performance for the most difficult phonemes /b/, /d/, and /g/.
Keyword:
phoneme classification
speech analysis
speech recognition
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Signal Processing Magazine 封面图
IEEE Signal Processing Magazine
IF:
9.6
论文数:
1.1W
被引数:
1.7W

机构

暂无机构信息
引用论文

引用论文

err
IF0
err
err0
PREAI
err
err分享
err收藏
err分享
err收藏
没有更多内容