返回
Speech feature analysis using variational Bayesian PCA
DOI:10.1109/LSP.2003.810017.png)
摘要
En 中文
In most hidden Markov model-based automatic speech recognition systems, one of the fundamental questions is to determine the intrinsic speech feature dimensionality and the number of clusters used in the Gaussian mixture model. We analyzed mel-frequency band energies using a variational Bayesian principal, component analysis method to estimate the feature dimensionality as well as the number of Gaussian mixtures by learning a maximum lower, bound of the evidence instead of maximizing the likelihood function as used in conventional speech recognition systems. In analyzing the Texas Instruments/Massachusetts Institute of Technology (TIMIT) speech database, our method revealed the intrinsic structures of vowels and consonants. The usefulness of this method is demonstrated in the superior classification performance for the most difficult phonemes /b/, /d/, and /g/.
Keyword:
phoneme classification
speech analysis
speech recognition
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
9.6
论文数:
1.1W
被引数:
1.7W
机构
暂无机构信息
引用论文
Soft computing-based investigation of mechanical properties of concrete using ready-mix concrete waste water as partial replacement of mixing portable water基于软计算方法的混凝土力学性能研究,使用预拌混凝土废水作为拌合饮用水的部分替代品
没有更多内容

