arrow
Return

Data-Driven and Feedback Based Spectro-Temporal Features for Speech Recognition

delete2010-11-01
delete14
PRE
AI
G
G. S. V. S. Sivaram *
S
Sridhar Krishna Nemala
N
Nima Mesgarani
H
Hynek Heřmanský
DOI:10.1109/LSP.2010.2079930delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
This paper proposes novel data-driven and feedback based discriminative spectro-temporal filters for feature extraction in automatic speech recognition (ASR). Initially a first set of spectro-temporal filters are designed to separate each phoneme from the rest of the phonemes. A hybrid Hidden Markov Model/Multilayer Perceptron (HMM/MLP) phoneme recognition system is trained on the features derived using these filters. As a feedback to the feature extraction stage, top confusions of this system are identified, and a second set of filters are designed specifically to address these confusions. Phoneme recognition experiments on TIMIT show that the features derived from the combined set of discriminative filters outperform conventional speech recognition features, and also contain significant complementary information.
Keywords:
Confusion analysis
discriminative filters
spectro-temporal features
speech recognition
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Signal Processing Magazine cover
IEEE Signal Processing Magazine
IF:
9.6
Papers:
1.1W
Citations:
1.7W

Organization

J
Johns Hopkins University
Scholars:
10.2W
Papers: 8.8W
Citations: 13.0W