arrow
返回

Content-based audio classification and segmentation by using support vector machines

delete2003-04-01
delete167
PRE
AI
L
Lie Lu
Z
Zhang, HJ
S
Stan Z. Li
DOI:10.1007/s00530-002-0065-0delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Content-based audio classification and segmentation is a basis for further audio/video analysis. In this paper, we present our work on audio segmentation and classification which employs support vector machines (SVMs). Five audio classes are considered in this paper: silence, music, background sound, pure speech, and non-pure speech which includes speech over music and speech over noise. A sound stream is segmented by classifying each sub-segment into one of these five classes. We have evaluated the performance of SVM on different audio type-pairs classification with testing unit of different-length and compared the performance of SVM, K-Nearest Neighbor (KNN), and Gaussian Mixture Model (GMM). We also evaluated the effectiveness of some new proposed features. Experiments on a database composed of about 4-hour audio data show that the proposed classifier is very efficient on audio classification and segmentation. It also shows the accuracy of the SVM-based method is much better than the method based on KNN and GMM.
Keyword:
audio content analysis
audio classification and segmentation
support vector machines

期刊

Multimedia Systems 封面图
Multimedia Systems
IF:
3.1
论文数:
2.8K
被引数:
2.7K

机构

暂无机构信息
引用论文

引用论文

暂无论文信息