arrow
返回

Adaptive Mid-Term Representations for Robust Audio Event Classification

delete2018-12-01
delete5
PRE
AI
I
Irene Martín-Morató *
M
Máximo Cobos
F
Francesc J. Ferri
DOI:10.1109/TASLP.2018.2865615delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Low-level audio features are commonly used in many audio analysis tasks, such as audio scene classification or acoustic event detection. Due to the variable length of audio signals, it is a common approach to create fixed-length feature vectors consisting of a set of statistics that summarize the temporal variability of such short-term features. To avoid the loss of temporal information, the audio event can be divided into a set of mid-term segments or texture windows. However, such an approach requires to estimate accurately the onset and offset times of the audio events in order to obtain a robust mid-term statistical description of their temporal evolution. This paper proposes the use of an alternative event representation based on nonlinear time normalization prior to the extraction of mid-term statistics. The short-term features are transformed into a new fixed-length representation that considers uniform distance subsampling over a defined feature space in contrast to the classical short-term temporal framing. The results show that the use of distance-based texture windows provides an improved statistical description of the event robust to errors in the event segmentation stage under noisy conditions.
Keyword:
Audio event classification
support vector machines
trace-segmentation
mid-term statistics
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

I
IEEE-ACM Transactions on Audio Speech and Language Processing
IF:
5.1
论文数:
2.6K
被引数:
1.1W

机构

U
University of Valencia
学者数:
2.5W
论文数: 2.1W
被引数: 24
引用论文

引用论文

err分享
err收藏
Evolution of Floral Nectaries in Iridaceae
err2003-01-01
err0
errOAAI
errPaula J. Rudall; John C. Manning; Peter Goldblatt
err分享
err收藏
err分享
err收藏
学者 查看更多内容