arrow
返回

Robust Bayesian Pitch Tracking Based on the Harmonic Model

delete2019-11-01
delete23
delete
OA
AI
L
Liming Shi *
J
Jesper Kjær Nielsen
J
Jesper Rindom Jensen
M
Max A. Little
M
Mads Græsbøll Christensen
DOI:10.1109/TASLP.2019.2930917delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Fundamental frequency is one of the most important characteristics of speech and audio signals. Harmonic model-based fundamental frequency estimators offer a higher estimation accuracy and robustness against noise than the widely used autocorrelation-based-methods. However, the traditional harmonic model-based estimators do not take the temporal smoothness of the fundamental frequency, the model order, and the voicing into account as they process each data segment independently. In this paper, a fully Bayesian fundamental frequency tracking algorithm based on the harmonic model and a first-order Markov process model is proposed. Smoothness priors are imposed on the fundamental frequencies, model orders, and voicing using first-order Markov process models. Using these Markov models, fundamental frequency estimation and voicing detection errors can be reduced. Using the harmonic model, the proposed fundamental frequency tracker has an improved robustness to noise. An analytical form of the likelihood function, which can be computed efficiently, is derived. Compared to the state-of-the-art neural network and nonparametric approaches, the proposed fundamental frequency tracking algorithm has superior performance in almost all investigated scenarios, especially in noisy conditions. For example, under 0 dB white Gaussian noise, the proposed algorithm reduces the mean absolute errors and gross errors by 15% and 20% on the Keele pitch database and 36% and 26% on sustained /a/ sounds from a database of Parkinson's disease voices. AMATLAB version of the proposed algorithm is made freely available for reproduction of the results.
Keyword:
Fundamental frequency or pitch tracking
harmonic model
Markov process
harmonic order
voiced-unvoiced detection
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

I
IEEE-ACM Transactions on Audio Speech and Language Processing
IF:
5.1
论文数:
2.6K
被引数:
1.1W

机构

A
Aston University
学者数:
4.8K
论文数: 5.6K
被引数: 8.8K
A
aalborg university
学者数:
1.6W
论文数: 1.7W
被引数: 22
引用论文

引用论文

err分享
err收藏
Shapes and Other Things
err2015-08-22
err0
errOAAI
errTerry Knight
err分享
err收藏
Primary Conjunctival Lymphoma: Response to Chemotherapy in 4 Cases
err2009-02-26
err0
errOAAI
errG. Bellesi; S. di Lollo; A. Bosi; F. Bernardi; G. Campana; S. Romagnani; P. Rossi Ferrini
err分享
err收藏
Fast fundamental frequency estimation: Making a statistically efficient estimator computationally efficient
err2017-06-01
err49
errOAAI
errNielsen, Jesper Kjaer; Jensen, Tobias Lindstrom; Jensen, Jesper Rindom; Christensen, Mads Graesboll; Jensen, Soren Holdt
err分享
err收藏
err分享
err收藏
Rehabilitation of executive dysfunction following brain injury: “Content-free” cueing improves everyday prospective memory performance
err2007-01-01
err0
PREAI
errJessica Fish; Jonathan J. Evans; Morag Nimmo; Emma Martin; Denyse Kersel; Andrew Bateman; Barbara A. Wilson; Tom Manly
err分享
err收藏
Multi-pitch estimation多基音估计
err2008-04-01
err81
errOAAI
errChristensen, Mads Graesboll; Stoica, Petre; Jakobsson, Andreas; Jensen, Soren Holdt
err分享
err收藏
学者 查看更多内容