返回
Speech enhancement using sparse dictionary learning in wavelet packet transform domain
DOI:10.1016/j.csl.2017.01.009.png)
摘要
En 中文
Sparse coding, as a successful representation method for many signals, has been recently employed in speech enhancement. This paper presents a new learning-based speech enhancement algorithm via sparse representation in the wavelet packet transform domain. We propose sparse dictionary learning procedures for training data of speech and noise signals based on a coherence criterion, for each subband of decomposition level. Using these learning algorithms, self-coherence between atoms of each dictionary and mutual coherence between speech and noise dictionary atoms are minimized along with the approximation error. The speech enhancement algorithm is introduced in two scenarios, supervised and semi-supervised. In each scenario, a voice activity detector scheme is employed based on the energy of sparse coefficient matrices when the observation data is coded over corresponding dictionaries. In the proposed supervised scenario, we take advantage of domain adaptation techniques to transform a learned noise dictionary to a dictionary adapted to noise conditions captured based on the test environment circumstances. Using this step, observation data is sparsely coded, based on the current situation of the noisy space, with low sparse approximation error. This technique has a prominent role in obtaining better enhancement results particularly when the noise is non-stationary. In the proposed semi-supervised scenario, adaptive thresholding of wavelet coefficients is carried out based on the variance of the estimated noise in each frame of different subbands. The proposed approaches lead to significantly better speech enhancement results in comparison with the earlier methods in this context and the traditional procedures, based on different objective and subjective measures as well as a statistical test. (C) 2017 Elsevier Ltd. All rights reserved.
Keyword:
Speech enhancement
Dictionary learning
Sparse representation
Domain adaptation
Voice activity detector
Wavelet packet transform
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
C
IF:
3.4
论文数:
1.5K
被引数:
2.6K
机构
引用论文
The future state of patient engagement? Personal health information use, attitudes towards health, and health behavior患者参与的未来状态?个人健康信息使用、对健康的看法以及健康行为
Use of situational judgment tests for assessing non‐cognitive attributes of final year dental students情景判断测试在评估应届牙科学生非认知属性中的应用
A micromachined efficient parametric array loudspeaker with a wide radiation frequency band具有宽辐射频带的微机械高效参量阵列扬声器
How pharmacists perceive their professional identity: a scoping review and discursive analysis药剂师如何感知其专业身份:一项范围综述和话语分析

