arrow
返回

Speech enhancement using sparse dictionary learning in wavelet packet transform domain

delete2017-07-01
delete12
PRE
AI
S
Samira Mavaddaty *
S
Seyed Mohammad Ahadi
S
Sanaz Seyedin
DOI:10.1016/j.csl.2017.01.009delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Sparse coding, as a successful representation method for many signals, has been recently employed in speech enhancement. This paper presents a new learning-based speech enhancement algorithm via sparse representation in the wavelet packet transform domain. We propose sparse dictionary learning procedures for training data of speech and noise signals based on a coherence criterion, for each subband of decomposition level. Using these learning algorithms, self-coherence between atoms of each dictionary and mutual coherence between speech and noise dictionary atoms are minimized along with the approximation error. The speech enhancement algorithm is introduced in two scenarios, supervised and semi-supervised. In each scenario, a voice activity detector scheme is employed based on the energy of sparse coefficient matrices when the observation data is coded over corresponding dictionaries. In the proposed supervised scenario, we take advantage of domain adaptation techniques to transform a learned noise dictionary to a dictionary adapted to noise conditions captured based on the test environment circumstances. Using this step, observation data is sparsely coded, based on the current situation of the noisy space, with low sparse approximation error. This technique has a prominent role in obtaining better enhancement results particularly when the noise is non-stationary. In the proposed semi-supervised scenario, adaptive thresholding of wavelet coefficients is carried out based on the variance of the estimated noise in each frame of different subbands. The proposed approaches lead to significantly better speech enhancement results in comparison with the earlier methods in this context and the traditional procedures, based on different objective and subjective measures as well as a statistical test. (C) 2017 Elsevier Ltd. All rights reserved.
Keyword:
Speech enhancement
Dictionary learning
Sparse representation
Domain adaptation
Voice activity detector
Wavelet packet transform
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

C
Computer Speech and Language
IF:
3.4
论文数:
1.5K
被引数:
2.6K

机构

A
Amirkabir University of Technology
学者数:
1.1W
论文数: 1.1W
被引数: 1.0W
引用论文

引用论文

Learning Dictionaries With Bounded Self-Coherence
err2012-12-01
err44
errOAAI
errSigg, Christian D.; Dikk, Tomas; Buhmann, Joachim M.
err分享
err收藏
Least angle regression
err2004-04-01
err7.5K
errOAAI
errEfron, B; Hastie, T; Johnstone, I; Tibshirani, R
err分享
err收藏
Impact of comorbid Sjögren syndrome in anti-aquaporin-4 antibody-positive neuromyelitis optica spectrum disorders
err2021-01-08
err0
PREAI
errTetsuya Akaishi; Toshiyuki Takahashi; Kazuo Fujihara; Tatsuro Misu; Juichi Fujimori; Yoshiki Takai; Shuhei Nishiyama; Michiaki Abe; Tadashi Ishii; Masashi Aoki; Ichiro Nakashima
err分享
err收藏
How pharmacists perceive their professional identity: a scoping review and discursive analysis药剂师如何感知其专业身份:一项范围综述和话语分析
err2021-05-12
err0
errOAAI
errJamie Kellar; Lachmi Singh; Glyneva Bradley-Ridout; Maria Athina Martimianakis; Cees P M van der Vleuten; Mirjam G A oude Egbrink; Zubin Austin
err分享
err收藏
学者 查看更多内容