arrow
返回

Parametric Spatial Sound Processing

delete2015-03-01
delete42
PRE
AI
K
Konrad Kowalczyk *
O
Oliver Thiergart
M
Maja Taseska
G
Giovanni Del Galdo
V
Ville Pulkki
E
Emanuël A. P. Habets
DOI:10.1109/MSP.2014.2369531delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Flexible and efficient spatial sound acquisition and subsequent processing are of paramount importance in communication and assisted listening devices such as mobile phones, hearing aids, smart TVs, and emerging wearable devices (e.g., smart watches and glasses). In application scenarios where the number of sound sources quickly varies, sources move, and nonstationary noise and reverberation are commonly encountered, it remains a challenge to capture sounds in such a way that they can be reproduced with a high and invariable sound quality. In addition, the objective in terms of what needs to be captured, and how it should be reproduced, depends on the application and on the user's preferences. Parametric spatial sound processing has been around for two decades and provides a flexible and efficient solution to capture, code, and transmit, as well as manipulate and reproduce spatial sounds. Instrumental to this type of processing is a parametric model that can describe a sound field in a compact and general way. In most cases, the sound field can be decomposed into a direct sound component and a diffuse sound component. These two components together with parametric side information such as the direction-of-arrival (DOA) of the direct sound component or the position of the sound source, provide a perceptually motivated description of the acoustic scene [1]-[3]. In this article, we provide an overview of recent advances in spatial sound capturing, manipulation, and reproduction based on such parametric descriptions of the sound field. In particular, we focus on two established parametric descriptions presented in a unified way and show how the signals and parameters can be obtained using multiple microphones. Once the sound field is analyzed, the sound scene can be transmitted, manipulated, and synthesized depending on the application. For example, sounds can be extracted from a specific direction or from a specific arbitrary two-dimensional or even three-dimensional region of interest. Furthermore, the sound scene can be manipulated to create an acoustic zoom effect in which direct sounds within the listening angular range are amplified depending on the zoom factor, while other sounds are suppressed. In addition, the signals and parameters can be used to create surround sound signals. As the manipulation and synthesis are highly application dependent, we focus in this article on three illustrative assisted listening applications: spatial audio communication, virtual classroom, and binaural hearing aids.
Keyword:
DIFFUSE SOUND
REPRODUCTION
COHERENCE
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Signal Processing Magazine 封面图
IEEE Signal Processing Magazine
IF:
9.6
论文数:
1.1W
被引数:
1.7W

机构

A
Aalto University
学者数:
1.6W
论文数: 1.5W
被引数: 2.1W
U
University of Erlangen Nuremberg
学者数:
3.2W
论文数: 2.6W
被引数: 29
F
fraunhofer gesellschaft
学者数:
1.6W
论文数: 1.2W
被引数: 24
学者 查看更多机构
引用论文

引用论文

The petrology and paragenesis of fracture mineralization in the Sellafield area, west Cumbria
err2022-06-06
err0
PREAI
errA. E. Milodowski; M. R. Gillespie; J. Naden; N. J. Fortey; T. J. Shepherd; J. M. Pearce; R. Metcalfe
err分享
err收藏
err分享
err收藏
学者 查看更多内容