arrow
返回

A framework for generating large-scale microphone array data for machine learning

delete2023-09-25
delete2
delete
OA
AI
A
Adam Kujawski *
A
Art J. R. Pelling
S
Simon Jekosch
E
Ennes Sarradj
DOI:10.1007/s11042-023-16947-wdelete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The use of machine learning for localization of sound sources from microphone array data has increased rapidly in recent years. Newly developed methods are of great value for hearing aids, speech technologies, smart home systems or engineering acoustics. The existence of openly available data is crucial for the comparability and development of new data-driven methods. However, the literature review reveals a lack of openly available datasets, especially for large microphone arrays. This contribution introduces a framework for generation of acoustic data for machine learning. It implements tools for the reproducible random sampling of virtual measurement scenarios. The framework allows computations on multiple machines, which significantly speeds up the process of data generation. Using the framework, an example of a development dataset for sound source characterization with a 64-channel array is given. A containerized environment running the simulation source code is openly available. The presented approach enables the user to calculate large datasets, to store only the features necessary for training, and to share the source code which is needed to reproduce datasets instead of sharing the data itself. This avoids the problem of distributing large datasets and enables reproducible research.
Keyword:
Acoustic source localization (ASL)
Acoustic source characterization (ASC)
Microphone array
Machine learning

期刊

Multimedia Tools and Applications 封面图
Multimedia Tools and Applications
IF:
3
论文数:
1.9W
被引数:
3.2W

机构

T
Technical University of Berlin
学者数:
1.3W
论文数: 1.1W
被引数: 18
引用论文

引用论文

The expression of multidrug resistance protein in human gastrointestinal tract carcinomas
err1998-02-15
err0
errOAAI
errYuji Takebayashi; Shin-ichi Akiyama; Shoji Natsugoe; Shuichi Hokita; Kiyoshi Niwa; Masaki Kitazono; Tomoyuki Sumizawa; Ayako Tani; Tatsuhiko Furukawa; Takashi Aikou
err分享
err收藏
Predictors of compartment syndrome of the foot after fracture of the calcaneus
err2018-03-01
err0
PREAI
errY. H. Park; J. W. Lee; J. Y. Hong; G. W. Choi; H. J. Kim
err分享
err收藏
Deep Audio-Visual Beamforming for Speaker Localization
err2022-01-01
err9
PREAI
errQian, Xinyuan; Zhang, Qiquan; Guan, Guohui; Xue, Wei
err分享
err收藏
Preparation of a novel anti-fouling β-cyclodextrin–PVDF membrane
err2015-01-01
err0
PREAI
errZongxue Yu; Yang Pan; Yi He; Guangyong Zeng; Heng Shi; Haihui Di
err分享
err收藏
Rapid communication: sequencing of the porcine agouti-related protein (AGRP) gene
err2002-05-01
err0
PREAI
errM. J. Halverson; K. J. Donelan; N. H. Granholm; T. M. Cheesbrough; C. A. Westby; D. M. Marshall
err分享
err收藏
学者 查看更多内容