arrow
返回

Toward deep drum source separation

delete2024-07-01
delete1
delete
OA
AI
A
Alessandro Ilic Mezza *
R
Riccardo Giampiccolo
A
Alberto Bernardini
A
Augusto Sarti
DOI:10.1016/j.patrec.2024.04.026delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In the past, the field of drum source separation faced significant challenges due to limited data availability, hindering the adoption of cutting -edge deep learning methods that have found success in other related audio applications. In this letter, we introduce StemGMD, a large-scale audio dataset of isolated single -instrument drum stems. Each audio clip is synthesized from MIDI recordings of expressive drum performances using ten real -sounding acoustic drum kits. Totaling 1224 h, StemGMD is the largest audio dataset of drums to date and the first to comprise isolated audio clips for every instrument in a canonical nine -piece drum kit. We leverage StemGMD to develop LarsNet, a novel deep drum source separation model. Through a bank of dedicated U -Nets, LarsNet can separate five stems from a stereo drum mixture faster than real-time and is shown to considerably outperform state-of-the-art nonnegative spectro-temporal factorization methods.
Keyword:
Deep learning
Drums
Music decomposition
Source separation
U-Net
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
7.9K
被引数:
1.6W

机构

P
Polytechnic University of Milan
学者数:
2.0W
论文数: 1.8W
被引数: 24