arrow
Return

Methods for Combining Time-Frequency Representations: A Python Package

delete2026-05-01
delete0
PRE
AI
B
Boechat, Bernardo A. *
D
da Costa, Mauricio do V. M.
B
Biscainho, Luiz W. P.
DOI:10.17743/jaes.2022.0266delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
This paper presents the main ideas behind ctfr, an extensible, user-friendly Python package for efficiently combining time-frequency representations (TFRs) of audio signals into a single representation that captures the best aspects of each, achieving high resolutions in both time and frequency. The authors develop and evaluate algorithmic tweaks and approximation schemes for existing TFR combination methods, with significant performance improvements over baseline implementations. In addition, combined TFRs are employed in training a deep learning system for note transcription from audio performances, showing improved results over traditional TFRs, thus demonstrating the effectiveness of using combination methods in audio processing and music information retrieval pipelines.
Keywords:
audio signal processing
combined time-frequency representation
constant Q transform
music information retrieval
short-time Fourier transform

Journal

J
Journal of the Audio Engineering Society
IF:
0
Papers:
33
Citations:
0

Organization

U
Universidade Federal do Rio de Janeiro
Scholars:
2.9W
Papers: 1.8W
Citations: 1.6W
U
University Osnabruck
Scholars:
3.0K
Papers: 2.6K
Citations: 15