1
Return

Example-Based Framework for Perceptually Guided Audio Texture Generation

delete2024-01-01
delete0
delete
OA
AI
P
Purnima Kamath *
C
Chitralekha Gupta
L
Lonce Wyse
S
Suranga Nanayakkara
DOI:10.1109/TASLP.2024.3393741delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Controllable generation in StyleGANs is usually achieved by training the model using labeled data. For audio textures, however, there is currently a lack of large semantically labeled datasets. Therefore, to control generation, we develop a method for semantic control over an unconditionally trained StyleGAN in the absence of such labeled datasets. In this paper, we propose an example-based framework to determine guidance vectors for audio texture generation based on user-defined semantic attributes. Our approach leverages the semantically disentangled latent space of an unconditionally trained StyleGAN. By using a few synthetic examples to indicate the presence or absence of a semantic attribute, we infer the guidance vectors in the latent space of the StyleGAN to control that attribute during generation. Our results show that our framework can find user-defined and perceptually relevant guidance vectors for controllable generation for audio textures. Furthermore, we demonstrate an application of our framework to other tasks, such as selective semantic attribute transfer.
Keywords:
Vectors
Semantics
Aerospace electronics
Training
Generative adversarial networks
Controllability
Feature extraction
Audio textures
controllability
analysis-by-synthesis
gaver sounds
stylegan
latent space exploration

Journal

I
IEEE-ACM Transactions on Audio Speech and Language Processing
IF:
5.1
Papers:
2.6K
Citations:
1.1W

Organization

P
Pompeu Fabra University
Scholars:
9.3K
Papers: 6.8K
Citations: 11
N
National University of Singapore
Scholars:
7.4W
Papers: 6.4W
Citations: 11.4W
Cited Papers

Cited Papers

Citing Papers

Citing Papers