arrow
返回

Generalized Score Comparison-Based Learning Objective for Deep Speaker Embedding

delete2025-01-01
delete0
delete
OA
AI
M
Min Hyun Han
S
Sung Hwan Mun
N
Nam Soo Kim *
DOI:10.1109/ACCESS.2025.3552790delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In state-of-the-art speaker verification systems, speaker embeddings are trained to be closer to the target speaker prototype, which is either obtained from the other speech samples or constructed with trainable parameters. This can be considered a classification task since the network is trying to learn the features that are most relevant to the corresponding speaker from the input speech. Although classification-based learning demonstrates the ability to extract speaker-related information, it does not guarantee optimal speaker verification performance. In this paper, we propose a score comparison-based learning objective, which guides the training framework to be more consistent with the verification task, enforcing the embedding space to have lower intra-class variance compared to inter-class variance in terms of similarity scores. Furthermore, we propose a generalized loss function for score comparison-based learning, encompassing many conventional training losses and regularization techniques. The proposed technique is compared with the conventional methods using the VoxCeleb, VOiCES, CN-Celeb, and Common Voice datasets. Experimental results demonstrate that the proposed method can boost the performance and make the system more robust to over-fitting in speaker verification tasks.
Keyword:
Training
Measurement
Prototypes
Vectors
Feature extraction
Data mining
Object recognition
Kernel
Focusing
Data models
Speaker verification
deep speaker embedding
metric learning
embedding space

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

S
samsung
学者数:
8.6K
论文数: 6.4K
被引数: 8
S
seoul national university (snu)
学者数:
7.2W
论文数: 6.6W
被引数: 86
引用论文

引用论文

暂无论文信息