arrow
Return

A Forced Decoding-based Approach for Enhancing Low-resource ASR

delete2025-05-05
delete0
delete
OA
AI
Y
Yunpeng Liu
X
Xukui Yang
D
Dan Qu *
DOI:10.1007/s11063-025-11759-5delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Multilingual automatic speech recognition represents a crucial research direction in tackling the challenges associated with low-resource scenarios. To effectively incorporate language-specific information in the joint training of a model across multiple languages and leverage linguistic similarities to enhance performance on the target low-resource language, this paper introduces a language similarity evaluation approach based on forced decoding. Specifically, when the target language is specified, the speeches of the source language are decoded into transcription in the target language, and the normalized posterior is utilized as the foundation for evaluating language similarity. Comprehensive experiments and analyses conducted on six low-resource languages reveal that the proposed approach achieves an average word error rate relative reduction of 21.74, 7.68, and 3.45% compared to three widely used benchmark methods, respectively, thereby validating the effectiveness of our approach.
Keywords:
Multilingual ASR
Low-resource languages
Forced decoding
Language similarity

Journal

Neural Processing Letters cover
Neural Processing Letters
IF:
2.8
Papers:
174
Citations:
5.5K

Organization

No organization information available