Discriminative pronunciation learning for speech recognition for resource scarce languages

ACM DEV '12 Pub Date : 2012-03-11 DOI:10.1145/2160601.2160618

H. Y. Chan, R. Rosenfeld

引用次数: 20

Abstract

In this paper, we describe a method to create speech recognition capability for small vocabularies in resource-scarce languages. By resource-scarce languages, we mean languages that have a small or economically disadvantaged user base which are typically ignored by the commercial world. We use a high-quality well-trained speech recognizer as our baseline to remove the dependence on large audio data for an accurate acoustic model. Using cross-language phoneme mapping, the baseline recognizer effectively recognizes words in our target language. We automate the generation of pronunciations and generate a set of initial pronunciations for each word in the vocabulary. Next, we remove potential conflicts in word recognition by discriminative training.

查看原文本刊更多论文

资源稀缺语言语音识别的辨析语音学习

在本文中，我们描述了一种在资源稀缺的语言中创建小词汇的语音识别能力的方法。所谓资源稀缺语言，我们指的是那些用户基数较小或经济上处于不利地位的语言，这些语言通常被商业世界所忽视。我们使用一个高质量的训练有素的语音识别器作为我们的基线，以消除对大型音频数据的依赖，以获得准确的声学模型。使用跨语言音素映射，基线识别器可以有效地识别目标语言中的单词。我们自动生成发音，并为词汇表中的每个单词生成一组初始发音。接下来，我们通过判别训练消除词识别中的潜在冲突。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

ACM DEV '12

自引率

0.00%

发文量