Automatic Segmentation and Labeling for Mandarin Chinese Speech Corpora for Concatenation-based TTS

Int. J. Comput. Linguistics Chin. Lang. Process. Pub Date : 2005-06-01 DOI:10.30019/IJCLCLP.200507.0001

Chengyuan Lin, J. Jang, Kuan-Ting Chen

引用次数: 25

Abstract

Precise phone/syllable boundary labeling of the utterances in a speech corpus plays an important role in constructing a corpus-based TTS (text-to-speech) system. However, automatic labeling based on Viterbi forced alignment does not always produce satisfactory results. Moreover, a suitable labeling method for one language does not necessarily produce desirable results for another language. Hence in this paper, we propose a new procedure for refining the boundaries of utterances in a Mandarin speech corpus. This procedure employs different sets of acoustic features for four different phonetic categories. In addition, a new scheme is proposed to deal with the ”periodic voiced + periodic voiced” case, which produced most of the segmentation errors in our experiment. Several experiments were conducted to demonstrate the feasibility of the proposed approach.

查看原文本刊更多论文

基于拼接TTS的汉语语音语料库自动分割与标注

语音语料库中语音的精确语音/音节边界标注对于构建基于语料库的文本到语音(TTS)系统至关重要。然而，基于维特比强制对齐的自动标注并不总是产生令人满意的结果。此外，适合一种语言的标注方法不一定对另一种语言产生理想的结果。因此，本文提出了一种新的汉语语音语料库中语音边界的提炼方法。这个过程采用了四种不同语音类别的不同声学特征集。此外，我们还提出了一种新的分割方案来处理“周期浊音+周期浊音”的情况，这是我们实验中产生大部分分割错误的原因。通过实验验证了该方法的可行性。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Int. J. Comput. Linguistics Chin. Lang. Process.

自引率

0.00%

发文量