The Sogou System for Blizzard Challenge 2020

Joint Workshop for the Blizzard Challenge and Voice Conversion Challenge 2020 Pub Date : 2020-10-30 DOI:10.21437/vcc_bc.2020-8

Fanbo Meng, Ruimin Wang, Peng Fang, Shuangyuan Zou, Wenjun Duan, Ming Zhou, Kai Liu, Wei Chen

引用次数: 1

Abstract

In this paper, we introduce the text-to-speech system from Sogou team submitted to Blizzard Challenge 2020. The goal of this year’s challenge is to build a natural Mandarin Chinese speech synthesis system from the 10-hours corpus by a native Chinese male speaker. We will discuss the major modules of the submitted system: (1) the front-end module to analyze the pronunciation and prosody of text; (2) the FastSpeech-based sequence-to-sequence acoustic model to predict acoustic features; (3) the WaveRNN based neural vocoder to reconstruct waveforms. Evaluation results provided by the challenge organizer are also discussed

查看原文本刊更多论文

2020暴雪挑战赛的搜狗系统

在本文中，我们介绍了搜狗团队提交给暴雪挑战赛2020的文本转语音系统。今年挑战赛的目标是由一名母语为汉语的男性从10小时的语料库中构建一个自然的汉语语音合成系统。我们将讨论提交系统的主要模块:(1)前端模块对文本进行语音韵律分析;(2)基于fastspeech的序列对序列声学模型预测声学特征;(3)基于WaveRNN的神经声码器重构波形。讨论了挑战赛组织者提供的评估结果

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Joint Workshop for the Blizzard Challenge and Voice Conversion Challenge 2020

自引率

0.00%

发文量