Perceptual Linear Predictive (PLP) Analysis-Resynthesis Technique

Final Program and Paper Summaries 1991 IEEE ASSP Workshop on Applications of Signal Processing to Audio and Acoustics Pub Date : 1991-09-24 DOI:10.21437/Eurospeech.1991-88

H. Hermansky, L. Cox

引用次数: 37

Abstract

A common wisdom in speech re-synthesis is that while the vocal tract excitation can be modified to represent the message prosody, the accurate preservation of the formants is needed in order to ensure that both the linguistic message and the speaker-dependent information is well represented in the synthesized speech. Formants are speaker-dependent. A further decomposition of the formant-based speech representation into its message-bearing and the speaker-dependent parts and the inverse problem of combining those two sources of speech information is of interest. The current paper addresses this issues.

查看原文本刊更多论文

感知线性预测(PLP)分析-再合成技术

在语音重新合成中，一个普遍的观点是，虽然可以修改声道激发来表示信息韵律，但需要准确地保留共振峰，以确保在合成的语音中既能很好地表示语言信息，也能很好地表示与说话人相关的信息。共振峰依赖于说话者。将基于共振峰的语音表示进一步分解为其消息承载部分和依赖于说话人的部分，以及将这两个语音信息源结合起来的逆问题是我们感兴趣的。本文解决了这个问题。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Final Program and Paper Summaries 1991 IEEE ASSP Workshop on Applications of Signal Processing to Audio and Acoustics

自引率

0.00%

发文量