使用隐马尔可夫模型识别说话人

ICSP '98. 1998 Fourth International Conference on Signal Processing (Cat. No.98TH8344) Pub Date : 1998-10-12 DOI:10.1109/ICOSP.1998.770285

M. Inman, D. Danforth, S. Hangai, K. Sato

{"title":"使用隐马尔可夫模型识别说话人","authors":"M. Inman, D. Danforth, S. Hangai, K. Sato","doi":"10.1109/ICOSP.1998.770285","DOIUrl":null,"url":null,"abstract":"In this study, we show that the use of hidden Markov models (HMMs) significantly enhances the success rate of speaker identification over time. The segment boundary information derived from HMMs provides a means of normalizing the formant patterns obtained from a digital cochlear filter, which we also describe. The use of the digital cochlear filter and HMMs in our study was motivated by two well-known problems in speech recognition generally, i.e. phonetic tempo variability and variability over temporal units of a given length, typically days. We show how these problems can be minimized to achieve more robust speaker identification.","PeriodicalId":145700,"journal":{"name":"ICSP '98. 1998 Fourth International Conference on Signal Processing (Cat. No.98TH8344)","volume":"13 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"1998-10-12","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"10","resultStr":"{\"title\":\"Speaker identification using hidden Markov models\",\"authors\":\"M. Inman, D. Danforth, S. Hangai, K. Sato\",\"doi\":\"10.1109/ICOSP.1998.770285\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"In this study, we show that the use of hidden Markov models (HMMs) significantly enhances the success rate of speaker identification over time. The segment boundary information derived from HMMs provides a means of normalizing the formant patterns obtained from a digital cochlear filter, which we also describe. The use of the digital cochlear filter and HMMs in our study was motivated by two well-known problems in speech recognition generally, i.e. phonetic tempo variability and variability over temporal units of a given length, typically days. We show how these problems can be minimized to achieve more robust speaker identification.\",\"PeriodicalId\":145700,\"journal\":{\"name\":\"ICSP '98. 1998 Fourth International Conference on Signal Processing (Cat. No.98TH8344)\",\"volume\":\"13 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"1998-10-12\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"10\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"ICSP '98. 1998 Fourth International Conference on Signal Processing (Cat. No.98TH8344)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/ICOSP.1998.770285\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"ICSP '98. 1998 Fourth International Conference on Signal Processing (Cat. No.98TH8344)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ICOSP.1998.770285","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 10

摘要

在这项研究中，我们发现随着时间的推移，隐马尔可夫模型(hmm)的使用显著提高了说话人识别的成功率。从hmm中得到的段边界信息提供了一种归一化从数字耳蜗滤波器中得到的峰模式的方法，我们也描述了这一点。在我们的研究中，数字耳蜗滤波器和hmm的使用是由语音识别中两个众所周知的问题所驱动的，即语音节奏变异性和给定长度(通常是天)的时间单位的变异性。我们展示了如何将这些问题最小化以实现更稳健的说话人识别。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

Speaker identification using hidden Markov models

In this study, we show that the use of hidden Markov models (HMMs) significantly enhances the success rate of speaker identification over time. The segment boundary information derived from HMMs provides a means of normalizing the formant patterns obtained from a digital cochlear filter, which we also describe. The use of the digital cochlear filter and HMMs in our study was motivated by two well-known problems in speech recognition generally, i.e. phonetic tempo variability and variability over temporal units of a given length, typically days. We show how these problems can be minimized to achieve more robust speaker identification.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

ICSP '98. 1998 Fourth International Conference on Signal Processing (Cat. No.98TH8344)

自引率

0.00%

发文量