快速去噪免提语音识别

2008 Hands-Free Speech Communication and Microphone Arrays Pub Date : 2008-05-06 DOI:10.1109/HSCMA.2008.4538706

R. Gomez, J. Even, H. Saruwatari, K. Shikano

{"title":"快速去噪免提语音识别","authors":"R. Gomez, J. Even, H. Saruwatari, K. Shikano","doi":"10.1109/HSCMA.2008.4538706","DOIUrl":null,"url":null,"abstract":"A robust dereverberation technique for real-time hands-free speech recognition application is proposed. Real-time implementation is made possible by avoiding time-consuming blind estimation. Instead, we use the impulse response by effectively identifying the late reflection components of it. Using this information, together with the concept of Spectral Subtraction (SS), we were able to remove the effects of the late reflection of the reverberant signal. After dereverberation, only the effects of the early component is left and used as input to the recognizer. In this method, multi-band SS is used in order to compensate for the error arising from approximation. We also introduced a training strategy to optimize the values of the multi-band coefficients to minimize the error.","PeriodicalId":129827,"journal":{"name":"2008 Hands-Free Speech Communication and Microphone Arrays","volume":"37 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2008-05-06","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"18","resultStr":"{\"title\":\"Fast Dereverberation for Hands-Free Speech Recognition\",\"authors\":\"R. Gomez, J. Even, H. Saruwatari, K. Shikano\",\"doi\":\"10.1109/HSCMA.2008.4538706\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"A robust dereverberation technique for real-time hands-free speech recognition application is proposed. Real-time implementation is made possible by avoiding time-consuming blind estimation. Instead, we use the impulse response by effectively identifying the late reflection components of it. Using this information, together with the concept of Spectral Subtraction (SS), we were able to remove the effects of the late reflection of the reverberant signal. After dereverberation, only the effects of the early component is left and used as input to the recognizer. In this method, multi-band SS is used in order to compensate for the error arising from approximation. We also introduced a training strategy to optimize the values of the multi-band coefficients to minimize the error.\",\"PeriodicalId\":129827,\"journal\":{\"name\":\"2008 Hands-Free Speech Communication and Microphone Arrays\",\"volume\":\"37 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2008-05-06\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"18\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2008 Hands-Free Speech Communication and Microphone Arrays\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/HSCMA.2008.4538706\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2008 Hands-Free Speech Communication and Microphone Arrays","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/HSCMA.2008.4538706","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 18

摘要

提出了一种用于实时免提语音识别的鲁棒去噪技术。通过避免耗时的盲目估计，实时实现成为可能。相反，我们通过有效地识别脉冲响应的后期反射分量来使用脉冲响应。利用这些信息，再加上谱减法(SS)的概念，我们能够消除混响信号后期反射的影响。去噪后，只留下早期分量的效果，并用作识别器的输入。在该方法中，为了补偿由近似引起的误差，采用了多波段SS。我们还引入了一种训练策略来优化多波段系数的值，以最小化误差。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

Fast Dereverberation for Hands-Free Speech Recognition

A robust dereverberation technique for real-time hands-free speech recognition application is proposed. Real-time implementation is made possible by avoiding time-consuming blind estimation. Instead, we use the impulse response by effectively identifying the late reflection components of it. Using this information, together with the concept of Spectral Subtraction (SS), we were able to remove the effects of the late reflection of the reverberant signal. After dereverberation, only the effects of the early component is left and used as input to the recognizer. In this method, multi-band SS is used in order to compensate for the error arising from approximation. We also introduced a training strategy to optimize the values of the multi-band coefficients to minimize the error.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

2008 Hands-Free Speech Communication and Microphone Arrays

自引率

0.00%

发文量