校正双耳比:用于稳健声音定位的复杂t分布特征

2016 24th European Signal Processing Conference (EUSIPCO) Pub Date : 2016-08-29 DOI:10.1109/EUSIPCO.2016.7760450

Antoine Deleforge, F. Forbes

{"title":"校正双耳比:用于稳健声音定位的复杂t分布特征","authors":"Antoine Deleforge, F. Forbes","doi":"10.1109/EUSIPCO.2016.7760450","DOIUrl":null,"url":null,"abstract":"Most existing methods in binaural sound source localization rely on some kind of aggregation of phase- and level-difference cues in the time-frequency plane. While different aggregation schemes exist, they are often heuristic and suffer in adverse noise conditions. In this paper, we introduce the rectified binaural ratio as a new feature for sound source localization. We show that for Gaussian-process point source signals corrupted by stationary Gaussian noise, this ratio follows a complex t-distribution with explicit parameters. This new formulation provides a principled and statistically sound way to aggregate binaural features in the presence of noise. We subsequently derive two simple and efficient methods for robust relative transfer function and time-delay estimation. Experiments on heavily corrupted simulated and speech signals demonstrate the robustness of the proposed scheme.","PeriodicalId":127068,"journal":{"name":"2016 24th European Signal Processing Conference (EUSIPCO)","volume":"39 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2016-08-29","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"4","resultStr":"{\"title\":\"Rectified binaural ratio: A complex T-distributed feature for robust sound localization\",\"authors\":\"Antoine Deleforge, F. Forbes\",\"doi\":\"10.1109/EUSIPCO.2016.7760450\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Most existing methods in binaural sound source localization rely on some kind of aggregation of phase- and level-difference cues in the time-frequency plane. While different aggregation schemes exist, they are often heuristic and suffer in adverse noise conditions. In this paper, we introduce the rectified binaural ratio as a new feature for sound source localization. We show that for Gaussian-process point source signals corrupted by stationary Gaussian noise, this ratio follows a complex t-distribution with explicit parameters. This new formulation provides a principled and statistically sound way to aggregate binaural features in the presence of noise. We subsequently derive two simple and efficient methods for robust relative transfer function and time-delay estimation. Experiments on heavily corrupted simulated and speech signals demonstrate the robustness of the proposed scheme.\",\"PeriodicalId\":127068,\"journal\":{\"name\":\"2016 24th European Signal Processing Conference (EUSIPCO)\",\"volume\":\"39 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2016-08-29\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"4\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2016 24th European Signal Processing Conference (EUSIPCO)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/EUSIPCO.2016.7760450\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2016 24th European Signal Processing Conference (EUSIPCO)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/EUSIPCO.2016.7760450","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 4

摘要

大多数现有的双耳声源定位方法依赖于时频平面上的某种相位和电平差线索的聚集。虽然存在不同的聚合方案，但它们往往是启发式的，并且受到不利噪声条件的影响。本文引入了校正双耳比作为声源定位的一种新特征。我们表明，对于被平稳高斯噪声破坏的高斯过程点源信号，该比率遵循具有显式参数的复t分布。这个新公式提供了一个原则和统计上合理的方式来聚集双耳特征在存在噪声。在此基础上推导了两种简单有效的鲁棒相对传递函数和时延估计方法。对严重损坏的模拟信号和语音信号的实验证明了该方法的鲁棒性。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

Rectified binaural ratio: A complex T-distributed feature for robust sound localization

Most existing methods in binaural sound source localization rely on some kind of aggregation of phase- and level-difference cues in the time-frequency plane. While different aggregation schemes exist, they are often heuristic and suffer in adverse noise conditions. In this paper, we introduce the rectified binaural ratio as a new feature for sound source localization. We show that for Gaussian-process point source signals corrupted by stationary Gaussian noise, this ratio follows a complex t-distribution with explicit parameters. This new formulation provides a principled and statistically sound way to aggregate binaural features in the presence of noise. We subsequently derive two simple and efficient methods for robust relative transfer function and time-delay estimation. Experiments on heavily corrupted simulated and speech signals demonstrate the robustness of the proposed scheme.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

2016 24th European Signal Processing Conference (EUSIPCO)

自引率

0.00%

发文量