用于背景音乐分离的具有扩张卷积的多波段多尺度DenseNet

IF 0.3 Q4 ACOUSTICS

Journal of the Acoustical Society of Korea Pub Date : 2019-11-01 DOI:10.7776/ASK.2019.38.6.697

Woon-Haeng Heo, Hyemi Kim, O. Kwon

{"title":"用于背景音乐分离的具有扩张卷积的多波段多尺度DenseNet","authors":"Woon-Haeng Heo, Hyemi Kim, O. Kwon","doi":"10.7776/ASK.2019.38.6.697","DOIUrl":null,"url":null,"abstract":"We propose a multi-band multi-scale DenseNet with dilated convolution that separates background music signals from broadcast content. Dilated convolution can learn the multi-scale context information represented by spectrogram. In computer simulation experiments, the proposed architecture is shown to improve Signal to Distortion Ratio (SDR) by 0.15 dB and 0.27 dB in 0dB and –10 dB Signal to Noise Ratio (SNR) environments, respectively.","PeriodicalId":42689,"journal":{"name":"Journal of the Acoustical Society of Korea","volume":"38 1","pages":"697-702"},"PeriodicalIF":0.3000,"publicationDate":"2019-11-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"Multi-band multi-scale DenseNet with dilated convolution for background music separation\",\"authors\":\"Woon-Haeng Heo, Hyemi Kim, O. Kwon\",\"doi\":\"10.7776/ASK.2019.38.6.697\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"We propose a multi-band multi-scale DenseNet with dilated convolution that separates background music signals from broadcast content. Dilated convolution can learn the multi-scale context information represented by spectrogram. In computer simulation experiments, the proposed architecture is shown to improve Signal to Distortion Ratio (SDR) by 0.15 dB and 0.27 dB in 0dB and –10 dB Signal to Noise Ratio (SNR) environments, respectively.\",\"PeriodicalId\":42689,\"journal\":{\"name\":\"Journal of the Acoustical Society of Korea\",\"volume\":\"38 1\",\"pages\":\"697-702\"},\"PeriodicalIF\":0.3000,\"publicationDate\":\"2019-11-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"Journal of the Acoustical Society of Korea\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.7776/ASK.2019.38.6.697\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"Q4\",\"JCRName\":\"ACOUSTICS\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"Journal of the Acoustical Society of Korea","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.7776/ASK.2019.38.6.697","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q4","JCRName":"ACOUSTICS","Score":null,"Total":0}

引用次数: 0

摘要

我们提出了一种多频带多尺度的扩展卷积DenseNet，用于从广播内容中分离背景音乐信号。展开卷积可以学习由谱图表示的多尺度上下文信息。在计算机仿真实验中，该结构在0dB和-10 dB信噪比(SNR)环境下分别提高了0.15 dB和0.27 dB的信失真比(SDR)。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

Multi-band multi-scale DenseNet with dilated convolution for background music separation

We propose a multi-band multi-scale DenseNet with dilated convolution that separates background music signals from broadcast content. Dilated convolution can learn the multi-scale context information represented by spectrogram. In computer simulation experiments, the proposed architecture is shown to improve Signal to Distortion Ratio (SDR) by 0.15 dB and 0.27 dB in 0dB and –10 dB Signal to Noise Ratio (SNR) environments, respectively.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

Journal of the Acoustical Society of Korea ACOUSTICS-

CiteScore

0.60

自引率

50.00%

发文量