Content-based music retrieval using linear scaling and branch-and-bound tree search

IEEE International Conference on Multimedia and Expo, 2001. ICME 2001. Pub Date : 2001-08-22 DOI:10.1109/ICME.2001.1237713

J. Jang, Hong-Ru Lee, M. Kao

引用次数: 47

Abstract

paper presents the use of linear scaling and tree search in a content-based music retrieval system that can take a user's acoustic input (8-second clip of singing or humming) via a microphone and then retrieve the intended song from over 3000 candidate songs in the database. The system, known as Super MBox, demonstrates the feasibility of real-time content-based music retrieval with a high recognition rate. Super MBox first takes the user's acoustic input from a microphone and converts it into a pitch vector. Then a fast comparison engine using linear scaling and tree search is employed to compute the similarity scores. We have tested Super MBox and found the top-20 recognition rate is about 73% with about 1000 clips of test inputs from people with mediocre singing skills.

查看原文本刊更多论文

基于内容的音乐检索，使用线性缩放和分支绑定树搜索

本文介绍了在基于内容的音乐检索系统中使用线性缩放和树搜索，该系统可以通过麦克风获取用户的声学输入(8秒的唱歌或哼唱片段)，然后从数据库中的3000多首候选歌曲中检索预期的歌曲。该系统被称为Super MBox，证明了基于内容的实时音乐检索具有高识别率的可行性。Super MBox首先从麦克风获取用户的声学输入，并将其转换为音调矢量。然后利用线性缩放和树搜索的快速比较引擎计算相似度分数。我们对Super MBox进行了测试，发现前20名的识别率约为73%，其中有大约1000个测试输入，来自唱歌技能一般的人。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

IEEE International Conference on Multimedia and Expo, 2001. ICME 2001.

自引率

0.00%

发文量