Unsupervised Music Structure Annotation by Time Series Structure Features and Segment Similarity
Unsupervised Music Structure Annotation by Time Series Structure Features and Segment Similarity
复制标题
DOI:
10.1109/tmm.2014.2310701
复制
发表时间:
2014-08-01
影响因子:
7.3
通讯作者:
Arcos, Josep Ll
中科院分区:
文献类型:
--
作者:
Serra, Joan;Mueller, Meinard;Arcos, Josep Ll
Automatically inferring the structural properties of raw multimedia documents is essential in today's digitized society. Given its hierarchical and multi-faceted organization, musical pieces represent a challenge for current computational systems. In this article, we present a novel approach to music structure annotation based on the combination of structure features with time series similarity. Structure features encapsulate both local and global properties of a time series, and allow us to detect boundaries between homogeneous, novel, or repeated segments. Time series similarity is used to identify equivalent segments, corresponding to musically meaningful parts. Extensive tests with a total of five benchmark music collections and seven different human annotations show that the proposed approach is robust to different ground truth choices and parameter settings. Moreover, we see that it outperforms previous approaches evaluated under the same framework.