Don’t hide in the frames: Note- and pattern-based evaluation of automated melody extraction algorithms
Don’t hide in the frames: Note- and pattern-based evaluation of automated melody extraction algorithms
复制标题
不要隐藏在框架中:基于音符和模式的自动旋律提取算法评估
DOI:
10.1145/3358664.3358672
复制
发表时间:
2019
期刊:
影响因子:
--
通讯作者:
G. Peeters
中科院分区:
文献类型:
--
作者:
K. Frieler;D. Başaran;F. Höger;H.-C. Crayencour;G. Peeters
In this paper, we address how to evaluate and improve the performance of automatic dominant melody extraction systems from a pattern mining perspective with a focus on jazz improvisations. Traditionally, dominant melody extraction systems estimate the melody on the frame-level, but for real-world musicological applications note-level representations are needed. For the evaluation of estimated note tracks, the current frame-wise metrics are not fully suitable and provide at most a first approximation. Furthermore, mining melodic patterns (n-grams) poses another challenge because note-wise errors propagate geometrically with increasing length of the pattern. On the other hand, for certain derived metrics such as pattern commonalities between performers, extraction errors might be less critical if at least qualitative rankings can be reproduced. Finally, while searching for similar patterns in a melody database the number of irrelevant patterns in the result set increases with lower similarity thresholds. For reasons of usability, it would be interesting to know the behavior using imperfect automated melody extractions. We propose three novel evaluation strategies for estimated note-tracks based on three application scenarios: Pattern mining, pattern commonalities, and fuzzy pattern search. We apply the proposed metrics to one general state-of-the-art melody estimation method (Melodia) and to two variants of an algorithm that was optimized for the extraction of jazz solos melodies. A subset of the Weimar Jazz Database with 91 solos was used for evaluation. Results show that the optimized algorithm clearly outperforms the reference algorithm, which quickly degrades and eventually breaks down for longer n-grams. Frame-wise metrics provide indeed an estimate for note-wise metrics, but only for sufficiently good extractions, whereas F1 scores for longer n-grams cannot be predicted from frame-wise F1 scores at all. The ranking of pattern commonalities between performers can be reproduced with the optimized algorithms but not with the reference algorithm. Finally, the size of result sets of pattern similarity searches decreases for automated note extraction and for larger similarity thresholds but the difference levels out for smaller thresholds.
登录
查看更多内容
影响因子:
2.4
作者:
A. Volk;Peter van Kranenburg
通讯作者:
Peter van Kranenburg
DOI:
--
发表时间:
1974
期刊:
影响因子:
--
作者:
Thomas Owens
通讯作者:
Thomas Owens
DOI:
--
发表时间:
2020
期刊:
B Jenkins
影响因子:
--
作者:
Carl Woideck
通讯作者:
Carl Woideck
DOI:
--
发表时间:
2018
期刊:
--
影响因子:
--
作者:
K. Frieler
通讯作者:
K. Frieler
DOI:
--
发表时间:
2014
期刊:
--
影响因子:
--
作者:
Colin Raffel;Brian McFee;Eric J. Humphrey;J. Salamon;Oriol Nieto;Dawen Liang;D. Ellis
通讯作者:
Colin Raffel;Brian McFee;Eric J. Humphrey;J. Salamon;Oriol Nieto;Dawen Liang;D. Ellis