Singing Voice Melody Transcription Using Deep Neural Networks

Singing Voice Melody Transcription Using Deep Neural Networks
复制标题

DOI:
--
复制
发表时间:
2016
期刊:
--
影响因子:
--
通讯作者:
François Rigaud;Mathieu Radenen
François Rigaud;Mathieu Radenen
中科院分区:
其他
文献类型:
--
作者:
François Rigaud;Mathieu Radenen

文献摘要

被引文献

相似文献

本文提出了一种基于深度神经网络(DNN)模型的复调音乐信号中演唱声音旋律的转录系统。特别是,一个新的DNN系统引入执行f 0估计的旋律,另一个DNN,从最近的研究中得到启发,学习分割声乐序列。描述了与这两项任务的具体情况相关的数据准备和学习配置。的旋律f 0估计系统的性能进行了比较,一个国家的最先进的方法,并表现出最高的精度,通过更好的概括两个不同的音乐数据库。深入了解这DNN的全球运作提出。最后,评价的全球系统相结合的两个DNN的歌声旋律转录。
This paper presents a system for the transcription of singing voice melodies in polyphonic music signals based on Deep Neural Network (DNN) models. In particular, a new DNN system is introduced for performing the f 0 estimation of the melody, and another DNN, inspired from recent studies, is learned for segmenting vocal sequences. Preparation of the data and learning configurations related to the specificity of both tasks are described. The performance of the melody f 0 estimation system is compared with a state-of-the-art method and exhibits highest accuracy through a better generalization on two different music databases. Insights into the global functioning of this DNN are proposed. Finally, an evaluation of the global system combining the two DNNs for singing voice melody transcription is presented.