Separation of Singing Voice Using Nonnegative Matrix Partial Co-Factorization for Singer Identification
Separation of Singing Voice Using Nonnegative Matrix Partial Co-Factorization for Singer Identification
复制标题
使用非负矩阵部分协分解分离歌声以进行歌手识别
DOI:
10.1109/taslp.2015.2396681
复制
发表时间:
2015-04
影响因子:
5.4
通讯作者:
Liu, Guizhong
中科院分区:
文献类型:
--
作者:
Hu, Ying;Liu, Guizhong
In order to improve the performance of singer identification, we propose a system to separate singing voice from music accompaniment for monaural recordings. Our system consists of two key stages. The first stage exploits the nonnegative matrix partial co-factorization (NMPCF), which is a joint matrix decomposition integrating prior knowledge of singing voice and pure accompaniment to separate the mixture signal into singing voice portion and accompaniment portion. In the second stage, based on the separated singing voice obtained by the first stage, the pitches of singing voice are first estimated and then the harmonic components of singing voice can be distinguished. For a frame, the distinguished harmonic components are regarded as reliable while other frequency components unreliable, thus the spectrum is incomplete. With those harmonic components, the complete spectrums of singing voice can be reconstructed by a missing feature method, spectrum reconstruction, obtaining a refined signal with more clean singing voice. Experimental results demonstrate that, from the point view of source separation, the singing voice refinement can further improve ΔSNR in contrast with the singing voice separation using NMPCF, while for the point view of singer identification, the singing voice separated by NMPCF is more appropriate than the refined singing voice.
登录
查看更多内容
DOI:
10.1109/tasl.2009.2034186
发表时间:
2010-03
期刊:
IEEE Transactions on Audio, Speech, and Language Processing
影响因子:
--
作者:
E. Vincent;N. Bertin;R. Badeau
通讯作者:
E. Vincent;N. Bertin;R. Badeau
DOI:
10.1109/aspaa.2005.1540227
发表时间:
2005-11
期刊:
IEEE Workshop on Applications of Signal Processing to Audio and Acoustics, 2005.
影响因子:
--
作者:
Anssi Klapuri
通讯作者:
Anssi Klapuri
DOI:
10.1109/tasl.2010.2042124
发表时间:
2010-11
期刊:
IEEE Transactions on Audio, Speech, and Language Processing
影响因子:
--
作者:
V. Rao;P. Rao
通讯作者:
V. Rao;P. Rao
DOI:
--
发表时间:
2008
期刊:
--
影响因子:
--
作者:
T. Virtanen;A. Mesaros;M. Ryynänen
通讯作者:
T. Virtanen;A. Mesaros;M. Ryynänen
DOI:
10.1016/j.specom.2004.03.007
发表时间:
2004-09
期刊:
Speech Commun.
影响因子:
--
作者:
B. Raj;M. Seltzer;R. Stern
通讯作者:
B. Raj;M. Seltzer;R. Stern