Exemplar-based voice conversion in noisy environment
Exemplar-based voice conversion in noisy environment
复制标题
DOI:
10.1109/slt.2012.6424242
复制
发表时间:
2012-12
期刊:
影响因子:
--
通讯作者:
R. Takashima;T. Takiguchi;Y. Ariki
中科院分区:
文献类型:
--
作者:
R. Takashima;T. Takiguchi;Y. Ariki
This paper presents a voice conversion (VC) technique for noisy environments, where parallel exemplars are introduced to encode the source speech signal and synthesize the target speech signal. The parallel exemplars (dictionary) consist of the source exemplars and target exemplars, having the same texts uttered by the source and target speakers. The input source signal is decomposed into the source exemplars, noise exemplars obtained from the input signal, and their weights (activities). Then, by using the weights of the source exemplars, the converted signal is constructed from the target exemplars. We carried out speaker conversion tasks using clean speech data and noise-added speech data. The effectiveness of this method was confirmed by comparing its effectiveness with that of a conventional Gaussian Mixture Model (GMM)-based method.