A segment-based approach to voice conversion

A segment-based approach to voice conversion
复制标题

基于分段的语音转换方法

DOI:
10.1109/icassp.1991.150451
复制
发表时间:
1991
期刊:
[Proceedings] ICASSP 91: 1991 International Conference on Acoustics, Speech, and Signal Processing
影响因子:
--
通讯作者:
M. Abe
M. Abe
中科院分区:
--
文献类型:
--
作者:
M. Abe

文献摘要

被引文献

相似文献

提出了一种以语音段为转换单元的语音转换算法。输入语音被语音识别模块分解成语音段,并且这些语音段被另一说话者发出的语音段替换。该算法不仅可以转换说话人个性的静态特征,而且可以转换说话人个性的动态特征。所提出的语音转换算法被用于两个男性扬声器。目标语音和转换语音之间的频谱失真被减小到两个扬声器之间的自然频谱失真的三分之一。听力实验表明,在说话人识别精度方面,由段大小的单位转换的语音比逐帧转换的语音高出20%。&lt;<ETX>&gt;
A voice conversion algorithm that uses speech segments as conversion units is proposed. Input speech is decomposed into speech segments by a speech recognition module, and the segments are replaced by speech segments uttered by another speaker. This algorithm makes it possible to convert not only the static characteristics but also the dynamic characteristics of speaker individuality. The proposed voice conversion algorithm was used with two male speakers. Spectrum distortion between target speech and the converted speech was reduced to one-third the natural spectrum distortion between the two speakers. A listening experiment showed that, in terms of speaker identification accuracy, the speech converted by segment-sized units gave a score 20% higher than the speech converted frame-by-frame.<<ETX>>