New rule-based and data-driven strategy to incorporate Fujisaki's F/sub 0/ model to a text-to-speech system in Castillian Spanish

New rule-based and data-driven strategy to incorporate Fujisaki's F/sub 0/ model to a text-to-speech system in Castillian Spanish
复制标题

新的基于规则和数据驱动的策略,将 Fujisaki 的 F/sub 0/ 模型纳入卡斯蒂利亚西班牙语的文本转语音系统

DOI:
10.1109/icassp.2001.941041
复制
发表时间:
2001
期刊:
2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No.01CH37221)
影响因子:
--
通讯作者:
J. Pardo
J. Pardo
中科院分区:
--
文献类型:
--
作者:
J. Gutiérrez;J. Montero;D. Saiz;J. Pardo

文献摘要

被引文献

相似文献

我们通过估计Fujisaki(1981)的F/sub0/Contour模型的参数来分析西班牙语韵律数据库。这些参数根据语言特征进行分类,并形成分析数据库。在合成F/sub0/轮廓时,我们从文本中提取语言特征,并进行k近邻搜索。使用韵律数据库中的数据来训练语言特征比较距离。为了避免伪影,我们对合成参数执行基于规则的过滤。评估测试的结果表明,该系统明显优于以前的神经网络方法。这一评估证实了Fujisaki的模型基于语言特征来表征韵律信息的能力。
We present the analysis of a Spanish prosody database by estimating the parameters of Fujisaki's (1981) model for F/sub 0/ contours. These parameters are classified attending to linguistic features and they form the analysis database. When synthesizing F/sub 0/ contours we extract the linguistic features from the text and perform a k-nearest neighbour search. Linguistic feature comparison distance is trained using data from the prosody database. To avoid artifacts we perform a rule-base filtering on synthesis parameters. The results of our evaluation test show that the proposed system is significantly better than the previous neural network approach. This evaluation confirms the ability of Fujisaki's model to represent prosody information based on linguistic features.