A method for automatic extraction of model parameters from fundamental frequency contours of speech

A method for automatic extraction of model parameters from fundamental frequency contours of speech
复制标题

一种从语音基频轮廓自动提取模型参数的方法

DOI:
10.1109/icassp.2002.5743766
复制
发表时间:
2002
期刊:
2002 IEEE International Conference on Acoustics, Speech, and Signal Processing
影响因子:
--
通讯作者:
H. Fujisaki
H. Fujisaki
中科院分区:
--
文献类型:
--
作者:
S. Narusawa;N. Minematsu;K. Hirose;H. Fujisaki

文献摘要

被引文献

相似文献

生成语音的F0轮廓的过程已经由Fujisaki和他的同事在数学术语中相当准确地建模,但是从观察到的F0轮廓中提取底层命令的参数是一个逆问题,只能通过逐次逼近来解决。为了保证有效和准确地搜索解决方案,需要从一组足够接近最优值的初始值开始。本文提出了一种对实测的F0轮廓进行预处理的方法,以获得由处处连续且可微的三阶多项式段组成的近似轮廓。结果表明,所提出的方法允许一个获得一阶近似的重音命令的参数约90%的所有口音命令,和短语命令的约84%的所有短语命令。
The process of generating the F0 contour of speech has been modeled quite accurately in mathematical tenns by Fujisaki and his coworkers, but the extraction of parameters of the underlying commands from an observed F0 contour is an inverse problem that can be solved only by successive approximation. In order to guarantee an efficient and accurate search for the solution, one needs to start with a set of initial values that are close enough to the optimum. This paper presents a method for pre-processing a measured F0 contour to obtain its approximation consisting of third-order polynomial segments that are continuous and differentiable everywhere. It is shown that the proposed method allows one to obtain first-order approximations to the parameters of accent commands for about 90% of all the accent commands, and of phrase commands for about 84% of all the phrase commands.