A statistical model of speech F0 contours

A statistical model of speech F0 contours
复制标题

语音F0轮廓的统计模型

DOI:
--
复制
发表时间:
2010
期刊:
--
影响因子:
--
通讯作者:
Yasunori Ohishi
Yasunori Ohishi
中科院分区:
--
文献类型:
--
作者:
H. Kameoka;Jonathan Le Roux;Yasunori Ohishi

文献摘要

参考文献

被引文献

相似文献

本文基于藤崎模型离散时间随机过程版本的表述,提出了一种语音基频(F0)轮廓的统计模型,藤崎模型被认为是代表声带振动控制机制的有根据的数学模型。这种统计公式有两个重要的动机。一是为 Fujisaki 模型导出通用参数估计框架,允许引入强大的统计方法,二是通过概率分布假设引入 F0 轮廓的语音自然度度量,该度量可以纳入许多统计语音处理问题,例如语音分析、合成、分离、去噪和去混响。
This paper proposes a statistical model of speech fundamental frequency (F0) contours, based on the formulation of the discrete-time stochastic process version of the Fujisaki model, which is known as a well-founded mathematical model representing the control mechanism of vocal fold vibration. There are two important motivations for this statistical formulation. One is to derive a general parameter estimation framework for the Fujisaki model, allowing for the introduction of powerful statistical methods, and the other is to introduce a measure of speech naturalness in terms of an F0 contour through a probability distribution assumption, that can be incorporated into many statistical speech processing problems such as speech analysis, synthesis, separation, denoising and dereverberation.
DOI: 10.1121/1.1458024
发表时间: 2002-04-01
影响因子: 2.4
作者:
de Cheveigné, A;Kawahara, H
通讯作者: Kawahara, H