Estimation of articulatory movements from speech acoustics using an HMM-based speech production model

Estimation of articulatory movements from speech acoustics using an HMM-based speech production model
复制标题

DOI:
10.1109/tsa.2003.822636
复制
发表时间:
2004-04
期刊:
IEEE Transactions on Speech and Audio Processing
影响因子:
--
通讯作者:
S. Hiroya;M. Honda
S. Hiroya;M. Honda
中科院分区:
其他
文献类型:
--
作者:
S. Hiroya;M. Honda

文献摘要

被引文献

相似文献

我们提出了一种方法,确定发音运动从语音声学使用隐马尔可夫模型(HMM)为基础的语音生产模型。该模型从一个给定的音素串统计生成语音频谱和发音参数。它由每个音素的发音参数的HALTH和每个HMM状态的发音到声学映射组成。对于给定的语音频谱,统计模型的发音参数的最大后验估计。通过比较估计的发音参数与观察到的参数来评估句子的性能。估计的发音参数的平均RMS误差与语音声学和话语中的音素信息的误差为1.50 mm,与语音声学的误差为1.73 mm。
We present a method that determines articulatory movements from speech acoustics using a Hidden Markov Model (HMM)-based speech production model. The model statistically generates speech spectrum and articulatory parameters from a given phonemic string. It consists of HMMs of articulatory parameters for each phoneme and an articulatory-to-acoustic mapping for each HMM state. For a given speech spectrum, maximum a posteriori estimation of the articulatory parameters of the statistical model is presented. The performance on sentences was evaluated by comparing the estimated articulatory parameters with the observed parameters. The average RMS errors of the estimated articulatory parameters were 1.50 mm from the speech acoustics and the phonemic information in an utterance and 1.73 mm from the speech acoustics only.