Utilising spontaneous conversational speech in HMM-based speech synthesis

Utilising spontaneous conversational speech in HMM-based speech synthesis
复制标题

在基于 HMM 的语音合成中利用自发会话语音

DOI:
--
复制
发表时间:
2010
期刊:
Speech Synthesis Workshop
影响因子:
--
通讯作者:
R. Clark
R. Clark
中科院分区:
--
文献类型:
--
作者:
Sebastian Andersson;J. Yamagishi;R. Clark

文献摘要

参考文献

被引文献

相似文献

自发会话语音有许多特​​征,目前在单元选择和基于 HMM 的语音合成中还没有很好地建模。但为了构建更适合交互的合成语音,我们需要比常用的朗读句子表现出更多会话特征的数据。在本文中,我们将展示从自发对话中精心挑选的话语如何有助于构建基于 HMM 的合成语音,该语音比基于仔细朗读句子的语音具有更自然的对话特征。我们还研究了一种风格混合技术,作为自发语音数据中语音覆盖固有问题的解决方案。但是,缺乏对自发语音现象的适当表示可能导致结果表明我们尚无法与语法句子所实现的语音质量竞争。
Spontaneous conversational speech has many characteristics that are currently not well modelled in unit selection and HMM-based speech synthesis. But in order to build synthetic voices more suitable for interaction we need data that exhibits more conversational characteristics than the generally used read aloud sentences. In this paper we will show how carefully selected utterances from a spontaneous conversation was instrumental for building an HMM-based synthetic voices with more natural sounding conversational characteristics than a voice based on carefully read aloud sentences. We also investigated a style blending technique as a solution to the inherent problem of phonetic coverage in spontaneous speech data. But the lack of an appropriate representation of spontaneous speech phenomena probably contributed to results showing that we could not yet compete with the speech quality achieved for grammatical sentences.
DOI: 10.21437/blizzard.2009-1
发表时间: 2009-09
期刊: The Blizzard Challenge 2009
影响因子: --
作者:
Simon King;Vasilis Karaiskos
通讯作者: Simon King;Vasilis Karaiskos