Merging data driven and rule based prosodic models for unit selection TTS
Merging data driven and rule based prosodic models for unit selection TTS
复制标题
合并数据驱动和基于规则的韵律模型以进行单元选择 TTS
DOI:
--
复制
发表时间:
2004
期刊:
影响因子:
--
通讯作者:
M. Aylett
中科院分区:
文献类型:
--
作者:
M. Aylett
Data driven models suffer from data sparsity and can be difficult to generalise. Rule based models suffer from being over prescriptive and insensitive to the contents of the unit selection database. To further complicate matters the space of acceptable prosody for any one utterance is large. However in some cases prosodic patterns for a particular speaker can be very homogeneous, for example the prosodic pattern used to read out a zip code. In this paper we describe a method for exploring and analysing the prosodic space within a limited domain, and a method for merging a simple rule based prosodic model with a set of data driven mini prosodic models. A listening test was carried out on the synthesis of zip codes with and without the mini models with promising results. The approach could be applied effectively to domains varying from numerical amounts to personal names.