Multi-lingual duration modeling
Multi-lingual duration modeling
复制标题
多语言持续时间建模
DOI:
10.21437/eurospeech.1997-669
复制
发表时间:
1997
期刊:
影响因子:
--
通讯作者:
M. Tanenblatt
中科院分区:
文献类型:
--
作者:
J. V. Santen;Chilin Shih;Bernd Möbius;E. Tzoukermann;M. Tanenblatt
Controlling timing in text-to-speech synthesis systems is complicated, because there are many contextual factors that affect timing; moreover, which factors matter and what their precise effects are varies among languages. We describe here a language-independent approach for duration control. At run time, a language-independent timing module accesses languagespecific tables. These tables specify which sub-classes of the feature space (i.e., all combinations of context and phone identity) are homogeneous in the specific sense that the same factors have similar effects on the cases in a sub-class. Within a sub-class, durations are modeled by simple arithmetic models such as multiplicative, additive, or – more generally – sums-ofproducts models. Exploratory statistical methods (supervised) and parameter estimation techniques (unsupervised) are used for