Consonant confusions and perceptual spaces for natural and synthetic speech

Consonant confusions and perceptual spaces for natural and synthetic speech
复制标题

自然语音和合成语音的辅音混淆和感知空间

DOI:
10.1121/1.2023027
复制
发表时间:
1985
影响因子:
2.4
通讯作者:
D. Pisoni
D. Pisoni
中科院分区:
物理与天体物理3区
文献类型:
--
作者:
M. Yuchtman;H. Nusbaum;D. Pisoni

文献摘要

被引文献

相似文献

我们实验室的早期研究表明,合成语音比自然语音更难理解,对容量的要求也更高。这些差异似乎与负责将输入信号编码成分段音素表示的过程有关。有几个假设可以解释合成语音感知涉及的更大困难。一种假设是,合成语音在结构上等同于被噪声降级的自然语音。另一种假设是,与自然语音相比,合成语音的声学-语音结构是贫乏的,因为使用最小的声学提示集来实现语音分段。这两个假说导致了对合成辅音混淆性质的不同预测,而合成辅音混淆与自然语音被噪声退化的混淆有关。为了解决这个问题,我们对DEC-Talk产生的合成辅音的混淆矩阵进行了多维尺度分析。
Earlier research in our laboratory has demonstrated that synthetic speech is less intelligible and more capacity demanding than natural speech. These differences appear to be related to processes responsible for encoding the input signal into a segmental phonemic representation. There are several hypotheses that could account for the greater difficulty involved in synthetic speech perception. One hypothesis is that synthetic speech is structurally equivalent to natural speech degraded by noise. An alternative hypothesis is that the acoustic‐phonetic structure of synthetic speech is impoverished in comparison to natural speech in that a minimal set of acoustic cues are used to implement phonetic segments. The two hypotheses lead to different predictions about the nature of synthetic consonant confusions in relation to confusions of natural speech degraded by noise. To resolve this issue, we carried out multidimensional scaling analyses of confusion matrices for synthetic consonants produced by DEC‐Talk, Pr...