Hierarchical categorization of coarticulated phonemes: A theoretical analysis

Hierarchical categorization of coarticulated phonemes: A theoretical analysis
复制标题

协同音素的层次分类:理论分析

DOI:
--
复制
发表时间:
2001
期刊:
Perception & Psychophysics
影响因子:
--
通讯作者:
R. Smits
R. Smits
中科院分区:
--
文献类型:
--
作者:
R. Smits

文献摘要

参考文献

被引文献

相似文献

本文关注的是听众如何识别协同发音的问题。该问题是从模式分类的角度来解决的。首先,协同发音的潜在声学效果是根据形成分类器输入的模式来定义的。接下来,引入了一种称为 HICAT 的分类模型,该模型结合了分层依赖关系以最佳地处理此输入。该模型允许一个音素边界的位置、方向和陡度取决于相邻音素的感知值。有人认为,如果听众的行为确实像统计模式识别器,他们可能会使用模型中包含的分类策略。将 HICAT 模型与现有的分类模型进行比较,其中包括感知的模糊逻辑模型和 Nearey 的双音素偏向二次线索模型。最后,提出了一种方法,通过该方法可以根据自然语音中出现的声学线索的分布来预测听众可能使用的分类策略。
This article is concerned with the question of how listeners recognize coarticulated phonemes. The problem is approached from a pattern classification perspective. First, the potential acoustical effects of coarticulation are defined in terms of the patterns that form the input to a classifier. Next, a categorization model called HICAT is introduced that incorporates hierarchical dependencies to optimally deal with this input. The model allows the position, orientation, and steepness of one phoneme boundary to depend on the perceived value of a neighboring phoneme. It is argued that, if listeners do behave like statistical pattern recognizers, they may use the categorization strategies incorporated in the model. The HICAT model is compared with existing categorization models, among which are the fuzzylogical model of perception and Nearey’s diphone-biased secondary-cue model. Finally, a method is presented by which categorization strategies that are likely to be used by listeners can be predicted from distributions of acoustical cues as they occur in natural speech.
DOI: 10.1121/1.418179
发表时间: 1997
期刊: The Journal of the Acoustical Society of America
影响因子: --
作者:
Kingston,J;Macmillan,NA;Dickey,LW;Thorburn,R;Bartels,C
通讯作者: Bartels,C
DOI: 10.1121/1.428113
发表时间: 1999
期刊: The Journal of the Acoustical Society of America
影响因子: --
作者:
Macmillan,NA;Kingston,J;Thorburn,R;Dickey,LW;Bartels,C
通讯作者: Bartels,C