Varying acoustic-phonemic ambiguity reveals that talker normalization is obligatory in speech processing.

Varying acoustic-phonemic ambiguity reveals that talker normalization is obligatory in speech processing.
复制标题

DOI:
10.3758/s13414-017-1395-5
复制
发表时间:
2018-04
期刊:
Attention, perception & psychophysics
影响因子:
--
通讯作者:
Perrachione TK
Perrachione TK
中科院分区:
其他
文献类型:
--
作者:
Choi JY;Hu ER;Perrachione TK

文献摘要

参考文献

被引文献

相似文献

语音声学和抽象音素表征之间的不确定关系给听者带来了一个挑战,即尽管语音的声学实现高度可变,但仍要保持知觉的一致性。说话者标准化通过减少在遇到的语音和音素表示之间映射的自由度来促进语音处理。虽然这一过程被提出是为了促进对歧义语音的感知,但目前还不清楚说话者的归一化是否受到声学-音素映射中潜在歧义程度的影响。我们在一系列快速分类范式中探索了说话者归一化对语音处理的影响,对说话者之间辅音和元音的声学-音素关系不一致的可能性进行了参数控制。听者在不同说话者之间识别具有不同潜在声学音素歧义的单词(例如,甜菜/船与靴子/船),这些词由单一或混合说话者说出。与单一说话者相比,当听到混合说话者时,即使目标声音之间没有潜在的声学模糊,对单词的听觉分类也总是慢于单一说话者。此外,当单词在不同说话者之间具有最大的潜在声学-音素重叠时,混合说话者施加的处理成本最大。目标语音之间的声学差异模型没有解释结果的模式。这些结果表明:(I)说话者归一化在消除高度混淆的声音时产生最大的处理成本,以及(Ii)说话者归一化似乎是言语感知的一个必要组成部分,即使在声音之间的声学-音素关系明确的情况下也是如此。
The nondeterministic relationship between speech acoustics and abstract phonemic representations imposes a challenge for listeners to maintain perceptual constancy despite the highly variable acoustic realization of speech. Talker normalization facilitates speech processing by reducing the degrees of freedom for mapping between encountered speech and phonemic representations. While this process has been proposed to facilitate the perception of ambiguous speech sounds, it is currently unknown whether talker normalization is affected by the degree of potential ambiguity in acoustic-phonemic mapping. We explored the effects of talker normalization on speech processing in a series of speeded classification paradigms, parametrically manipulating the potential for inconsistent acoustic-phonemic relationships across talkers for both consonants and vowels. Listeners identified words with varying potential acoustic-phonemic ambiguity across talkers (e.g., beet/boat vs. boot/boat) spoken by single or mixed talkers. Auditory categorization of words was always slower when listening to mixed talkers compared to a single talker, even when there was no potential acoustic ambiguity between target sounds. Moreover, the processing cost imposed by mixed talkers was greatest when words had the most potential acoustic-phonemic overlap across talkers. Models of acoustic dissimilarity between target speech sounds did not account for the pattern of results. These results suggest (i) that talker normalization incurs the greatest processing cost when disambiguating highly-confusable sounds and (ii) that talker normalization appears to be an obligatory component of speech perception, taking place even when the acoustic-phonemic relationships across sounds are unambiguous.
DOI: 10.1080/00437956.1964.11659830
发表时间: 1964-01-01
影响因子: 0.6
作者:
LISKER, L;ABRAMSON, AS
通讯作者: ABRAMSON, AS
DOI: 10.1037/0278-7393.17.1.152
发表时间: 1991-01-01
影响因子: 2.6
作者:
GOLDINGER, SD;PISONI, DB;LOGAN, JS
通讯作者: LOGAN, JS
DOI: 10.3758/bf03213123
发表时间: 1999-11-01
期刊: PERCEPTION & PSYCHOPHYSICS
影响因子: --
作者:
Huettel, SA;Lockhead, GR
通讯作者: Lockhead, GR
DOI: 10.1121/1.387579
发表时间: 1982-01-01
影响因子: 2.4
作者:
ASSMANN, PF;NEAREY, TM;HOGAN, JT
通讯作者: HOGAN, JT
DOI: 10.1037/a0038695
发表时间: 2015-04
影响因子: 5.4
作者:
Kleinschmidt DF;Jaeger TF
通讯作者: Jaeger TF