Sequential stream segregation of voiced and unvoiced speech sounds based on fundamental frequency.

Sequential stream segregation of voiced and unvoiced speech sounds based on fundamental frequency.
复制标题

DOI:
10.1016/j.heares.2016.11.016
复制
发表时间:
2017-02
期刊:
影响因子:
2.8
通讯作者:
Oxenham AJ
Oxenham AJ
中科院分区:
医学1区
文献类型:
--
作者:
David M;Lavandier M;Grimault N;Oxenham AJ

文献摘要

被引文献

相似文献

已知浊音之间的基频(F0)的差异是流分离的强烈提示。然而,语音由有声和无声声音组成,并且关于无声部分是否以及如何被分离的知之甚少。本研究测量了听者整合或分离辅音-元音标记序列的能力,包括无辅音摩擦音和元音,作为标记交错序列之间的F0差异的函数。一个基于性能的措施,其中听众检测到一个序列内或两个序列之间的重复令牌的存在(措施的自愿和强制性的流,分别)。结果显示,当两个交织序列之间的F0差异从0增加到13个单位音时,自愿流分离的系统性增加,表明F0差异允许听众分离语音声音,包括清音部分。相反,自愿流的一致影响,强制性流隔离的趋势在大F0差异未能达到显着性。当无声部分从刺激中移除时,听众不再能够可靠地执行非线性流任务,这表明无声部分在原始任务中被使用并正确地分离。结果表明,基于F0差异的流发生自然语音声音,并且清音部分被正确地分配给语音声音的对应的有声部分。
Differences in fundamental frequency (F0) between voiced sounds are known to be a strong cue for stream segregation. However, speech consists of both voiced and unvoiced sounds, and less is known about whether and how the unvoiced portions are segregated. This study measured listeners’ ability to integrate or segregate sequences of consonant-vowel tokens, comprising a voiceless fricative and a vowel, as a function of the F0 difference between interleaved sequences of tokens. A performance-based measure was used, in which listeners detected the presence of a repeated token either within one sequence or between the two sequences (measures of voluntary and obligatory streaming, respectively). The results showed a systematic increase of voluntary stream segregation as the F0 difference between the two interleaved sequences increased from 0 to 13 semitones, suggesting that F0 differences allowed listeners to segregate speech sounds, including the unvoiced portions. In contrast to the consistent effects of voluntary streaming, the trend towards obligatory stream segregation at large F0 differences failed to reach significance. Listeners were no longer able to perform the voluntary-streaming task reliably when the unvoiced portions were removed from the stimuli, suggesting that the unvoiced portions were used and correctly segregated in the original task. The results demonstrate that streaming based on F0 differences occurs for natural speech sounds, and that unvoiced portions are correctly assigned to corresponding voiced portions of the speech sounds.