Contribution of Stimulus Variability to Word Recognition in Noise Versus Two-Talker Speech for School-Age Children and Adults.

Contribution of Stimulus Variability to Word Recognition in Noise Versus Two-Talker Speech for School-Age Children and Adults.
复制标题

DOI:
10.1097/aud.0000000000000951
复制
发表时间:
2021
期刊:
影响因子:
3.7
通讯作者:
Leibold LJ
Leibold LJ
中科院分区:
医学1区
文献类型:
--
作者:
Buss E;Calandruccio L;Oleson J;Leibold LJ

文献摘要

被引文献

相似文献

语音识别分数往往比噪声中的语音识别分数更易变,无论是在听众内部还是在听众之间。这种变化可能是由于听者因素,如个体差异的听觉或敏感性信息掩蔽。这也可能是由于刺激的可变性,一些语音样本比其他样本更具挑战性。本实验的目的是检验两个假设:1)刺激的可变性影响成人在两个说话人的言语掩蔽中的单词识别,2)刺激的可变性在儿童的表现中起着较小的作用,由于相对较大的贡献听者因素。听力正常的儿童(5-10岁)和成人(18 - 41岁)。目标语音是一个语料库的30个双音节词,每个与一个明确的说明。掩蔽物是30个样本的两个谈话者的语音或语音形状的噪音。这个任务是一个四选一的被迫选择。自适应地测量语音接收阈值(SRT),这些结果用于确定每个听众和掩蔽者的~65%正确率相关的信噪比。然后在两种条件下完成两个30字的固定水平测试块:1)在每个块之前随机分配目标掩蔽物对,以及2)冻结目标掩蔽物对。成年人的SRT低于儿童,特别是对于两个谈话者的言语掩蔽。在固定水平测试中,我们评估了听众之间的一致性。目标样本是随机和冻结条件下语音形状噪声掩蔽器性能的最佳预测因子。相比之下,目标和掩蔽样本影响性能的两个谈话者掩蔽。儿童和成人的结果是定性相似的,在两个年龄组的掩蔽目标可听度的差异,在整个刺激样本的性能模式是一致的。虽然在语音形状的噪声中的单词识别在目标单词之间始终不同,但在两个说话者的语音掩蔽中的识别取决于目标和掩蔽样本。这些刺激效应与掩蔽目标可听度的简单模型大致一致。虽然语音识别的变异性通常被认为反映了信息掩蔽的差异,但目前的研究结果表明,能量掩蔽在刺激中的变异性可以在性能中发挥重要作用。
Speech-in-speech recognition scores tend to be more variable than those for speech-in-noise recognition, both within and across listeners. This variability could be due to listener factors, such as individual differences in audibility or susceptibility to informational masking. It could also be due to stimulus variability, with some speech-in-speech samples posing more of a challenge than others. The purpose of this experiment was to test two hypotheses: 1) that stimulus variability affects adults’ word recognition in a two-talker speech masker, and 2) that stimulus variability plays a smaller role in children’s performance due to relatively greater contributions of listener factors. Listeners were children (5–10 yrs) and adults (18 – 41 yrs) with normal hearing. Target speech was a corpus of 30 disyllabic words, each associated with an unambiguous illustration. Maskers were 30 samples of either two-talker speech or speech-shaped noise. The task was a four-alternative forced choice. Speech reception thresholds (SRTs) were measured adaptively, and those results were used to determine the signal-to-noise ratio associated with ~65% correct for each listener and masker. Two 30-word blocks of fixed-level testing were then completed in each of two conditions: 1) with the target-masker pairs randomly assigned prior to each block, and 2) with frozen target-masker pairs. The SRTs were lower for adults than children, particularly for the two-talker speech masker. Listener responses in fixed-level testing were evaluated for consistency across listeners. The target sample was the best predictor of performance in the speech-shaped noise masker for both the random and frozen conditions. In contrast, both the target and masker samples affected performance in the two-talker masker. Results were qualitatively similar for children and adults, and the pattern of performance across stimulus samples was consistent with differences in masked target audibility in both age groups. Whereas word recognition in speech-shaped noise differed consistently across target words, recognition in a two-talker speech masker depended on both the target and masker samples. These stimulus effects are broadly consistent with a simple model of masked target audibility. Although variability in speech-in-speech recognition is often thought to reflect differences in informational masking, the present results suggest that variability in energetic masking across stimuli can play an important role in performance.