Spatial release from informational masking in speech recognition

Spatial release from informational masking in speech recognition
复制标题

DOI:
10.1121/1.1354984
复制
发表时间:
2001-05-01
影响因子:
2.4
通讯作者:
Helfer, KS
Helfer, KS
中科院分区:
物理与天体物理3区
文献类型:
--
作者:
Freyman, RL;Balakrishnan, U;Helfer, KS

文献摘要

被引文献

相似文献

进行了三个实验,以确定在何种程度上感知分离的语音和干扰,提高语音识别在自由领域。目标言语刺激是320个语法正确但无意义的句子,由女性说话者说。在第一个实验中,干扰是一个或两个女性说话者背诵连续流类似的无意义的句子的录音。目标说话者总是从正前方(0度)的扬声器呈现。干扰来自前扬声器(F-F条件)或来自右扬声器(60度)和前扬声器,其中右扬声器领先前扬声器4 ms(F-RF条件)。由于优先效应,在F-RF条件下的干扰被感知为向右,而目标说话者从前面被听到。对于单说话者和两个说话者的干扰,有一个相当大的改善语音识别:F-RF条件相比,F-F条件。然而,第二个实验表明,当干扰是由两个说话者掩蔽的单通道或多通道包络调制的噪声时,没有F-RF优势。第三个实验的结果表明,感知分离的优势不仅限于干扰语音是可理解的条件。(C)2001年,美国声学学会。
Three experiments were conducted to determine the extent to which perceived separation of speech and interference improves speech recognition in the free field. Target speech stimuli were 320 grammatically correct but nonmeaningful sentences spoken by a female talker. In the first experiment the interference was a recording of either one or two female talkers reciting a continuous stream of similar nonmeaningful sentences. The target talker was always presented from a loudspeaker directly in front (0 degrees). The interference was either presented from the front loudspeaker (the F-F condition) or from both a right loudspeaker (60 degrees) and the front loudspeaker, with the right leading the front by 4 ms (the F-RF condition). Due to the precedence effect, the interference in the F-RF condition was perceived to be well to the right, while the target talker was heard from the front. For both the single-talker and two-talker interference, there was a sizable improvement in speech recognition in the: F-RF condition compared with the F-F condition. However, a second experiment showed that there was no F-RF advantage when the interference was noise modulated by the single- or multi-channel envelope of the two-talker masker. Results of the third experiment indicated that the advantage of perceived separation is not limited to conditions where the interfering speech is understandable. (C) 2001 Acoustical Society of America.