Prediction of the Intelligibility for Speech in Real-Life Background Noises for Subjects With Normal Hearing

Prediction of the Intelligibility for Speech in Real-Life Background Noises for Subjects With Normal Hearing
复制标题

预测听力正常受试者在现实生活背景噪声中的语音清晰度

DOI:
10.1097/aud.0b013e31816476d4
复制
发表时间:
2008
期刊:
影响因子:
3.7
通讯作者:
W. Dreschler
W. Dreschler
中科院分区:
医学1区
文献类型:
--
作者:
K. S. Rhebergen;N. Versfeld;W. Dreschler

文献摘要

参考文献

被引文献

相似文献

目的:语音接收阈值(SRT)传统上是在具有目标语音的长期平均语音频谱的平稳噪声中测量的。然而,在现实生活中,背景噪声的瞬时谱很可能不同于平稳的长期平均语音谱噪声。为了更深入地了解现实生活中背景噪声对语言清晰度的影响,在一组在频谱和时间域都不同的噪声中测量了听力正常的听者的SRT。本文考察了Rhebergen等人提出的扩展语音清晰度指数(ESII)的能力。来解释这些现实生活背景噪音中的SRT。设计:对12名听力正常的受试者进行噪声环境下的SRTS测试。干扰噪声由各种现实生活中的噪声组成,这些噪声是从数据库中选择的,并根据它们的频谱-时间差异进行选择。测量的SRT被转换为ESII值并进行比较。理想情况下,在阈值下,ESII值应该相同,因为ESII表示收听者可用的语音信息量。结果:SRT的变化范围为−6dBSNR(平稳噪声)至−21dBSNR(机枪噪声)。换算成ESII值得到的平均值为0.34,标准偏差为0.06。ESII模型对SRT的预测优于传统的SII(ANSI S3.5-1997)模型。在干扰语音的情况下,ESII模型的预测较差,因为人们认为发生了额外的、非能量的(信息)掩蔽。结论:对于目前的一组掩蔽噪声,代表了各种现实生活中的噪声,Rhebergen等人的ESII模型。能够以合理的准确度预测听力正常的受试者的SRT。可以得出结论,ESII模型可以在一些日常情况下为语音清晰度提供有价值的预测。
Objectives: The speech reception threshold (SRT) traditionally is measured in stationary noise that has the long-term average speech spectrum of the target speech. However, in real life the instantaneous spectrum of the background noise is likely to be different from the stationary long-term average speech spectrum noise. To gain more insight into the effect of real-life background noises on speech intelligibility, the SRT of listeners with normal hearing was measured in a set of noises that varied in both the spectral and the temporal domain. This article investigates the ability of the extended speech intelligibility index (ESII), proposed by Rhebergen et al. to account for SRTs in these real-life background noises. Design: SRTs in noise were measured in 12 subjects with normal hearing. Interfering noises consisted of a variety of real-life noises, selected from a database, and chosen on the basis of their spectro-temporal differences. Measured SRTs were converted to ESII values and compared. Ideally, at threshold, ESII values should be the same, because the ESII represents the amount of speech information available to the listener. Results: SRTs ranged from −6 dB SNR (in stationary noise) to −21 dB SNR (in machine gun noise). Conversion to ESII values resulted in an average value of 0.34, with a standard deviation of 0.06. SRT predictions with the ESII model were better than those obtained with the conventional SII (ANSI S3.5-1997) model. In case of interfering speech, the ESII model predictions were poorer, because additional, nonenergetic (informational) masking is thought to occur. Conclusions: For the present set of masking noises, being representative for a variety of real-life noises, the ESII model of Rhebergen et al. is able to predict the SRTs of subjects with normal hearing with reasonable accuracy. It may be concluded that the ESII model can provide valuable predictions for the speech intelligibility in some everyday situations.
DOI: 10.1121/1.1531983
发表时间: 2003-02-01
影响因子: 2.4
作者:
Nelson, PB;Jin, SH;Nelson, DA
通讯作者: Nelson, DA
DOI: 10.1121/1.1555611
发表时间: 2003-04-01
影响因子: 2.4
作者:
Dubno, JR;Horwitz, AR;Ahlstrom, JB
通讯作者: Ahlstrom, JB
DOI: 10.1121/1.2400666
发表时间: 2007-01-01
影响因子: 2.4
作者:
Van Engen, Kristin J.;Bradlow, Ann R.
通讯作者: Bradlow, Ann R.