Reasons why current speech-enhancement algorithms do not improve speech intelligibility and suggested solutions.

Reasons why current speech-enhancement algorithms do not improve speech intelligibility and suggested solutions.
复制标题

当前语音增强算法无法提高语音清晰度的原因以及建议的解决方案。

DOI:
10.1109/tasl.2010.2045180
复制
发表时间:
2011
期刊:
IEEE transactions on audio, speech, and language processing
影响因子:
--
通讯作者:
Kim G
Kim G
中科院分区:
其他
文献类型:
--
作者:
Loizou PC;Kim G

文献摘要

被引文献

相似文献

现有的语音增强算法可以提高语音质量,但不能提高语音清晰度,其原因尚不清楚。在本文中,我们提出了一个理论框架,可用于分析影响处理后语音清晰度的潜在因素。更具体地说,该框架侧重于对语音增强算法引入的失真进行细粒度分析。据推测,如果这些失真得到适当控制,则可以实现清晰度的大幅提升。为了检验这一假设,我们对人类听众进行了清晰度测试,其中我们呈现经过控制的语音失真的处理后的语音。这些测试的目的是评估语音增强算法可能引入的各种失真对语音清晰度的感知影响。三种不同增强算法的结果表明,某些失真比其他失真对语音清晰度下降的影响更大。然而,当这些失真得到适当控制时,人类听众即使使用已知会降低语音质量和清晰度的频谱相减算法,也能获得很大的清晰度。
Existing speech enhancement algorithms can improve speech quality but not speech intelligibility, and the reasons for that are unclear. In the present paper, we present a theoretical framework that can be used to analyze potential factors that can influence the intelligibility of processed speech. More specifically, this framework focuses on the fine-grain analysis of the distortions introduced by speech enhancement algorithms. It is hypothesized that if these distortions are properly controlled, then large gains in intelligibility can be achieved. To test this hypothesis, intelligibility tests are conducted with human listeners in which we present processed speech with controlled speech distortions. The aim of these tests is to assess the perceptual effect of the various distortions that can be introduced by speech enhancement algorithms on speech intelligibility. Results with three different enhancement algorithms indicated that certain distortions are more detrimental to speech intelligibility degradation than others. When these distortions were properly controlled, however, large gains in intelligibility were obtained by human listeners, even by spectral-subtractive algorithms which are known to degrade speech quality and intelligibility.