Bayesian Feature Enhancement for Reverberation and Noise Robust Speech Recognition
Bayesian Feature Enhancement for Reverberation and Noise Robust Speech Recognition
复制标题
用于混响和噪声鲁棒语音识别的贝叶斯特征增强
DOI:
--
复制
发表时间:
2013
期刊:
影响因子:
--
通讯作者:
Reinhold Häb
中科院分区:
文献类型:
--
作者:
Volker Leutnant;A. Krueger;Reinhold Häb
In this contribution we extend a previously proposed Bayesian approach for the enhancement of reverberant logarithmic mel power spectral coefficients for robust automatic speech recognition to the additional compensation of background noise. A recently proposed observation model is employed whose time-variant observation error statistics are obtained as a side product of the inference of the a posteriori probability density function of the clean speech feature vectors. Further a reduction of the computational effort and the memory requirements are achieved by using a recursive formulation of the observation model. The performance of the proposed algorithms is first experimentally studied on a connected digits recognition task with artificially created noisy reverberant data. It is shown that the use of the time-variant observation error model leads to a significant error rate reduction at low signal-to-noise ratios compared to a time-invariant model. Further experiments were conducted on a 5000 word task recorded in a reverberant and noisy environment. A significant word error rate reduction was obtained demonstrating the effectiveness of the approach on real-world data.