A pitch based noise estimation technique for robust speech recognition with Missing Data

A pitch based noise estimation technique for robust speech recognition with Missing Data
复制标题

DOI:
10.1109/icassp.2011.5947431
复制
发表时间:
2011-05
期刊:
2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
影响因子:
--
通讯作者:
J. A. Morales-Cordovilla;Ning Ma;V. Sánchez;J. L. Carmona;A. Peinado;J. Barker
J. A. Morales-Cordovilla;Ning Ma;V. Sánchez;J. L. Carmona;A. Peinado;J. Barker
中科院分区:
其他
文献类型:
--
作者:
J. A. Morales-Cordovilla;Ning Ma;V. Sánchez;J. L. Carmona;A. Peinado;J. Barker

文献摘要

被引文献

相似文献

本文提出了一种基于基音信息的噪声估计技术,用于鲁棒语音识别。在第一阶段中,通过从认为不存在语音的帧外推噪声来估计噪声。这些帧被检测与建议的音高为基础的VAD(语音活动检测器)。在第二阶段中,噪声估计在有声帧中使用谐波隧道技术进行修正。隧道噪声估计在高SNR下用作噪声的上限而不是合适的估计。一个谱图MD(缺失数据)识别系统被选择来评估所提出的噪声估计。建议的系统进行了比较,在极光-2与其他类似的技术,如倒谱SS(频谱减法)。
This paper presents a noise estimation technique based on knowledge of pitch information for robust speech recognition. In the first stage the noise is estimated by means of extrapolating the noise from frames where speech is believed to be absent. These frames are detected with a proposed pitch based VAD (Voice Activity Detector). In the second stage the noise estimation is revised in voiced frames using harmonic tunnelling thechnique. The tunnelling noise estimation is used at high SNRs as an upper bound of the noise rather than a suitable estimation. A spectrogram MD (Missing Data) recognition system is chosen to evaluate the proposed noise estimation. The proposed system is compared in Aurora-2 with other similar techniques like cepstral SS (Spectral Subtraction).