A generalized subspace approach for enhancing speech corrupted by colored noise

A generalized subspace approach for enhancing speech corrupted by colored noise
复制标题

DOI:
10.1109/tsa.2003.814458
复制
发表时间:
2003-07-01
期刊:
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
影响因子:
--
通讯作者:
Loizou, PC
Loizou, PC
中科院分区:
其他
文献类型:
--
作者:
Hu, Y;Loizou, PC

文献摘要

被引文献

相似文献

提出了一种广义子空间方法来增强有色噪声干扰下的语音。基于清洁语音和噪声协方差矩阵的同时对角化,采用非酉变换将噪声信号投影到信号加噪声子空间和噪声子空间上。通过消除噪声子空间中的信号分量并保留信号子空间中的分量来估计干净信号。应用的变换有内置的预白,因此可以用于有色噪声。所提出的方法被证明是对Ephraim和Van Trees提出的白噪声方法的推广。推导了两个基于非酉变换的估计量,一个基于时域约束,一个基于谱域约束。在使用被语音形状噪声和多说话者胡言乱语损坏的TIMIT句子进行测试时,客观和主观测量显示出比其他基于子空间的方法的改进。
A generalized subspace approach is proposed for enhancement of speech corrupted by colored noise. A nonunitary transform, based on the simultaneous diagonalization of the clean speech and noise covariance matrices, is used to project the noisy signal onto a signal-plus-noise subspace and a noise subspace. The clean signal is estimated by nulling the signal components in the noise subspace and retaining the components in the signal subspace. The applied, transform has built-in prewhitening and can therefore be used in general for colored noise. The proposed approach is shown to be a generalization of the approach proposed by Ephraim and Van Trees for white noise. Two estimators were derived based on the nonunitary transform, one based on time-domain constraints and one based on spectral domain constraints. Objective and subjective measures demonstrated improvements over other subspace-based methods when tested with TIMIT sentences corrupted with speech-shaped noise and multi-talker babble.