Recognizing the message and the messenger: biomimetic spectral analysis for robust speech and speaker recognition.
Recognizing the message and the messenger: biomimetic spectral analysis for robust speech and speaker recognition.
复制标题
DOI:
10.1007/s10772-012-9184-y
复制
发表时间:
2013
影响因子:
--
通讯作者:
Elhilali M
中科院分区:
文献类型:
--
作者:
Nemala SK;Patil K;Elhilali M
Humans are quite adept at communicating in presence of noise. However most speech processing systems, like automatic speech and speaker recognition systems, suffer from a significant drop in performance when speech signals are corrupted with unseen background distortions. The proposed work explores the use of a biologically-motivated multi-resolution spectral analysis for speech representation. This approach focuses on the information-rich spectral attributes of speech and presents an intricate yet computationally-efficient analysis of the speech signal by careful choice of model parameters. Further, the approach takes advantage of an information-theoretic analysis of the message and speaker dominant regions in the speech signal, and defines feature representations to address two diverse tasks such as speech and speaker recognition. The proposed analysis surpasses the standard Mel-Frequency Cepstral Coefficients (MFCC), and its enhanced variants (via mean subtraction, variance normalization and time sequence filtering) and yields significant improvements over a state-of-the-art noise robust feature scheme, on both speech and speaker recognition tasks.