Fast NMF based approach and improved VQ based approach for speech recognition from mixed sound
Fast NMF based approach and improved VQ based approach for speech recognition from mixed sound
复制标题
DOI:
--
复制
发表时间:
2012-12
期刊:
影响因子:
--
通讯作者:
Shoichi Nakano;Kazumasa Yamamoto;S. Nakagawa
中科院分区:
文献类型:
--
作者:
Shoichi Nakano;Kazumasa Yamamoto;S. Nakagawa
We have considered a speech recognition method for mixed sound, consisting of speech and music, that removes only the music based on vector quantization (VQ) and non-negative matrix factorization (NMF). This paper describe fast calculation technique of music removal based on NMF and improvement using a VQ method. For isolated word recognition using the clean speech model, an improvement of 46% word error reduction rate was obtained compared with the case of not removing music. Furthermore, a high recognition rate, close to clean speech recognition was obtained at 10 dB. For the case of the multi-conditions, our proposed method reduced the error rate of 50% compared with the multi-conditions model.