An Improved Speech / Nonspeech Classification Based on Feature Combination for Audio Indexing
An Improved Speech / Nonspeech Classification Based on Feature Combination for Audio Indexing
复制标题
一种改进的基于音频索引特征组合的语音/非语音分类
DOI:
10.1587/transfun.e93.a.830
复制
发表时间:
2010
期刊:
影响因子:
--
通讯作者:
M. Hagiwara
中科院分区:
文献类型:
--
作者:
Ji;Hyon;M. Hagiwara
In this letter, we propose an improved speech/nonspeech classification method to effectively classify a multimedia source. To improve performance, we introduce a feature based on spectral duration analysis, and combine recently proposed features such as high zero crossing rate ratio (HZCRR), low short time energy ratio (LSTER), and pitch ratio (PR). According to the results of our experiments on speech, music, and environmental sounds, the proposed method obtained high classification results when compared with conventional approaches.