Real time speech / speaker recognition by using digital cochlea system
Real time speech / speaker recognition by using digital cochlea system
批准号:
12650397
负责人:
HANGAI Seiichiro
金额:
$2.05万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
2000
资助国家:
日本
项目状态:
已结题
起止时间:
2000 至 2001
中文摘要
在本研究中,我们讨论了基于数字耳蜗模型的语音/说话人识别系统。数字耳蜗机模型的优化我们对安装在数字信号处理器上的数字耳蜗机进行了优化。它由行波滤波、速度变换滤波和二次滤波共16个部分组成。数字耳蜗语音/说话人识别算法的研究提出了动态时间规整算法和增强相邻输出之间的差值。它们可以提高识别性能和对噪声的鲁棒性。数字耳蜗片的实现我们在32块DSP板上设计了数字耳蜗板。数字信号处理器TMS320C3xDSK。利用该数字耳蜗式滤波器可以进行实时语音处理。在实时语音识别中的应用我们测试了在各种噪声环境下使用该系统进行的实时语音识别。实验结果表明,在静音环境下,识别率达到99.2%。10dBSXR时为90.6%,5dBSXR时为41.0%。在实时说话人识别中的应用我们使用该系统对18个人的实时说话人识别进行了检验。达到了92.2%的说话人识别率。此外,我们还通过调整每一段数字耳蜗滤波器的增益达到了98.9%。
英文摘要
In this research, we discuss the speech/speaker recognition system using Digital Cochlear Model as follows.1. Optimization of Digital Cochlear ModelWe optimize the Digital Cochlear Model for installation on DSP. It has 16 section consists of traveling-wave filter, velocity transformation filter and second filter.2. Investigation of speech/speaker recognition algorithm for Digital CochleaWe propose the Dynamic Time Warping algorithm and the enhancement of difference between adjacent outputs of Digital Cochlea. They can improve the recognition performance and the robustness against noise.3. Realization of Digital Cochlea filterWe design the digital cochlea on 32 DSP boards. TMS320C3xDSK. Real-time speech processing can be done by this Digital Cochlear filter.4. Application to real-time speech recognitionWe examine the real-time speech recognition using this system under various noisy environment. From experimental results, we achieve 99.2% recognition rate under silent environment. 90.6% under 10dB SXR and 41.0% under 5dB SXR.5. Application to real-time speaker recognitionWe examine the real-time speaker recognition for 18 persons using this system. We achieve 92.2% speaker recognition rate. In addition, we achieve 98.9% by adjusting the gain of each section of Digital Cochlear filter.
期刊论文(20)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
T.YOSHIDA, T.HAMAMOTO and S.HANGAI: "A Multi-modal HMM for Spoken Word Recognition under Noisy Environment"IEEE Int. conf. on Acoustics, Speech and Signal Processing (ICASSP'01). SPEECHSF1.10. (2001)
T.YOSHIDA、T.HAMAMOTO 和 S.HANGAI:“噪声环境下口语识别的多模态 HMM”IEEE Int。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
吉田孝博: "雑音環境下の単語音声認識のための視聴覚融合HMMについて"信学総大. SD-3-2. (2001)
Takahiro Yoshida:“关于嘈杂环境中单词语音识别的视听融合 HMM”SD-3-2。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
T.YOSHIDA, T.HAMAMOTO and S.HANGAI: "A Study on Multi-Modal HMM for Word Recognition under Noisy Environment"IEICE general conference. SD-3-2. (2001)
T.YOSHIDA、T.HAMAMOTO 和 S.HANGAI:“噪声环境下单词识别的多模态 HMM 研究”IEICE 大会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
T.YOSHIDA, T.HAMAMOTO and S.HANGAI: "Speaker Recognition using Improved Digital Cochlear Filter"IEICE general conference. D-14-4. (2002)
T.YOSHIDA、T.HAMAMOTO 和 S.HANGAI:“使用改进的数字耳蜗滤波器进行说话人识别”IEICE 大会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
M.Namiki: "Spoken word recognition with digital cochlea using 32 DSP-boards"IEEE ICASSP. ITT-L3.5. 1-4 (2001)
M.Namiki:“使用 32 个 DSP 板通过数字耳蜗进行口语识别”IEEE ICASSP。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 15 条
A study on information hiding method with motion vectors of standard video compression
-
批准号:21500179
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.83万
-
财政年份:2009
-
负责人:HANGAI Seiichiro
-
依托单位:
Adaptive Scan Method for Moving Picture Coding System
-
批准号:14550375
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.6万
-
财政年份:2002
-
负责人:HANGAI Seiichiro
-
依托单位:
Real-time writer verification system using tablet information
-
批准号:10650380
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.11万
-
财政年份:1998
-
负责人:HANGAI Seiichiro
-
依托单位:
海外基金