Real time speech / speaker recognition by using digital cochlea system
Real time speech / speaker recognition by using digital cochlea system
批准号:
12650397
负责人:
HANGAI Seiichiro
金额:
$2.05万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
2000
资助国家:
日本
项目状态:
已结题
起止时间:
2000 至 2001
中文摘要
在本研究中,我们讨论了使用数字相干辐射模型的语音/说话人识别系统。数字坐标系模型的优化我们对数字坐标系模型进行了优化,以便在DSP上安装。它由行波滤波器、速度变换滤波器和二次滤波器等16个部分组成.数字耳蜗语音/说话人识别算法的研究我们提出了动态时间规整算法和增强数字耳蜗相邻输出之间的差异。它们可以提高识别性能和对噪声的鲁棒性.数字耳蜗滤波器的实现我们在32片DSP板上设计了数字耳蜗。TMS320C3xDSK。该数字相干滤波器可实现实时语音处理.应用于实时语音识别我们研究了在各种噪声环境下使用该系统的实时语音识别。从实验结果来看,我们在无声环境下取得了99.2%的识别率。10 dB SXR下为90.6%,5dB SXR下为41.0%。应用于实时说话人识别我们测试了18人使用该系统的实时说话人识别。我们实现了92.2%的说话人识别率。另外,通过调整数字相干滤波器各段的增益,我们可以达到98.9%。
英文摘要
In this research, we discuss the speech/speaker recognition system using Digital Cochlear Model as follows.1. Optimization of Digital Cochlear ModelWe optimize the Digital Cochlear Model for installation on DSP. It has 16 section consists of traveling-wave filter, velocity transformation filter and second filter.2. Investigation of speech/speaker recognition algorithm for Digital CochleaWe propose the Dynamic Time Warping algorithm and the enhancement of difference between adjacent outputs of Digital Cochlea. They can improve the recognition performance and the robustness against noise.3. Realization of Digital Cochlea filterWe design the digital cochlea on 32 DSP boards. TMS320C3xDSK. Real-time speech processing can be done by this Digital Cochlear filter.4. Application to real-time speech recognitionWe examine the real-time speech recognition using this system under various noisy environment. From experimental results, we achieve 99.2% recognition rate under silent environment. 90.6% under 10dB SXR and 41.0% under 5dB SXR.5. Application to real-time speaker recognitionWe examine the real-time speaker recognition for 18 persons using this system. We achieve 92.2% speaker recognition rate. In addition, we achieve 98.9% by adjusting the gain of each section of Digital Cochlear filter.
期刊论文(20)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
T.YOSHIDA, T.HAMAMOTO and S.HANGAI: "A Multi-modal HMM for Spoken Word Recognition under Noisy Environment"IEEE Int. conf. on Acoustics, Speech and Signal Processing (ICASSP'01). SPEECHSF1.10. (2001)
T.YOSHIDA、T.HAMAMOTO 和 S.HANGAI:“噪声环境下口语识别的多模态 HMM”IEEE Int。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
吉田孝博: "雑音環境下の単語音声認識のための視聴覚融合HMMについて"信学総大. SD-3-2. (2001)
Takahiro Yoshida:“关于嘈杂环境中单词语音识别的视听融合 HMM”SD-3-2。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
T.YOSHIDA, T.HAMAMOTO and S.HANGAI: "A Study on Multi-Modal HMM for Word Recognition under Noisy Environment"IEICE general conference. SD-3-2. (2001)
T.YOSHIDA、T.HAMAMOTO 和 S.HANGAI:“噪声环境下单词识别的多模态 HMM 研究”IEICE 大会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
T.YOSHIDA, T.HAMAMOTO and S.HANGAI: "Speaker Recognition using Improved Digital Cochlear Filter"IEICE general conference. D-14-4. (2002)
T.YOSHIDA、T.HAMAMOTO 和 S.HANGAI:“使用改进的数字耳蜗滤波器进行说话人识别”IEICE 大会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
M.Namiki: "Spoken word recognition with digital cochlea using 32 DSP-boards"IEEE ICASSP. ITT-L3.5. 1-4 (2001)
M.Namiki:“使用 32 个 DSP 板通过数字耳蜗进行口语识别”IEEE ICASSP。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 15 条
A study on information hiding method with motion vectors of standard video compression
-
批准号:21500179
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.83万
-
财政年份:2009
-
负责人:HANGAI Seiichiro
-
依托单位:
Adaptive Scan Method for Moving Picture Coding System
-
批准号:14550375
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.6万
-
财政年份:2002
-
负责人:HANGAI Seiichiro
-
依托单位:
Real-time writer verification system using tablet information
-
批准号:10650380
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.11万
-
财政年份:1998
-
负责人:HANGAI Seiichiro
-
依托单位:
海外基金