Development of high-accuracy system for recognizing spontaneous speech
Development of high-accuracy system for recognizing spontaneous speech
批准号:
22500144
负责人:
KOSAKA Tetsuo
金额:
$2.5万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
2010
资助国家:
日本
项目状态:
已结题
起止时间:
2010 至 2012
中文摘要
在我们的研究中,我们的目标是提高系统的性能,识别自发语音,这被认为是比识别阅读语音更困难。我们集中讨论了三个技术问题:(1)声学和语言模型,(2)系统组合技术,(3)说话人索引。为了提高声学模型的性能,我们研究了一种基于区分训练的离散混合隐马尔可夫模型,说话人类模型,五音子模型和混响类模型。研究了连续模型与离散模型相结合、多种五音子相结合、混响级模型相结合等系统组合技术。针对语言模型的问题,提出了交叉自适应和交叉验证自适应技术。此外,我们还改进了基于说话人自适应过程中所需的说话人向量的说话人索引技术的性能。
英文摘要
In our research, we aimed to improve the system performance for recognizing spontaneousspeech, which was considered to be more difficult than recognizing read speech. We focused on three technical issues: (1) acoustic and language models, (2) system combinationtechniques, and (3) speaker indexing. For improving the performance of acoustic models,we investigated a discrete-mixture hidden Markov model based on discriminative training, speaker-class model, quinphone, and a reverberation-class model. Some systemco(a) mbinationtechniquesw(a) ere investigated, such as the combination of continuous anddiscrete models, the combination of various quinphones, and the combination of reverberation-class models. For the issues of language models, we proposed the cross adaptation and cross-validation adaptation techniques. In addition, we improved theperformance of speaker indexing techniques based on speaker vectors required during theexecution of speaker adaptation.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
識別学習を用いた離散混合分布HMMによる音声認識
使用离散混合分布 HMM 进行判别学习的语音识别
DOI:
--
发表时间:
2013
期刊:
情報処理学会論文誌
影响因子:
--
作者:
[Miyazaki, K., 小坂哲夫,加藤正治]
通讯作者:
小坂哲夫,加藤正治
入力音声の韻律情報を用いたHMM音声合成
使用输入语音的韵律信息进行 HMM 语音合成
DOI:
--
发表时间:
2013
期刊:
影响因子:
--
作者:
[Tomoko Nariai, Kazuyo Tanaka, Tatsuya Kawahara, 栗原大樹,加藤正治,小坂哲夫]
通讯作者:
栗原大樹,加藤正治,小坂哲夫
Performance Improvement in Automatic Evaluation System of English Pronunciation by Using Various Normalization Methods
多种归一化方法提高英语发音自动评价系统的性能
DOI:
--
发表时间:
2010
期刊:
Proc. of International Congress on Acoustics 2010
影响因子:
--
作者:
[Masaru Kusumi, Masaharu Kato, Tetsuo Kosaka and Itaru Matsunaga]
通讯作者:
Tetsuo Kosaka and Itaru Matsunaga
Speaker Adaptation Based on System Combination Using Speaker-Class Models
基于使用扬声器类模型的系统组合的扬声器自适应
DOI:
--
发表时间:
2010
期刊:
Proc. of Interspeech2010
影响因子:
--
作者:
[Tetsuo Kosaka, Takashi Ito, Masaharu Kato and Masaki Kohda]
通讯作者:
Masaharu Kato and Masaki Kohda
Lecture Speech Recognition by Combining Word Graphs of Various Acoustic Models
结合各种声学模型的词图进行讲座语音识别
DOI:
--
发表时间:
2010
期刊:
Proc. of Interspeech2010
影响因子:
--
作者:
[Tetsuo Kosaka, Keisuke Goto, Takashi Ito and Masaharu Kato]
通讯作者:
Takashi Ito and Masaharu Kato
共 33 条
Development of Noise Robust Speech Recognition and Its Application on Mobile Environment
-
批准号:16500097
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.86万
-
财政年份:2004
-
负责人:KOSAKA Tetsuo
-
依托单位:
海外基金