Formulation of Prosodic Features of Speech and its Application to Continuous Speech Recognition
Formulation of Prosodic Features of Speech and its Application to Continuous Speech Recognition
批准号:
06452397
负责人:
HIROSE Keikichi
金额:
$4.8万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
1994
资助国家:
日本
项目状态:
已结题
起止时间:
1994 至 1996
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Prosodic feature-based methods were developed for word identification, syntactic boundary detection, and so on. These methods were utillized to aid continuous speech recognition. Followings are the major results.1. Prosodic rules for read speech were modified and improved. Prosodic rules for dialogue speech were also construcled based on the comparative study with read speech. Prosodic features were clarified for speakers' intention, attitude and emotion.2. A method was develped to detect syntactic boundaries in continuous speech using fundamental frequency contours and their macroscopic features.3. A method was develped to extract phrase component onsels from fundamental frequency contours by suppressing local undulations due to accent components. By combining this with AbS method based on the superpositional model, automatic extraction of fundamental frequency contour features was realized and was applied to important word detection successfully.4. A method was develped to estimate t … More he feasibility of recognition candidates where a fundamental frequency contour was generated for each candidate, and was compared to the observed contour. The method was shown to be effective in detecting recognition error accompanied by accent type changes and/or syntactic boundary changes.5. A method was develped to model fundamental frequency contours statistically after representing them with several codes in moraic unit. The method was proved to be able to detect syntactic boundaries and to recognize accent types effectively.6. A method was develped to divide training data phoneme HMM into several clusters by inspecting HMM path of each data. By arranging a new HMM for each cluster, recognition rate was clearly shown to increase.7. A robust speech recognition method was developed based on Viterbi Baysian predictive classification. Validity of the method was shown with word recognition experiments under noisy conditions, where more than 10% improvement was observed as compared to conventional methods.After incorporating the developed methods above into a continuous speech recognition system, and their positive effects on recognition were proved. Less
期刊论文(87)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Keikichi Hirose, Atsuhiro Sakurai and Hiroyuki Konno: "USe of prosodic features in the recognition of continuous speech" Proc. International Conference on Spoken Language Processing, 3. 1123-1126 (1994)
Keikichi Hirose、Atsuhiro Sakurai 和 Hiroyuki Konno:“在连续语音识别中使用韵律特征”Proc。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Sumio Ohno, Hiyoya Fujisaki and Keikichi Hirose: "A method for word spotting in continious speech using both segmental and contextual likelihood scores" Proc. International Conference on Spoken Language Processing, 4. 2199-2202 (1994)
Sumio Ohno、Hiyoya Fujisaki 和 Keikichi Hirose:“一种使用分段和上下文似然分数在连续语音中进行单词识别的方法”Proc。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Nobuaki Minematsu and Keikichi Hirose: "Duration modeling with decreased intra-group temporal variation for HMM-based phoneme recognition" IEICE Trans. Information and Systems. E78-D-6. 654-661 (1995)
Nobuaki Minematsu 和 Keikichi Hirose:“基于 HMM 的音素识别,组内时间变化减少的持续时间建模”IEICE Trans。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
川波弘道: "対話音声の韻律的特徴に影響を与える要因の定量的分析" 日本音響学会秋季研究発表会講演論文集. I. 201-202 (1996)
Hiromichi Kawanami:“影响对话语音韵律特征的因素的定量分析”日本声学学会秋季研究会议论文集 I. 201-202 (1996)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
江輝: "A robust speech recognition method based on Bayes classification approach" 日本音響学会秋季研究発表会講演論文集. I. 149-150 (1996)
Eki:“基于贝叶斯分类方法的鲁棒语音识别方法”日本声学学会秋季会议论文集 I. 149-150 (1996)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 79 条
Pronunciation education system based on the systematization of non-mothor tongue speech prosody using generation process model and speech synthesis
-
批准号:24652115
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.33万
-
财政年份:2012
-
负责人:HIROSE Keikichi
-
依托单位:
Advanced method of prosody control in statistical-based speech synthesis using generation process model of fundamental frequency contours
-
批准号:24300068
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.4万
-
财政年份:2012
-
负责人:HIROSE Keikichi
-
依托单位:
Expressive Multi-language Speech Synthesis Based on the Generation Process Model and Its Use for Automatic Speech Translation
-
批准号:21300061
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.23万
-
财政年份:2009
-
负责人:HIROSE Keikichi
-
依托单位:
Synthesis of speech in any speaking styles based on corpus-based generation of prosodic features using the generation process model
-
批准号:17300055
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$10.79万
-
财政年份:2005
-
负责人:HIROSE Keikichi
-
依托单位:
High-quality Speech Synthesis based on Accurate Analysis Method and Statistical Method
-
批准号:12480079
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$6.4万
-
财政年份:2000
-
负责人:HIROSE Keikichi
-
依托单位:
Naturally Sounding Speech Synthesis and Recognition Based on the Formulation of Prosody
-
批准号:09480061
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$5.25万
-
财政年份:1997
-
负责人:HIROSE Keikichi
-
依托单位:
Develoment of Spoken Dialogue System for Japanese and Chinese
-
批准号:08558028
-
项目类别:Grant-in-Aid for Scientific Research (A)
-
资助金额:$5.12万
-
财政年份:1996
-
负责人:HIROSE Keikichi
-
依托单位:
Rule-Synthesis of Spoken Sentences for the Speech Dialogue Systems
-
批准号:03452288
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$3.97万
-
财政年份:1991
-
负责人:HIROSE Keikichi
-
依托单位:
Development of Output System of Announcing Speech with Input of Kanji-Kana Sentences
-
批准号:01850073
-
项目类别:Grant-in-Aid for Developmental Scientific Research (B).
-
资助金额:$3.33万
-
财政年份:1989
-
负责人:HIROSE Keikichi
-
依托单位:
海外基金