课题基金 / 基金详情

Synthesis of speech in any speaking styles based on corpus-based generation of prosodic features using the generation process model

Synthesis of speech in any speaking styles based on corpus-based generation of prosodic features using the generation process model
使用生成过程模型基于语料库生成韵律特征来合成任何说话风格的语音
批准号:
17300055
负责人:
HIROSE Keikichi
金额:
$10.79万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2005
资助国家:
日本
项目状态:
已结题
起止时间:
2005 至 2007

项目摘要

项目成果

HIROSE Keikichi的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Research works were conducted to establish a corpus-based speech synthesis method, which is based on generation process model of fundamental frequency contours and can generate high-quality speech in any speaking styles. The original research plan was fulfilled with the following results :1. A method was developed to predict the command parameters of the generation process model using binary decision trees with inputs such as linguistic information available by parsing texts, and thus to synthesize fundamental frequency contours. An integrated method of prosodic control was realized by integrating the above method with other methods using binary decision trees to predict pause positions and lengths and phoneme durations. The validity of the method was shown through experiments on speech synthesis of various styles including emotional speech. A method was also developed to automatically extract the command parameters from observed fundamental frequency contours using binary decision tre … More es. It was shown that the accuracy of extraction increased by including linguistic information of the text into inputs of the trees.2. Binary decision trees were constructed to predict deviations in phrase and accent commands of the utterances with specific focuses from those without. Their inputs are accent types and positions in sentences of the focused words, and command values of the corresponding parts of the utterances without specific focus. An appropriate focus control was realized by modifying the phrase and accent commands predicted by the method in section 1 based on the predicted deviations.3. A two-step method was developed for generating fundamental frequency contours of Standard Chinese. It first generates phrase components in a corpus-based way, and then generates tone components in a corpus-based way. The method has a high flexibility in synthesizing fundamental frequency contours. As an example of flexible control, it was shown that proper focus control could be realized in a simple set of rules.4. Speech synthesis systems were constructed for Japanese and Chinese by integrating methods developed in sections 1 and 2 above with HMM speech synthesis. It was shown that synthetic speech with higher natural ness could be realized by our system than using "full" HMM synthesizer, where prosodic control was done in HMM framework. It was also shown that various styles of synthetic speech could be realized by our system.5. Spoken dialogue systems for road guidance and TV program guidance were constructed using the above speech synthesis systems. The validity of the developed speech synthesis method was proved through experiments on the control of speaking styles of reply speech depending on the user's characters and situations. Less
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Estimation of intonation variation with constrained tone transformations
通过受约束的声调变换来估计声调变化
DOI: --
发表时间: 2005
期刊: Proc. 9^<th> European Conference on Speech Communication and Technology (INTERSPEECH) CD-ROM
影响因子: --
作者: [Keikichi Hirose, Yusuke Furuyama, Nobuaki Minematsu, Keikichi Hirose, Keikichi Hirose, 広瀬啓吉, Keikichi Hirose, Keikichi Hirose, Quinghua Sun, Jinfu Ni]
通讯作者: Jinfu Ni
日本語テキスト音声合成用アクセント結合規則の改良
改进日语文本语音合成的重音组合规则
DOI: --
发表时间: 2005
期刊: 日本音響学会講演論文集 CD-ROM
影响因子: --
作者: [Keikichi Hirose, Yusuke Furuyama, Nobuaki Minematsu, Keikichi Hirose, Keikichi Hirose, 広瀬啓吉, Keikichi Hirose, Keikichi Hirose, Quinghua Sun, Jinfu Ni, Keikichi Hirose, 黒岩 龍]
通讯作者: 黒岩 龍
DOI: --
发表时间: 2006
期刊: Journal of Acoustical Society of America 119・3
影响因子: --
作者: [Corinne Touati, Atsushi Inoie, Hisao Kameda, H.Kameda, Jinfu Ni]
通讯作者: Jinfu Ni
道案内音声対話システムへの概念音声合成に基づく応答生成手法の実装とその評価
基于概念语音合成的路线引导语音对话系统响应生成方法的实现与评估
DOI: --
发表时间: 2007
期刊: 情報処理学会論文誌 48
影响因子: --
作者: [Qinghua Sun, Keikichi Hirose, Nobuaki Minematsu, 八木裕司]
通讯作者: 八木裕司
36
    Pronunciation education system based on the systematization of non-mothor tongue speech prosody using generation process model and speech synthesis
    • 批准号:
      24652115
    • 项目类别:
      Grant-in-Aid for Challenging Exploratory Research
    • 资助金额:
      $2.33万
    • 财政年份:
      2012
    • 负责人:
      HIROSE Keikichi
    • 依托单位:
    Advanced method of prosody control in statistical-based speech synthesis using generation process model of fundamental frequency contours
    • 批准号:
      24300068
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $11.4万
    • 财政年份:
      2012
    • 负责人:
      HIROSE Keikichi
    • 依托单位:
    Expressive Multi-language Speech Synthesis Based on the Generation Process Model and Its Use for Automatic Speech Translation
    • 批准号:
      21300061
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $11.23万
    • 财政年份:
      2009
    • 负责人:
      HIROSE Keikichi
    • 依托单位:
    High-quality Speech Synthesis based on Accurate Analysis Method and Statistical Method
    • 批准号:
      12480079
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $6.4万
    • 财政年份:
      2000
    • 负责人:
      HIROSE Keikichi
    • 依托单位:
    海外基金