课题基金 / 基金详情

Realization of average-voice-based speech synthesis with diverse voices and speaking styles

Realization of average-voice-based speech synthesis with diverse voices and speaking styles
实现基于平均语音的多种语音和说话风格的语音合成
批准号:
15300055
负责人:
KOBAYASHI Takao
金额:
$5.76万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2003
资助国家:
日本
项目状态:
已结题
起止时间:
2003 至 2005

项目摘要

项目成果

KOBAYASHI Takao的其他基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
The purpose of this research is the realization of text-to-speech synthesis that can generate speech with an arbitrarily given speaker's voice and diverse speaking styles and/or emotional expressions. We have obtained the following results.1. Speech synthesis with arbitrary speaker's voice based on average voice modelWe have proposed a new training method of average voice model for speech synthesis in which an arbitrary speaker's voice is generated based on speaker adaptation. We have also proposed new speaker adaptation techniques based on hidden semi-Markov model (HSMM) that can model phone duration more precisely than the conventional hidden Markov model (HMM). From the results of objective and subjective evaluation tests, it has been shown that the average-voice-model-based speech synthesis can generates natural sounding speech of the target speaker.2. Speech synthesis with various speaking styles and emotional expressionsWe have proposed several approaches to the realization of emotional expressivity and speaking style variability in text-to-speech synthesis. We investigated two methods for modeling speaking styles and/or emotional expressions based on an HMM-based speech synthesis framework, and then proposed some approaches to adding various styles to synthetic speech, such as style interpolation, style morphing, style adaptation, and style control techniques. From results of subjective experiments, we have shown that the effectiveness of the proposed approaches.3. ProsodyWe have developed a robust fundamental frequency estimation and voice/unvoiced determination technique based on instantaneous frequency amplitude spectrum. We have also proposed modeling techniques for phone duration and pause for high quality text-to-speech synthesis.
期刊论文(147)
专著(0)
科研奖励(0)
会议论文
DOI: --
发表时间: 2003
期刊:
影响因子: --
作者: [J. Yamagishi;T. Masuko;Takao Kobayashi]
通讯作者: J. Yamagishi;T. Masuko;Takao Kobayashi
重回帰HSMMを用いた合成音声のスタイル制御
使用多元回归 HSMM 对合成语音进行风格控制
DOI: --
发表时间: 2006
期刊: 電子情報通信学会技術研究報告,SP2005-160 105・572
影响因子: --
作者: [Yoshihide Kato, Tomohiro Ohno, Nobuo Kawaguchi, Makoto Tachibana, Yoshihide Kato, 能勢 隆]
通讯作者: 能勢 隆
A style adaptation technique for speech synthesis using HSMM and suprasegamental features
使用 HSMM 和超段特征的语音合成风格适应技术
DOI: --
发表时间: 2006
期刊: IEICE Trans.Information and Systems E89-D・3
影响因子: --
作者: [Makoto Tachibana]
通讯作者: Makoto Tachibana
DOI: --
发表时间: 2006
期刊: Proc.the 2006 Spring Meeting of the Acoustical Society of Japan
影响因子: --
作者: [Makoto Tachibana]
通讯作者: Makoto Tachibana
103
    Research on speech synthesis using non-parametric modeling based on Gaussian process regression
    • 批准号:
      25540065
    • 项目类别:
      Grant-in-Aid for Challenging Exploratory Research
    • 资助金额:
      $2.41万
    • 财政年份:
      2013
    • 负责人:
      KOBAYASHI Takao
    • 依托单位:
    Research on advanced robust speech synthesis and its applications to multi-lingual speech communication
    • 批准号:
      24300071
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $9.15万
    • 财政年份:
      2012
    • 负责人:
      KOBAYASHI Takao
    • 依托单位:
    Research on robust spoken language interfaces for diverse voice variability and expressivity
    • 批准号:
      21300063
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $6.24万
    • 财政年份:
      2009
    • 负责人:
      KOBAYASHI Takao
    • 依托单位:
    Development of highly sensitive optical fiber sensor system for wide-area disaster prevention
    • 批准号:
      18510145
    • 项目类别:
      Grant-in-Aid for Scientific Research (C)
    • 资助金额:
      $2.39万
    • 财政年份:
      2006
    • 负责人:
      KOBAYASHI Takao
    • 依托单位: