Automatic voice building for flexible speech synthesis
Automatic voice building for flexible speech synthesis
批准号:
14380160
负责人:
TOKUDA Keikhi
金额:
$5.95万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2002
资助国家:
日本
项目状态:
已结题
起止时间:
2002 至 2004
中文摘要
随着大型语音数据库的不断增加,人们可以通过应用统计学习算法来构建语音合成系统,这被称为数据驱动或基于语料库的方法。这些系统可以自动训练,不仅可以生成自然和高质量的合成语音,而且可以再现原始说话人的语音特征。然而,为了使整个语音构建过程完全自动化,我们需要以自动的方式构建语音数据库。在这项研究工作中,我们研究了自动语音建设技术的HMM为基础的语音合成系统,可以合成语音与各种语音质量。首先,我们实现了一个基于GUI的标签工具,称为PLEd(韵律和语言标签编辑器)。然后,为了构建一个自动语音构建系统,我们开发了一个自动口音标注技术。实验结果表明,利用该系统,我们成功地标注了口音信息。
英文摘要
The increasing availability of large speech databases makes it possible to construct speech synthesis systems, which are referred to as data-driven or corpus-based approach, by applying statistical learning algorithms. These systems, which can be automatically trained, not only generate natural and high quality synthetic speech but also can reproduce voice characteristics of the original speaker. However, to make the whole voice building process fully-automatic, we need to construct speech databases in an automatic way. In this research work, we investigate automatic voice building techniques for an HMM-based speech synthesis system which can synthesize speech with various voice qualities. First, we implemented an GUI-based labeling tool, called PLEd (Prosody and Linguistic Label Editor). Then, in order to construct an automatic voice building system, we have developed an automatic accent labeling technique. It has been shown that by using the developed system, we have successfully label accent information.
期刊论文(341)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
10.1109/wss.2002.1224415
发表时间:
2002
期刊:
Proceedings of 2002 IEEE Workshop on Speech Synthesis, 2002.
影响因子:
--
作者:
[Keiichi Tokuda;H. Zen;Alan W. Black]
通讯作者:
Keiichi Tokuda;H. Zen;Alan W. Black
DOI:
--
发表时间:
2005
期刊:
2005 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2005)
影响因子:
--
作者:
[Yusuke Kida, Hiroyoshi Yamamoto, Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura]
通讯作者:
Tadashi Kitamura
スペクトル・FO・継続長決定木の同時バックオフに基づくHMM音声合成
基于频谱、FO 和持续时间树同时退避的 HMM 语音合成
DOI:
--
发表时间:
2005
期刊:
日本音響学会2005年春季研究発表会講演論文集 vol.I
影响因子:
--
作者:
[片岡俊介, 全 炳河, 南角吉彦, 徳田恵一, 北村 正]
通讯作者:
北村 正
全 炳河, 徳田 恵一, 北村 正: "静的・動的特徴の明示的な関係によりHMMから導出されるトラジェクトリモデル"電子情報通信学会技術研究報告. vol.103, no.519. 55-60 (2003)
Binghe Zen、Keiichi Tokuda、Tadashi Kitamura:“使用静态和动态特征之间的显式关系导出的轨迹模型”IEICE 技术研究报告,第 103 卷,第 55-60 期(2003 年)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
全炳河, 徳田恵一, 北村正: "トラジェクトリHMMのためのViterbiアルゴリズム"日本音響学会2003年秋季研究発表会講演論文集. vol.I, 2-6-3. 65-66 (2003)
Zen Binghe、Keiichi Tokuda、Tadashi Kitamura:“轨迹 HMM 的维特比算法”日本声学学会 2003 年秋季研究会议论文集,第 I 卷,2-6-3(2003 年)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 156 条
海外基金