High-quality Speech Synthesis based on Accurate Analysis Method and Statistical Method
High-quality Speech Synthesis based on Accurate Analysis Method and Statistical Method
批准号:
12480079
负责人:
HIROSE Keikichi
金额:
$6.4万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2000
资助国家:
日本
项目状态:
已结题
起止时间:
2000 至 2002
中文摘要
点击翻译按钮获取中文摘要
英文摘要
The original research plan, which aims at realizing high-quality speech synthesis through utilizing accurate pole-zero information of vocal transfer function for segmental feature generation and applying the functional model constraints for prosodic feature generation, was accomplished with the following results :1. A successive approximation was applied to ARX analysis enabling accurate pole-zero estimation. The method was combined with our formerly developed terminal analogue synthesizer to construct a analysis-synthesis workbench. Using this, we succeeded to improve the quality of liquid sound.2. A speech synthesizer, hybrid of terminal analogue and waveform concatenation, was developed. A high-quality speech synthesis was realized.3. A method was developed for stable formant extraction, which was based on AR-HMM modeling, representing source waveform using HMM. Result of speech synthesis experiment showed that the method could generate high-quality even for a large F0 (fundamental … More frequency) change.4. By adding natural waveform of junction periods in the spectral domain with appropriate weighting to the concatenated speech, we successfully realized a smooth spectral transition. Also we developed a method to effectively reduce the corpus size for concatenative synthesis by the weighted VQ according to the frequency.5. The necessary data size for speaker adaptation was investigated form the viewpoint of speech quality after developing a HMM speech synthesizer. It was shown that a good quality was obtainable 10 and more sentences.6. F0 contour generation was realized by estimating the generation process model parameters using statistical methods. A high speech quality was realized only from a small speech corpus by using linguistic information such as on direct modification relations of words. Also we succeeded to estimate the accent phrase boundaries form text using the same statistical framework. Furthermore, F0 contour generation and phoneme length estimation were realized for emotional speech with a good result.7. A method for automatically estimating F0 contour generation process model commands was realized. Using the method, a prosodic corpus was made. This corpus is indispensable for the above F0 contour generation.8. A rule for controlling mora duration for dialogue-like speech synthesis was constructed. The result of the speech synthesis experiment showed the validity of the rule. Less
期刊论文(142)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
広瀬啓吉: "音声情報処理におけるパラ・非言語情報"日本音響学会秋季講演論文集. I. 243-246 (2002)
Keikichi Hirose:“语音信息处理中的副/非语言信息”日本声学学会秋季会议记录 I. 243-246 (2002)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
広瀬啓吉: "音声合成研究への招待 -自由な合成の実現に向けて-"情報処理. 43・3. 321-324 (2002)
广濑圭吉:“语音合成研究的邀请 - 实现自由合成 -” 信息处理 321-324。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Nobuyuki Nishizawa, Nobuaki Minematsu, and Keikichi Hirose: "Formant speech synthesis partly using waveform concatenative synthesis -Experimental study on VCV sounds-"IEICE Technical Report. SP2001-20. 35-42 (2001)
Nobuyuki Nishizawa、Nobuaki Minematsu 和 Keikichi Hirose:“部分使用波形连接合成的共振峰语音合成 -VCV 声音的实验研究 -”IEICE 技术报告。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
江藤雅哉: "生成過程モデルと統計的手法による統語構造を考慮した基本周波数パターンの生成"電子情報通信学会技術研究報告(音声研究会). 17-22 (2002)
Masaya Eto:“使用生成过程模型和统计方法考虑句法结构的基频模式的生成”IEICE 技术报告(语音研究组)17-22 (2002)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
西沢信行: "Development of a formant-based analysis-synthesis system and generation of high quality liquid sounds of Japanese"Proc.International Conf.on Spoken Language Processing. 1. 725-728 (2000)
Nobuyuki Nishizawa:“开发基于共振峰的分析合成系统并生成高质量的日语液体声音”Proc.International Conf.on Spoken Languageprocessing 1. 725-728 (2000)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 62 条
Pronunciation education system based on the systematization of non-mothor tongue speech prosody using generation process model and speech synthesis
-
批准号:24652115
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.33万
-
财政年份:2012
-
负责人:HIROSE Keikichi
-
依托单位:
Advanced method of prosody control in statistical-based speech synthesis using generation process model of fundamental frequency contours
-
批准号:24300068
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.4万
-
财政年份:2012
-
负责人:HIROSE Keikichi
-
依托单位:
Expressive Multi-language Speech Synthesis Based on the Generation Process Model and Its Use for Automatic Speech Translation
-
批准号:21300061
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.23万
-
财政年份:2009
-
负责人:HIROSE Keikichi
-
依托单位:
Synthesis of speech in any speaking styles based on corpus-based generation of prosodic features using the generation process model
-
批准号:17300055
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$10.79万
-
财政年份:2005
-
负责人:HIROSE Keikichi
-
依托单位:
Naturally Sounding Speech Synthesis and Recognition Based on the Formulation of Prosody
-
批准号:09480061
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$5.25万
-
财政年份:1997
-
负责人:HIROSE Keikichi
-
依托单位:
Develoment of Spoken Dialogue System for Japanese and Chinese
-
批准号:08558028
-
项目类别:Grant-in-Aid for Scientific Research (A)
-
资助金额:$5.12万
-
财政年份:1996
-
负责人:HIROSE Keikichi
-
依托单位:
Formulation of Prosodic Features of Speech and its Application to Continuous Speech Recognition
-
批准号:06452397
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.8万
-
财政年份:1994
-
负责人:HIROSE Keikichi
-
依托单位:
Rule-Synthesis of Spoken Sentences for the Speech Dialogue Systems
-
批准号:03452288
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$3.97万
-
财政年份:1991
-
负责人:HIROSE Keikichi
-
依托单位:
Development of Output System of Announcing Speech with Input of Kanji-Kana Sentences
-
批准号:01850073
-
项目类别:Grant-in-Aid for Developmental Scientific Research (B).
-
资助金额:$3.33万
-
财政年份:1989
-
负责人:HIROSE Keikichi
-
依托单位: