课题基金 / 基金详情

Estimation of vocal tract configurations from magnetic resonance images and synthesizing speech sounds

Estimation of vocal tract configurations from magnetic resonance images and synthesizing speech sounds
从磁共振图像估计声道配置并合成语音
批准号:
07650506
负责人:
SONODA Yorinobu
金额:
$1.34万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
1995
资助国家:
日本
项目状态:
已结题
起止时间:
1995 至 1996

项目摘要

项目成果

SONODA Yorinobu的其他基金

相关文献

中文摘要
翻译
本项目的主要研究是调查通过(1)磁共振图像(MRI)和(2)真实的语音信号估计的声道配置。特别是在这一阶段,对从真实的语音信号中估计声道结构进行了实验研究,在模拟人类语音产生机制的计算机系统上,采用“综合分析”算法估计声道结构。该模拟器由声源、声道和唇部辐射三部分组成。为了简化损失项插入道模型的过程,将每一部分与一个由频域和时域表示的混合系统相结合。声道模型由20个等长、不同截面积的圆柱形管道组成,利用声道模拟器合成了5个日语元音,其中声道的结构由MRI估计,共振峰模式(频率)由MRI估计。元音/a/和/i/的第一共振峰频率被估计为分别比真实的语音的第一共振峰频率低约120 Hz和70 Hz。除第四共振峰频率/i/外,其他元音的相对误差均在5%以内。对由真实的语音估计的形状合成的声音,实验结果与真实的声音的谱图有较好的近似,其相对误差均在3%以内。然而,/a/、/u/和/e/的第一共振峰频率的误差相对较大,其值在8 - 9%之间。
英文摘要
A main research of this project is investigating vocal tract configurations estimated by (1) magnetic resonance images (MRIs) and by (2) real speech signals. Especially in this term, experiments were conducted to the estimation of the vocal tract configuration from real speech signals.Developing a simulator on a computer system which is analogous to a mechanism of speech production process of human beings, the configuration of the tract was estimated by using "Analysis by Synthesis" algorithm. The simulator consists of three parts ; vocal source, vocal tract and lip radiation. Each part was combined with a hybrid system represented by frequency domain and time domain for simplicity of insertion of loss-term into the tract model. The model of vocal tract consists of 20 cylindrical tubes which are equal in length and different in cross sectional area.Five Japanese vowels were synthesized by using the vocal tract simulator where the configuration of the vocal tract was estimated by the MRIs, and their formant patterns (frequencies) were estimated. First formant frequencies of vowel /a/ and /i/ were estimated lower than those of real speech sound by about 120 Hz and 70 Hz, respectively. Relative errors were shown within 5 % in other vowels except fourth formant frequency of /i/.On sounds synthesized by the shape estimated from real speech sounds, experimental results were shown rather good approximation to spectral patterns of real sounds, and their relative errors were shown within 3 %. However, errors in first formant frequencies of /a/, /u/ and /e/ were relatively large and their values ranged in 8 - 9 %.
期刊论文(8)
专著(0)
科研奖励(0)
会议论文
K.Mori and Y.Sonoda: ""Relationship between lip shapes and acoustical characteristics during speech"" Proceedings of 3rd Joint meeting of Acoustical Society of America (ASA) and Acoustical Society of Japan (ASJ). 879-882 (1996)
K.Mori 和 Y.Sonoda:“演讲时嘴唇形状与声学特性之间的关系”美国声学学会 (ASA) 和日本声学学会 (ASJ) 第三次联席会议论文集。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
Kohichi Ogata: "Development of articulatory measuring system by using magnetometer and optical sensors" Proceedings of 3rd Joint meeting of ASA and ASJ. 889-894 (1996)
Kohichi Ogata:“使用磁力计和光学传感器开发关节测量系统”ASA 和 ASJ 第三次联席会议记录。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
K.Ogata and Y.Sonoda: ""Development of articulatory measuring system by using magnetometer and opticcal sensors"" Proceedings of 3rd Joint meeting of Acoustical Society of America (ASA) and Acoustical Society of Japan (ASJ). 889-894 (1996)
K.Ogata 和 Y.Sonoda:“使用磁力计和光学传感器开发关节测量系统”美国声学学会 (ASA) 和日本声学学会 (ASJ) 第三届联席会议论文集。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
井上,園田: "発話時の姿勢を変えた時の音響特徴" 電気関係学会九州支部連合会大会論文集. 1245. 771 (1995)
园田井上:“演讲时改变姿势时的声学特性”九州电气工程学会分会联合会会议记录1245. 771(1995)。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
6
    Articulatory model based on articulatory data and speech synthesis
    • 批准号:
      11650430
    • 项目类别:
      Grant-in-Aid for Scientific Research (C)
    • 资助金额:
      $2.3万
    • 财政年份:
      1999
    • 负责人:
      SONODA Yorinobu
    • 依托单位: