A Novel Method for Constructing 3D Geometric Articulatory Models

A Novel Method for Constructing 3D Geometric Articulatory Models
复制标题

DOI:
10.1007/s11265-015-1002-8
复制
发表时间:
2015-05
期刊:
Journal of Signal Processing Systems
影响因子:
--
通讯作者:
Jianguo Wei;Jie Liu;Qiang Fang;Wenhuan Lu;J. Dang;K. Honda
Jianguo Wei;Jie Liu;Qiang Fang;Wenhuan Lu;J. Dang;K. Honda
中科院分区:
其他
文献类型:
--
作者:
Jianguo Wei;Jie Liu;Qiang Fang;Wenhuan Lu;J. Dang;K. Honda

文献摘要

被引文献

相似文献

本研究描述了一种新的方法,通过考虑语音器官的生理边界,构建基于磁共振成像数据的几何发音模型。对建模过程进行了两方面的改进:i)将从不同视点拍摄的图像进行组合,以提高轮廓标注的准确性。ii)参照解剖结构对语音器官的网格进行建模。定性和定量评价均表明,该方法优于传统方法。根据不同发音器官的网格,采用线性成分分析法提取控制参数。每个语音器官可以使用三个或更少的控制参数来描述。重建后的模型与真实的数据之间的平均误差小于1.0 mm。这也是首次基于中国MRI数据构建3D声道模型。这将有助于汉语言语生成相关问题的理论研究和实际应用。
This study describes a novel method of constructing a geometric articulatory model based on magnetic resonance imaging data by taking the physiological boundaries of speech apparatus into account. Two improvements have been made to the modeling process: i) Images taken from different viewpoints are combined to improve the accuracy of outline annotation. ii) Speech organs’ meshes are modeled with reference to the anatomical structures. Both qualitative and quantitative evaluations indicated that the proposed method surpasses the conventional method. Based on the meshes of the speech organs associated with different articulations, the linear component analysis was used to extract the control parameters. Each speech organ can be described using three control parameters or fewer. After the reconstruction, the average error between model and real data was less than 1.0 mm. This is also the first effort made to construct a 3D vocal tract model based on Chinese MRI data. It will facilitate the theoretical study and practical use in Chinese-speech-production related issues.