Modeling the articulatory space using a hypercube codebook for acoustic-to-articulatory inversion.

Modeling the articulatory space using a hypercube codebook for acoustic-to-articulatory inversion.
复制标题

使用超立方体密码本对发音空间进行建模,以进行声学到发音反转。

DOI:
10.1121/1.1921448
复制
发表时间:
2005
期刊:
The Journal of the Acoustical Society of America
影响因子:
--
通讯作者:
Y. Laprie
Y. Laprie
中科院分区:
--
文献类型:
--
作者:
Slim Ouni;Y. Laprie

文献摘要

被引文献

相似文献

声学到发音反演是一个困难的问题,主要是因为发音和声学空间之间的非线性和这种关系的非唯一性。为了解决这个问题,我们已经开发了一种反演方法,提供了一个完整的描述可能的解决方案,没有过多的限制,并检索现实的时间动态的声道形状。我们提出了一种自适应采样算法,以确保声学分辨率几乎是独立的发音空间中的区域正在考虑。这导致了以超立方体层次结构形式组织的码本,并确保在每个超立方体内,可以通过线性变换来近似发音到声学映射。反演过程从超立方体码本中检索对应于声学条目的发音向量。一个非线性平滑算法与正则化技术,然后使用恢复最佳的发音轨迹。该反演确保了反演发音参数以高精度和声道形状的真实序列生成原始共振峰轨迹。
Acoustic-to-articulatory inversion is a difficult problem mainly because of the nonlinearity between the articulatory and acoustic spaces and the nonuniqueness of this relationship. To resolve this problem, we have developed an inversion method that provides a complete description of the possible solutions without excessive constraints and retrieves realistic temporal dynamics of the vocal tract shapes. We present an adaptive sampling algorithm to ensure that the acoustical resolution is almost independent of the region in the articulatory space under consideration. This leads to a codebook that is organized in the form of a hierarchy of hypercubes, and ensures that, within each hypercube, the articulatory-to-acoustic mapping can be approximated by means of a linear transform. The inversion procedure retrieves articulatory vectors corresponding to acoustic entries from the hypercube codebook. A nonlinear smoothing algorithm together with a regularization technique is then used to recover the best articulatory trajectory. The inversion ensures that inverse articulatory parameters generate original formant trajectories with high precision and a realistic sequence of the vocal tract shapes.