Modeling the articulatory space using a hypercube codebook for acoustic-to-articulatory inversion.
Modeling the articulatory space using a hypercube codebook for acoustic-to-articulatory inversion.
复制标题
使用超立方体密码本对发音空间进行建模,以进行声学到发音反转。
DOI:
10.1121/1.1921448
复制
发表时间:
2005
期刊:
影响因子:
--
通讯作者:
Y. Laprie
中科院分区:
文献类型:
--
作者:
Slim Ouni;Y. Laprie
Acoustic-to-articulatory inversion is a difficult problem mainly because of the nonlinearity between the articulatory and acoustic spaces and the nonuniqueness of this relationship. To resolve this problem, we have developed an inversion method that provides a complete description of the possible solutions without excessive constraints and retrieves realistic temporal dynamics of the vocal tract shapes. We present an adaptive sampling algorithm to ensure that the acoustical resolution is almost independent of the region in the articulatory space under consideration. This leads to a codebook that is organized in the form of a hierarchy of hypercubes, and ensures that, within each hypercube, the articulatory-to-acoustic mapping can be approximated by means of a linear transform. The inversion procedure retrieves articulatory vectors corresponding to acoustic entries from the hypercube codebook. A nonlinear smoothing algorithm together with a regularization technique is then used to recover the best articulatory trajectory. The inversion ensures that inverse articulatory parameters generate original formant trajectories with high precision and a realistic sequence of the vocal tract shapes.