Towards Spontaneous Speech Translation
Towards Spontaneous Speech Translation
复制标题
走向自发语音翻译
DOI:
10.1016/j.procs.2016.04.032
复制
发表时间:
1994
影响因子:
4.7
通讯作者:
A. Waibel
中科院分区:
文献类型:
--
作者:
M. Woszczyna;N. Aoki;Finn Dag Buø;N. Coccaro;Keiko Horiguchi;T. Kemp;A. Lavie;A. McNair;T. Polzin;I. Rogina;C. Rosé;T. Schultz;B. Suhm;M. Tomita;A. Waibel
In this work we make use of unsupervised linear discriminant analysis (LDA) to support acoustic unit discovery in a zero resource scenario. The idea is to automatically find a mapping of feature vectors into a subspace that is more suitable for Dirichlet process Gaussian mixture model (DPGMM) based clustering, without the need of supervision. Supervised acoustic modeling typically makes use of feature transformations such as LDA to minimize intra-class discriminability, to maximize inter-class discriminability and to extract relevant informations from high-dimensional features spanning larger contexts. The need of class labels makes it difficult to use this technique in a zero resource setting where the classes and even their amount are unknown. To overcome this issue we use a first iteration of DPGMM clustering on standard features to generate labels for the data, that serve as basis for learning a proper transformation. A second clustering operates on the transformed features. The application of unsupervised LDA demonstrably leads to better clustering results given the unsupervised data. We show that the improved input features consistently outperform our baseline input features.