Automatic discovery of a phonetic inventory for unwritten languages for statistical speech synthesis
Automatic discovery of a phonetic inventory for unwritten languages for statistical speech synthesis
复制标题
自动发现非书面语言的语音库存以进行统计语音合成
DOI:
10.1109/icassp.2014.6854069
复制
发表时间:
2014
期刊:
影响因子:
--
通讯作者:
A. Black
中科院分区:
文献类型:
--
作者:
P. Muthukumar;A. Black
Speech synthesis systems are typically built with speech data and transcriptions. In this paper, we try to build synthesis systems when no transcriptions or knowledge about the language are available. It is usually necessary to at least possess phonetic knowledge about the language. In this paper, we propose an automated way of obtaining phones and phonetic knowledge about the corpus at hand by making use of Articulatory Features (AFs). An Articulatory Feature predictor is trained on a bootstrap corpus in an arbitrary other language using a three-hidden layer neural network. This neural network is run on the speech corpus to extract AFs. Hierarchical clustering is used to cluster the AFs into categories i.e. phones. Phonetic information about each of these inferred phones is obtained by computing the mean of the AFs in each cluster. Results of systems built with this framework in multiple languages are reported.