A Step-by-Step Process for Building TTS Voices Using Open Source Data and Frameworks for Bangla, Javanese, Khmer, Nepali, Sinhala, and Sundanese
A Step-by-Step Process for Building TTS Voices Using Open Source Data and Frameworks for Bangla, Javanese, Khmer, Nepali, Sinhala, and Sundanese
复制标题
使用孟加拉语、爪哇语、高棉语、尼泊尔语、僧伽罗语和巽他语的开源数据和框架构建 TTS 语音的分步过程
DOI:
--
复制
发表时间:
2018
期刊:
影响因子:
--
通讯作者:
Linne Ha
中科院分区:
文献类型:
--
作者:
Keshan Sanjaya Sodimana;Pasindu De Silva;Supheakmungkol Sarin;Oddur Kjartansson;Martin Jansche;Knot Pipatsrisawat;Linne Ha
The availability of language resources is vital for the development of text-to-speech (TTS) systems. Thus, open source resources are highly beneficial for TTS research communities focused on low-resourced languages. In this paper, we present data sets for 6 low-resourced languages that we open sourced to the public. The data sets consist of audio files, pronunciation lexicons, and phonology definitions for Bangla, Javanese, Khmer, Nepali, Sinhala, and Sundanese. These data sets are sufficient for building voices in these languages. We also describe a recipe for building a new TTS voice using our data together with openly available resources and tools.