Rapid Development of TTS Corpora for Four South African Languages

Rapid Development of TTS Corpora for Four South African Languages
复制标题

四种南非语言 TTS 语料库的快速开发

DOI:
10.21437/interspeech.2017-1139
复制
发表时间:
2017
期刊:
--
影响因子:
--
通讯作者:
Linne Ha
Linne Ha
中科院分区:
--
文献类型:
--
作者:
D. V. Niekerk;C. V. Heerden;Marelie Hattingh Davel;N. Kleynhans;Oddur Kjartansson;Martin Jansche;Linne Ha

文献摘要

被引文献

相似文献

本文描述了四种南非语言的文本转语音语料库的开发。随后的办法调查了使用低成本方法的可能性,包括非正式录音环境和未经训练的志愿发言者。为了实现这一目标,以及今后扩大语料库以增加南非11种正式语文的覆盖范围的另一个目标,必须试验多讲者和代码转换数据。整个过程和相关的观察都是详细的。最新版本的语料库可以在开源许可下下载,未来可能会有进一步的发展和完善。
This paper describes the development of text-to-speech corpora for four South African languages. The approach followed investigated the possibility of using low-cost methods including informal recording environments and untrained volunteer speakers. This objective and the additional future goal of expanding the corpus to increase coverage of South Africa’s 11 official languages necessitated experimenting with multi-speaker and code-switched data. The process and relevant observations are detailed throughout. The latest version of the corpora are available for download under an open-source licence and will likely see further development and refinement in future.