AGH corpus of Polish speech

AGH corpus of Polish speech
复制标题

AGH 波兰语语音语料库

DOI:
10.1007/s10579-015-9302-y
复制
发表时间:
2016
影响因子:
2.7
通讯作者:
D. Skurzok
D. Skurzok
中科院分区:
计算机科学4区
文献类型:
--
作者:
Piotr Żelasko;B. Ziółko;T. Jadczyk;D. Skurzok

文献摘要

参考文献

被引文献

相似文献

波兰语的语音,这已被收集的自动语音识别(ASR)和文本到语音(TTS)系统的应用程序的目的,语料库。语料库由几组录音组成:读句子,口语命令,语音平衡的TTS训练语料库,电话语音和其他。总之,记录持续时间超过25小时。唯一扬声器的数量达到166。他们大多数在20-35岁年龄组,其中三分之一是女性。 在较大的文本资源的独特的词出现频率的分析已经得出结论。其中,最常出现的单词被发现并呈现出来。语料库被用作ASR系统的训练数据。使用我们的语料库交叉验证训练和测试的SARMATA ASR系统的结果表明,短语识别率为91.9%。语料库进行了额外的评价,在对比测试对CORPORA语料库,这表明有利于我们的语料库的短语识别率的大幅增加。
A corpus of Polish speech, which has been collected for the purpose of automatic speech recognition (ASR) and text-to-speech (TTS) systems applications, is presented. The corpus consists of several groups of recordings: read sentences, spoken commands, a phonetically balanced TTS training corpus, telephonic speech and others. In summary duration of recordings is above 25 h. Number of unique speakers amounts to 166. The majority of them being in an age group of 20–35 and one third of them being female. Analysis of unique word occurrence frequency in relation to larger text resources has been concluded. From them, most commonly appearing words have been found and presented. The corpus was used as training data for the ASR system. Results of cross-validation training and testing the SARMATA ASR system using our corpus have shown that phrase recognition rate is 91.9 %. The corpus was additionally evaluated in comparative test against the CORPORA corpus, which had shown major increase in phrase recognition rate in favour of our corpus.
DOI: 10.21437/icslp.2002-151
发表时间: 2002-09
期刊: --
影响因子: --
作者:
Tanja Schultz
通讯作者: Tanja Schultz