DATA COLLECTION AND EVALUATION OF AURORA-2 JAPANESE CORPUS

DATA COLLECTION AND EVALUATION OF AURORA-2 JAPANESE CORPUS
复制标题

DOI:
--
复制
发表时间:
2003
期刊:
--
影响因子:
--
通讯作者:
Satoshi Nakamura;Kazumasa Yamamoto;K. Takeda;S. Kuroiwa;N. Kitaoka;Takeshi Yamada;M. Mizumachi;T. Nishiura;Masakiyo Fujimoto;A. Saso;Toshiki Endo
Satoshi Nakamura;Kazumasa Yamamoto;K. Takeda;S. Kuroiwa;N. Kitaoka;Takeshi Yamada;M. Mizumachi;T. Nishiura;Masakiyo Fujimoto;A. Saso;Toshiki Endo
中科院分区:
其他
文献类型:
--
作者:
Satoshi Nakamura;Kazumasa Yamamoto;K. Takeda;S. Kuroiwa;N. Kitaoka;Takeshi Yamada;M. Mizumachi;T. Nishiura;Masakiyo Fujimoto;A. Saso;Toshiki Endo

文献摘要

被引文献

相似文献

当语音识别系统暴露在噪声环境中时,它们仍然必须得到改进。为了实现这一目标,标准评价语料库和评价技术的发展至关重要。近年来,AURORA-2,3语料库及其评估方案对含噪语音识别研究产生了重要影响。介绍了一个日语带噪语音语料库AURORA-2 J及其评测脚本。AURORA-2 J是一个日语连接数字语料库。在ETSI AURORA集团的帮助下,数据收集和评估方案的设计方式与AURORA-2相同。此外,我们还收集了类似于AURORA-3的车内语音语料库。车内语音语料库包括在移动的汽车中收集的日语连接数字和命令词。本文描述了数据收集、基线脚本及其基线性能。
Speech recognition systems must still be improved when they are exposed to noisy environments. For this improvement, developments of the standard evaluation corpus and assessment technologies are essential. Recently the AURORA-2,3 corpus and their evaluation scenarios have had significant impacts on noisy speech recognition research. This paper introduces a Japanese noisy speech corpus and its evaluation scripts, called AURORA-2J. The AURORA-2J is a Japanese connected digits corpus. The data collection and evaluation scenarios are designed in the same way as AURORA-2 with the help of ETSI AURORA group. Furthermore, we have collected in-car speech corpus similar to AURORA-3. The in-car speech corpus includes Japanese connected digits and command words collected in a moving car. This paper describes the data collection, baseline scripts, and its baseline performance.