Japanese speech databases for robust speech recognition

Japanese speech databases for robust speech recognition
复制标题

用于强大语音识别的日语语音数据库

DOI:
10.1109/icslp.1996.607241
复制
发表时间:
1996
期刊:
Proceeding of Fourth International Conference on Spoken Language Processing. ICSLP '96
影响因子:
--
通讯作者:
Y. Sagisaka
Y. Sagisaka
中科院分区:
--
文献类型:
--
作者:
Atsushi Nakamura;S. Matsunaga;Tohru Shimizu;M. Tonomura;Y. Sagisaka

文献摘要

被引文献

相似文献

在ATR,下一代语音翻译系统正在开发中,奖励自然的跨语言交流。为了应对新系统对语音识别技术的各种要求,进一步的研究工作应强调对大词汇量、快速自发语音中经常出现的语音变化和说话人变化的鲁棒性。这些不仅是语音翻译需要解决的关键问题,也是语音识别在现实环境中的普遍应用所需要解决的关键问题。针对语音识别中存在的问题,设计了三个大型语音数据库,并介绍了数据采集的现状。
At ATR, a next-generation speech translation system is under development rewards natural trans-language communication. To cope with the various requirements to speech recognition technology for the new system, further research efforts should emphasize the robustness for large vocabulary, speaking variations often found in fast spontaneous speech and speaker variances. These are key problems to be solved nor only for speech translation but also for the general use of speech recognition in real environments. Three large speech databases are designed to cope with these problems in speech recognition and the current status of data collection is reported.