Japanese Dictation Toolkit-1997 version-

Japanese Dictation Toolkit-1997 version-
复制标题

日语听写工具包-1997年版-

DOI:
10.1250/ast.20.233
复制
发表时间:
1999
期刊:
The Journal of The Acoustical Society of Japan (e)
影响因子:
--
通讯作者:
K. Shikano
K. Shikano
中科院分区:
--
文献类型:
--
作者:
Tatsuya Kawahara;Akinobu Lee;Tetsunori Kobayashi;K. Takeda;N. Minematsu;K. Itou;Akinori Ito;Mikio Yamamoto;A. Yamada;T. Utsuro;K. Shikano

文献摘要

参考文献

被引文献

相似文献

日本听写工具包已被设计和开发为日本LVCSR的基线平台(大型词汇连续语音识别)。该平台由标准识别引擎,日本电话模型和日本统计语言模型组成。我们设置了各种日本手机HMM,从不依赖上下文的单声到成千上万个州的Triphone型号。他们接受了ASJ(日本声学学会)数据库的培训。词典和单词n-gram(2克和3克)模型是由Mainichi报纸的语料库构建的。识别引擎Julius是为了评估声学和语言模型的评估。作为这些模块的集成系统,我们实施了基线5,000字的命令系统,并评估了各种组件。该软件存储库可向公众使用。
The Japanese Dictation Toolkit has been designed and developed as a baseline platform for Japanese LVCSR (Large Vocabulary Continuous Speech Recognition). The platform consists of a standard recognition engine, Japanese phone models and Japanese statistical language models. We set up a variety of Japanese phone HMMs from a contextindependent monophone to a triphone model of thousands of states. They are trained with ASJ (The Acoustical Society of Japan) databases. A lexicon and word N-gram (2-gram and 3-gram) models are constructed with a corpus of Mainichi newspaper. The recognition engine JULIUS is developed for evaluation of both acoustic and language models. As an integrated system of these modules, we have implemented a baseline 5, 000-word dictation system and evaluated various components. The software repository is available to the public.
T.Kawahara,T.Kobayashi,K.Takeda,N.Minematsu,K.Itou,M.Yamamoto,A.Yamada,T.Utsuro,K.Shikano:“用于日语大词汇连续语音识别的可共享软件存储库”Proc。
DOI: --
发表时间: --
期刊:
影响因子: --
作者:
通讯作者: --