Fast bootstrapping of LVCSR systems with multilingual phoneme sets

Fast bootstrapping of LVCSR systems with multilingual phoneme sets
复制标题

具有多语言音素集的 LVCSR 系统的快速引导

DOI:
10.21437/eurospeech.1997-141
复制
发表时间:
1997
影响因子:
4.3
通讯作者:
A. Waibel
A. Waibel
中科院分区:
医学2区
文献类型:
--
作者:
Tanja Schultz;A. Waibel

文献摘要

被引文献

相似文献

本文描述了一种基于多语言音素集的大词汇量连续语音识别系统的高效方法。为了评估这种技术,我们收集了多语言数据库GlobalPhone,该数据库目前包含9种不同的语言。开发了基于德语、英语、日语和西班牙语四种语言的多语言识别器(MULTI)作为源系统。同样,该系统对语言识别也非常有用,实现了100%的语言识别率。基于MULTI系统,我们在汉语、克罗地亚语和土耳其语等完全不同的语言上评估了我们的引导技术。
In this paper we described an e cient method to bootstrap continuously spoken, large vocabulary speech recognition systems by multilingual phoneme sets. To evaluate this techniques we collected the multilingual database GlobalPhone which currently consists of 9 di erent languages. A multilingual recognizer (MULTI) based on the four languages German, English, Japanese and Spanish was developed to serve as a source system. Likewise this system is very useful for language identi cation and achieves 100% language identi cation rate. Based on the MULTI system we evaluated our bootstrap technique on such completely di erent languages as Chinese, Croatian, and Turkish.