Automatic Speech Recognition System Dedicated for Polish

Automatic Speech Recognition System Dedicated for Polish
复制标题

波兰语专用自动语音识别系统

DOI:
--
复制
发表时间:
2011
期刊:
Interspeech
影响因子:
--
通讯作者:
Mariusz Masior
Mariusz Masior
中科院分区:
--
文献类型:
--
作者:
M. Ziólko;Jakub Gałka;B. Ziółko;T. Jadczyk;D. Skurzok;Mariusz Masior

文献摘要

被引文献

相似文献

展示了一种自动语音识别系统。由于波兰语和英语之间的差异,我们系统的几层与流行方法不同。几十年前,关于自动语音识别(ASR)的研究始于几十年前。该领域的大多数进度都是用英语完成的。这导致了许多成功的设计,但是ASR系统始终低于人类语音识别能力的水平,即使对于英语也是如此。如果不那么受欢迎的语言,例如波兰语(大约有6000万扬声器),情况更糟。没有用于波兰的大型词汇ASR(LVR)软件。波兰语音包含非常高频的手机(摩擦和普罗逊语),该语言高度发音和非刻词。 PrimePeech开发的一些商业呼叫中心应用程序,但仅限于其域名。我们的系统基于修改的KNN分类器和小波。它针对波兰,而其他[1、6、7、5]更笼统,并且强烈基于HTK框架[9]。
An automatic speech recognition system for Polish is demonstrated. A few layers of our system are different from popular approaches as a result of differences between Polish and English languages. Research on automatic speech recognition (ASR) started several decades ago. Most of the progress in the field was done for English. It has resulted in many successful designs, however ASR systems are always below the level of human speech recognition capability, even for English. In case of less popular languages, like Polish (with around 60 million speakers), the situation is much worse. There is no large vocabulary ASR (LVR) software for Polish. Polish speech contains very high-frequency phones (fricatives and plosives) and the language is highly inflected and non-positional. There are some commercial call centre applications, developed by PrimeSpeech, but they are limited to their domain areas. Our system is based on modified kNN classifier and wavelets. It is targeted for Polish, while others [1, 6, 7, 5] are more general, and strongly based on HTK framework [9].