Chhattisgarhi speech corpus for research and development in automatic speech recognition

Chhattisgarhi speech corpus for research and development in automatic speech recognition
复制标题

用于自动语音识别研究和开发的恰蒂斯加尔语语音语料库

DOI:
10.1007/s10772-018-9496-7
复制
发表时间:
2018
影响因子:
--
通讯作者:
G. B. Kshirsagar
G. B. Kshirsagar
中科院分区:
--
文献类型:
--
作者:
N. Londhe;G. B. Kshirsagar

文献摘要

被引文献

相似文献

自动语音识别(ASR)是一种计算机化的接口,它允许人类以自然对话的方式与机器进行交流。ASR在各个领域有着广泛的应用,例如幼儿的语言发展、电信、作为听力受损者的辅助设备等。ASR系统的性能受用于其实现的数据库的影响很大。本文讨论的是一种罕见但重要的印度方言切蒂斯加里语的语音语料库的建立。该语音语料库由100个独立的单词和四个语音脚本组成,共67个句子,来自478名母语者。这些词选自恰蒂斯加尔Rajbhasha Aayog出版的英语到恰蒂斯加尔邦词典以及恰蒂斯加尔邦文学和报纸文章的脚本。该数据集是在恰蒂斯加尔州60%的地理区域内收集的。最后,首次为恰蒂斯加尔邦准备了一个有价值的语音语料库,旨在加强语音研究。在准备好的数据库上,已经证明了对孤立和连续语音样本的语音识别的成功消除。
Automatic speech recognition (ASR) is a computerized interface which allows humans to communicate with machine in a way of its natural conversation. ASR has wide range of applications in various fields such as language development in young children, telecommunications, as an assistive device for hearing impaired etc. Performance of ASR system is greatly influenced by the database used for its implementation. In this paper, we are discussing about building a speech corpus for a rare but important Indian dialect Chhattisgarhi. This speech corpus consists of 100 unique isolated words and four speech scripts aggregating 67 sentences, recorded from total 478 native speakers. These words were selected from English to Chhattisgarhi dictionary published by Chhattisgarh Rajbhasha Aayog and scripts from Chhattisgarhi literature and newspaper articles. This dataset has been collected travelling over 60% geographical area of the Chhattisgarh state. Finally, a valuable speech corpus for the first time have been prepared for Chhattisgarhi with an aim to enhance the speech research. The successful extermination of speech recognition for both isolated and continuous speech samples have been demonstrated on the prepared database.