Acoustic-phonetic labels in a Japanese speech database

Acoustic-phonetic labels in a Japanese speech database
复制标题

日语语音数据库中的声学语音标签

DOI:
10.21437/ecst.1987-152
复制
发表时间:
1987
期刊:
--
影响因子:
--
通讯作者:
S. Katagiri
S. Katagiri
中科院分区:
--
文献类型:
--
作者:
K. Takeda;Y. Sagisaka;S. Katagiri

文献摘要

被引文献

相似文献

介绍了一个大型日语语音数据库JSDB-ATR。这些语音数据以多种方式使用声学语音符号进行转录,用于各种数据访问请求和便于精细声学语音分析。对于多重转录,考虑了三种类型的类别:语言和音位类别,声学事件类别和一些单音变化类别。到目前为止,已收集了约8500字,分别由八个专业播音员说,其中一半是声学语音转录。引言最近,语音数据库的建设已经在许多语言中进行,以获得语音识别,感知和合成的许多知识[1]-[3]。本文介绍了ATR正在建立的一个大型日语语音数据库(JSDB-ATR),重点介绍了它的多声音转换特性。
A large sized Japanese speech database at ATR(JSDB-ATR) is introduced. Thesespeech data are transcribed in multiple ways using acoustic-phonetic symbols for various data access requests and for the convenience of fine acoustic-phonetic analysis. For multiple transcription, three types of categories are considered: linguistic and phonemic categories, acoustic event categories and some alophonic variation categories. To date, about 8500 words respectively uttered by eight professional announcers have been collected with half of them being acoustically-phonetically transcribed. INTRODUCTION Recently, the construction of speech databases has been undertaken in many languages to obtain much knowledge for speech recognition, perception and syntheses[l]-[3]. However, there are few Japanese speech database (JSDB) large enough for various research purposes.ln this paper, a large Japanese speech database that is being built at ATR (JSDB-ATR) is introduced focusing on its multiple acoustic-phonetic transcriptions.