The LaMIT database: A read speech corpus for acoustic studies of the Italian language toward lexical access based on the detection of landmarks and other acoustic cues to features.

The LaMIT database: A read speech corpus for acoustic studies of the Italian language toward lexical access based on the detection of landmarks and other acoustic cues to features.
复制标题

DOI:
10.1016/j.dib.2022.108275
复制
发表时间:
2022-06
期刊:
影响因子:
1.2
通讯作者:
Budoni, Sara
Budoni, Sara
中科院分区:
其他
文献类型:
--
作者:
Di Benedetto, Maria-Gabriella;Shattuck-Hufnagel, Stefanie;Choi, Jeung-Yoon;De Nardis, Luca;Arango, Javier;Chan, Ian;DeCaprio, Alec;Budoni, Sara

文献摘要

参考文献

相似文献

LaMIT数据库包含100个意大利语句子的录音。数据库中的句子被设计为包括意大利语的所有音素,并且还考虑到书面意大利语中每个音素的典型频率。四名在意大利罗马长大并生活的标准意大利语母语成年人,两名女性和两名男性,在两个不同的录音会话中发音句子;因此收集了每个说话者每个句子的两次重复,总共800次录音。该数据库是专门为LaMIT项目的应用而创建的,该项目的重点是Ken Stevens为美国英语提出的词汇访问模型在意大利语中的应用。该模型依赖于检测特定的声学不连续性,称为地标和其他声学线索的功能,表征每个音素。因此,处理每个记录以生成一组标记文件,其识别预测的地标和其他线索以及实际的地标/线索。根据Praat语音处理软件中使用的标注语法编译的标注文件也可作为LAMIT数据库的一部分提供。
The LaMIT database consists in recordings of 100 Italian sentences. The sentences in the database were designed so to include all phonemes of the Italian language, and also take into account the typical frequency of each phoneme in written Italian. Four native adult speakers of Standard Italian, raised and living in Rome, Italy, two female and two male, pronounced the sentences in two different recording sessions; two repetitions for each sentence per speaker were therefore collected, for a total of 800 recordings. The database was specifically created for application in the LaMIT project, that focuses on the application to the Italian language of the Lexical Access model proposed by Ken Stevens for American English. The model relies on the detection of specific acoustic discontinuities called landmarks and other acoustic cues to features that characterize each phoneme. Each recording was thus processed to generate a set of labeling files that identify both predicted landmarks and other cues, and actual landmarks/cues. The labeling files, compiled according to the labeling syntax used in the Praat speech processing software, are also made available as part of the LAMIT database.
DOI: 10.1121/1.1458026
发表时间: 2002-04-01
影响因子: 2.4
作者:
Stevens, KN
通讯作者: Stevens, KN
DOI: 10.1121/10.0004987
发表时间: 2021-05-01
影响因子: 2.4
作者:
Di Benedetto, Maria-Gabriella;Shattuck-Hufnagel, Stefanie;DeCaprio, Alec
通讯作者: DeCaprio, Alec