The analysis of Acoustic Phonetic Data: exploring differences in the spoken Romance languages

The analysis of Acoustic Phonetic Data: exploring differences in the spoken Romance languages
复制标题

声学语音数据分析:探索罗曼语口语的差异

DOI:
--
复制
发表时间:
2015
期刊:
影响因子:
--
通讯作者:
J. Aston
J. Aston
中科院分区:
--
文献类型:
--
作者:
D. Pigoli;P. Hadjipantelis;J. Coleman;J. Aston

文献摘要

参考文献

被引文献

相似文献

长期以来,人们一直从文本变化和语音转换的角度研究从古代语言到现代语言的变化过程,特别是理解历史和地理传播。然而,从声学的角度来分析这些是有点困难的,虽然这可能是主要的传播方法,而不是通过书面记录。在这里,我们提出了一种新的方法来分析的声学语音数据,其目的将是统计语音的声音模型。特别是,我们探索语音的变化和变化使用的时间-频率表示,即语音录音的对数频谱图。在对数据进行预处理以去除固有的个体差异后,我们将时间和频率协方差函数确定为语言的一个特征;相比之下,平均值主要取决于已经说出的特定单词。我们为均值和协方差建立模型(考虑到对此类对象的统计分析的限制),并使用它来定义语音转换,使我们能够对单个说话者在不同语言中的发音进行建模,从而探索语言之间的语音差异。最后,我们将这些转换映射回录音的域,使我们能够收听统计分析。所提出的方法证明使用录音的话对应的数字从“一”到“十”的发音的发言者从五个不同的罗曼语。
The process of change, particularly understanding the historical and geographical spread, from older to modern languages has long been studied from the point of view of textual changes and phonetic transcriptions. However, it is somewhat more difficult to analyze these from an acoustic point of view, although this is likely to be the dominant method of transmission rather than through written records. Here, we propose a novel approach to the analysis of acoustic phonetic data, where the aim will be to model statistically speech sounds. In particular, we explore phonetic variation and change using a time-frequency representation, namely the log-spectrograms of speech recordings. After preprocessing the data to remove inherent individual differences, we identify time and frequency covariance functions as a feature of the language; in contrast, the mean depends mostly on the particular word that has been uttered. We build models for the mean and covariances (taking into account the restrictions placed on the statistical analysis of such objects) and use this to define a phonetic transformation that allows us to model how an individual speaker would sound in a different language, allowing the exploration of phonetic differences between languages. Finally, we map back these transformations to the domain of sound recordings, allowing us to listen to statistical analysis. The proposed approach is demonstrated using the recordings of the words corresponding to the numbers from "one" to "ten" as pronounced by speakers from five different Romance languages.
DOI: 10.1121/1.4714345
发表时间: 2012
期刊: The Journal of the Acoustical Society of America
影响因子: --
作者:
Hadjipantelis PZ
通讯作者: Hadjipantelis PZ