The Statistical Analysis of Acoustic Phonetic Data: Exploring Differences Between Spoken Romance Languages

The Statistical Analysis of Acoustic Phonetic Data: Exploring Differences Between Spoken Romance Languages
复制标题

声学语音数据的统计分析:探索罗曼语口语之间的差异

DOI:
10.1111/rssc.12258
复制
发表时间:
2018
期刊:
Applied Statistics
影响因子:
--
通讯作者:
Pigoli D
Pigoli D
中科院分区:
--
文献类型:
--
作者:
Pigoli D

文献摘要

相似文献

长期以来,人们一直通过考察文本变化和音标变化来研究从古老语言到现代语言的历史和地理传播。然而,从声学的角度分析语言变化更为困难,尽管这通常是主要的传播方式。我们提出了一种新的声学语音数据分析方法,其目的是对口语的声学特性进行统计建模。我们通过使用时频表示,即语音记录的对数谱图来探索语音的变化和变化。我们将时间和频率协方差函数识别为语言的特征;相比之下,平均谱图主要取决于所发出的特定单词。我们为平均值和协方差建立了模型(考虑到对这些对象的统计分析的限制),并使用这些模型来定义语音转换,该转换可以模拟单个说话者在不同语言中的发音,从而允许探索语言之间的语音差异。最后,我们将这些转换映射回录音领域,使我们能够收听统计分析的输出。该方法通过使用来自五种不同罗曼语的说话者发音的1到10对应单词的录音来证明。
The historical and geographical spread from older to more modern languages has long been studied by examining textual changes and in terms of changes in phonetic transcriptions. However, it is more difficult to analyse language change from an acoustic point of view, although this is usually the dominant mode of transmission. We propose a novel analysis approach for acoustic phonetic data, where the aim will be to model the acoustic properties of spoken words statistically. We explore phonetic variation and change by using a time–frequency representation, namely the log-spectrograms of speech recordings. We identify time and frequency covariance functions as a feature of the language; in contrast, mean spectrograms depend mostly on the particular word that has been uttered. We build models for the mean and covariances (taking into account the restrictions placed on the statistical analysis of such objects) and use these to define a phonetic transformation that models how an individual speaker would sound in a different language, allowing the exploration of phonetic differences between languages. Finally, we map back these transformations to the domain of sound recordings, enabling us to listen to the output of the statistical analysis. The approach proposed is demonstrated by using recordings of the words corresponding to the numbers from 1 to 10 as pronounced by speakers from five different Romance languages.