Structural representation of the non-native pronunciations
Structural representation of the non-native pronunciations
复制标题
DOI:
10.21437/interspeech.2005-94
复制
发表时间:
2005
期刊:
影响因子:
--
通讯作者:
S. Asakawa;N. Minematsu;Toshiko Isei-Jaakkola;K. Hirose
中科院分区:
文献类型:
--
作者:
S. Asakawa;N. Minematsu;Toshiko Isei-Jaakkola;K. Hirose
Acoustic representation of speech provided by phonetics, spectrogram, is noisy representation in that it shows every acoustic aspect of speech. Age, gender, shape, microphone, room, line, etc. are completely irrelevant to the pronunciation assessment. However, the spectrogram is affected inevitably by these factors. Recently, a novel acoustic representation of speech was proposed, where dimensions of these non-linguistic factors can hardly be seen[1, 2]. Every acoustic substance of speech is discarded and only their interrelations are extracted to represent the pronunciation structurally. Using this method, individual learnersweredescribedasdistortedphonemicstructures[3]andautomatic scoring of the pronunciation was investigated[3, 4]. This paper describes two new analyses using the proposed method. The first analysis is done to examine whether the method can trace the development of a student’s pronunciation appropriately using only a limited amount of speech. The second one focuses on the prosodic aspect of the pronunciation, especially stressed and unstressed vowels. The former indicates that the proposed method can show history of the student’s development adequately and the latter clarifies that size of the pronunciation structure is highly correlated with the pronunciation proficiency.