Spectral- and Cepstral-Based Measures During Continuous Speech: Capacity to Distinguish Dysphonia and Consistency Within a Speaker

Spectral- and Cepstral-Based Measures During Continuous Speech: Capacity to Distinguish Dysphonia and Consistency Within a Speaker
复制标题

DOI:
10.1016/j.jvoice.2010.06.007
复制
发表时间:
2011-09-01
期刊:
影响因子:
2.2
通讯作者:
Hahn, Youngmee C.
Hahn, Youngmee C.
中科院分区:
医学3区
文献类型:
--
作者:
Lowell, Soren Y.;Colton, Raymond H.;Hahn, Youngmee C.

文献摘要

被引文献

相似文献

基于频谱和倒谱的声学测量优于基于时间的测量,用于在连续语音期间准确地表示发音困难的声音。虽然这些措施表现出有前途的关系,感知语音质量评级,少有人知道他们的能力,以区分正常的发音困难的声音在连续讲话和这些措施的一致性,在多个话语由同一发言人。本研究的目的是确定是否频谱矩的长期平均频谱(LTAS)(频谱平均值,标准差,偏度和峰度)和倒谱峰值突出措施显着不同的扬声器与无语音障碍时,在连续讲话。这些措施的一致性,在一个发言者在不同的话语也得到了解决。对27例无嗓音障碍和27例混合性嗓音障碍的受试者的连续语音样本进行声学分析。此外,声音样本的整体严重程度进行了感知评级。声学分析进行了三个连续的语音刺激从阅读通过:两个完整的句子和一个组成短语。组间差异显著的两个倒频谱测量和三个LTAS测量(P < 0.001):频谱平均值,偏度和峰度。这五项措施也显示出中度到强烈的相关性,整体声音的严重性。此外,在不同长度和音素内容的话语中,两个受试者组都表现出高度的说话者内一致性(相关系数>= 0.89)。
Spectral- and cepstral-based acoustic measures are preferable to time-based measures for accurately representing dysphonic voices during continuous speech. Although these measures show promising relationships to perceptual voice quality ratings, less is known regarding their ability to differentiate normal from dysphonic voice during continuous speech and the consistency of these measures across multiple utterances by the same speaker. The purpose of this study was to determine whether spectral moments of the long-term average spectrum (LTAS) (spectral mean, standard deviation, skewness, and kurtosis) and cepstral peak prominence measures were significantly different for speakers with and without voice disorders when assessed during continuous speech. The consistency of these measures within a speaker across utterances was also addressed. Continuous speech samples from 27 subjects without voice disorders and 27 subjects with mixed voice disorders were acoustically analyzed. In addition, voice samples were perceptually rated for overall severity. Acoustic analyses were performed on three continuous speech stimuli from a reading passage: two full sentences and one constituent phrase. Significant between-group differences were found for both cepstral measures and three LTAS measures (P < 0.001): spectral mean, skewness, and kurtosis. These five measures also showed moderate to strong correlations to overall voice severity. Furthermore, high degrees of within-speaker consistency (correlation coefficients >= 0.89) across utterances with varying length and phonemic content were evidenced for both subject groups.