Hybridizing conversational and clear speech to determine the degree of contribution of acoustic features to intelligibility

Hybridizing conversational and clear speech to determine the degree of contribution of acoustic features to intelligibility
复制标题

DOI:
10.1121/1.2967844
复制
发表时间:
2008-10-01
影响因子:
2.4
通讯作者:
Hosom, John-Paul
Hosom, John-Paul
中科院分区:
物理与天体物理3区
文献类型:
--
作者:
Kain, Alexander;Amano-Kusumoto, Akiko;Hosom, John-Paul

文献摘要

被引文献

相似文献

说话者自然会采用一种特殊的“清晰”(clear)说话风格,以便更好地被由于听力障碍、背景噪音的存在或两者兼而有之而在理解语音能力方面受到中度损害的听众所理解。相反,在安静的环境中为非受损的听众准备的语音被称为“会话”(CNV)。研究表明,在不利的环境下,非自然语音的可懂度通常高于CNV语音。目前还不知道哪些单独的声学特征或特征的组合导致了更高的可懂度的语音。本研究的目的是确定单个说话人的某些声学特征对可懂度的贡献。所提出的方法创建“混合”(HYB)的语音刺激,选择性地结合联合收割机的声学特征的一个句子中说的CNV和cardiac风格。这些刺激的可理解性,然后测量知觉测试,使用96个语音平衡的句子。一个扬声器的结果显示显着的可理解性改善CNV语音时,更换某些组合的短期频谱,音素身份,和音素持续时间的CNV语音与那些从语音语音的组合,但没有改善涉及基频,能量,或非语音事件(停顿)。(c)2008年,美国声学学会。
Speakers naturally adopt a special "clear" (CLR) speaking style in order to be better understood by listeners who are moderately impaired in their ability to understand speech due to a hearing impairment, the presence of background noise, or both. In contrast, speech intended for nonimpaired listeners in quiet environments is referred to as "conversational" (CNV). Studies have shown that the intelligibility of CLR speech is usually higher than that of CNV speech in adverse circumstances. It is not known which individual acoustic features or combinations of features cause the higher intelligibility of CLR speech. The objective of this study is to determine the contribution of some acoustic features to intelligibility for a single speaker. The proposed method creates "hybrid" (HYB) speech stimuli that selectively combine acoustic features of one sentence spoken in the CNV and CLR styles. The intelligibility of these stimuli is then measured in perceptual tests, using 96 phonetically balanced sentences. Results for one speaker show significant sentence-level intelligibility improvements over CNV speech when replacing certain combinations of short-term spectra, phoneme identities, and phoneme durations of CNV speech with those from CLR speech, but no improvements for combinations involving fundamental frequency, energy, or nonspeech events (pauses). (c) 2008 Acoustical Society of America.