Psychoacoustic cues to emotion in speech prosody and music

Psychoacoustic cues to emotion in speech prosody and music
复制标题

DOI:
10.1080/02699931.2012.732559
复制
发表时间:
2013-06-01
影响因子:
2.6
通讯作者:
Dibben, Nicola
Dibben, Nicola
中科院分区:
心理学3区
文献类型:
--
作者:
Coutinho, Eduardo;Dibben, Nicola

文献摘要

被引文献

相似文献

有强有力的证据表明,在音乐和演讲中表达情感时共有的声学特征,但对所涉及的特定心理声学特征的理解相对有限。本研究结合控制实验和计算模型来研究听觉域中与情感表达相关的感知代码。研究的实证阶段提供了连续的人类情感评级,在电影音乐和自然语音样本的摘录。计算阶段创建了一个计算机模型,该模型从声学刺激中检索相关信息,并对语音和音乐的情感表达进行预测,以接近人类受试者的反应。我们发现,一个显着的一部分,听众的第二次报告的情绪,音乐和语音韵律可以预测从一组七个心理声学特征:响度,克里思/语音速率,旋律/韵律轮廓,频谱质心,频谱通量,锐度和粗糙度。这些结果的影响进行了讨论的背景下,跨模态的相似性在声学域的情感沟通。
There is strong evidence of shared acoustic profiles common to the expression of emotions in music and speech, yet relatively limited understanding of the specific psychoacoustic features involved. This study combined a controlled experiment and computational modelling to investigate the perceptual codes associated with the expression of emotion in the acoustic domain. The empirical stage of the study provided continuous human ratings of emotions perceived in excerpts of film music and natural speech samples. The computational stage created a computer model that retrieves the relevant information from the acoustic stimuli and makes predictions about the emotional expressiveness of speech and music close to the responses of human subjects. We show that a significant part of the listeners' second-by-second reported emotions to music and speech prosody can be predicted from a set of seven psychoacoustic features: loudness, tempo/speech rate, melody/prosody contour, spectral centroid, spectral flux, sharpness, and roughness. The implications of these results are discussed in the context of cross-modal similarities in the communication of emotion in the acoustic domain.