Investigating emotional speech parameters for speech synthesis

Investigating emotional speech parameters for speech synthesis
复制标题

研究语音合成的情感语音参数

DOI:
10.1109/icecs.1996.584653
复制
发表时间:
1996
期刊:
Proceedings of Third International Conference on Electronics, Circuits, and Systems
影响因子:
--
通讯作者:
G. Kokkinakis
G. Kokkinakis
中科院分区:
--
文献类型:
--
作者:
D. Galanis;V. Darsinos;G. Kokkinakis

文献摘要

被引文献

相似文献

本文研究了人类情感语音的一些特征,并报告了分析结果。收集的语音材料进行分析和分析过程中,相对于应用研究结果的共振峰为基础的文本到语音系统在我们的实验室开发的角度。更具体地说,以下参数进行了研究:音高轮廓,强度,语速和音素持续时间为四种不同的情绪:愤怒,恐惧,喜悦和悲伤。将情感语音数据的分析结果与情感中性人类语音的相应参数进行比较。从情绪中性参数值的偏离指示产生情绪合成语音所需的修改因子。
In this paper some of the characteristics of human emotional speech are investigated and the analysis results are reported. The collection of speech material to be analyzed and the analysis procedure were done with respect to the perspective of applying the findings to the formant based Text-to-Speech system developed in our laboratory. More specifically the following parameters have been investigated: pitch contour, intensity, speech rate and phoneme duration for four different emotions: anger, fear, joy and grief. The results of the analysis of the emotional speech data are compared to the corresponding parameters of emotionally neutral human speech. The diversion from the emotionally neutral parameters' values indicates the modification factor necessary for the production of emotional synthetic speech.