Building a naturalistic emotional speech corpus by retrieving expressive behaviors from existing speech corpora
Building a naturalistic emotional speech corpus by retrieving expressive behaviors from existing speech corpora
复制标题
通过从现有语音语料库中检索表达行为来构建自然情感语音语料库
DOI:
10.21437/interspeech.2014-60
复制
发表时间:
2014
期刊:
影响因子:
--
通讯作者:
C. Busso
中科院分区:
文献类型:
--
作者:
Soroosh Mariooryad;Reza Lotfian;C. Busso
A key element in affective computing is to have large corpora of genuine emotional samples collected during natural conversations. Recording natural interactions through telephone is an appealing approach to build emotional databases. However, collecting real conversational data with expressive reactions is a challenging task, especially if the recordings are to be shared with the community (e.g., privacy concerns). This study explores a novel approach consisting in retrieving emotional reactions from existing spontaneous speech databases collected for general speech processing problems. Although most of the recordings in these databases are expected to have non-emotional expressions, given the naturalness of the interactions, the flow of the conversation can lead to emotional responses from conversation partners which we aim to retrieve. We use the IEMOCAP and SEMAINE databases to build emotion detector systems. We use these classifiers to identify emotional behaviors from the FISHER database, which is a large conversational speech corpus recorded over the phone. Subjective evaluations over the retrieved samples demonstrate the potential of the proposed scheme to build naturalistic emotional speech database. Index Terms: emotion recognition, expressive speech, information retrieval, emotional databases