Specifying Affect and Emotion for Expressive Speech Synthesis

Specifying Affect and Emotion for Expressive Speech Synthesis
复制标题

指定表达性语音合成的情感和情感

DOI:
10.1007/978-3-540-24630-5_47
复制
发表时间:
2004
期刊:
Conference on Intelligent Text Processing and Computational Linguistics
影响因子:
--
通讯作者:
N. Campbell
N. Campbell
中科院分区:
--
文献类型:
--
作者:
N. Campbell

文献摘要

被引文献

相似文献

语音合成不一定是文本到语音的同义词。本文描述了一种原型说话机,它使用基于图标的输入,从说话者、语言、说话风格和内容信息的组合中产生合成语音。本文结合语言和说话人的信息,从概念符号的组合出发,讨论了会话话语的文本内容和输出实现的具体问题。本文的结论是,为了充分指定演讲内容(即文本细节和演讲风格),需要选择演讲者-承诺和演讲者-听众关系的选项。本文最后描述了一种基于约束的方法,用于选择用于连接语音合成的情感标记语音样本。
Speech synthesis is not necessarily synonymous with text-to-speech. This paper describes a prototype talking machine that produces synthesised speech from a combination of speaker, language, speaking-style, and content information, using icon-based input. The paper addresses the problems of specifying the text-content and output realisation of a conversational utterance from a combination of conceptual icons, in conjunction with language and speaker information. It concludes that in order to specify the speech content (i.e., both text details and speaking-style) adequately, selection options for speaker-commitment and speaker-listener relations will be required. The paper closes with a description of a constraint-based method for selection of affect-marked speech samples for concatenative speech synthesis.