Onoma-to-wave: Environmental sound synthesis from onomatopoeic words

Onoma-to-wave: Environmental sound synthesis from onomatopoeic words
复制标题

DOI:
10.1561/116.00000049
复制
发表时间:
2021-02
期刊:
ArXiv
影响因子:
--
通讯作者:
Yuki Okamoto;Keisuke Imoto;Shinnosuke Takamichi;Ryosuke Yamanishi;Takahiro Fukumori;Y. Yamashita
Yuki Okamoto;Keisuke Imoto;Shinnosuke Takamichi;Ryosuke Yamanishi;Takahiro Fukumori;Y. Yamashita
中科院分区:
其他
文献类型:
--
作者:
Yuki Okamoto;Keisuke Imoto;Shinnosuke Takamichi;Ryosuke Yamanishi;Takahiro Fukumori;Y. Yamashita

文献摘要

相似文献

在本文中,我们提出了一个框架,从拟声词的环境声音合成。作为表达环境声音的一种方式,我们可以使用拟声词,这是一个语音模仿声音的字符序列。拟声词能有效地描述各种语音特征。因此,使用拟声词进行环境声音合成将使我们能够生成多样化的环境声音。为了生成不同的声音,我们提出了一种方法的基础上的序列到序列的框架合成环境的声音拟声词。我们还提出了一种使用拟声词和声音事件标签的环境声音合成方法。除了拟声词之外,声音事件标签的使用使我们能够根据输入的声音事件标签来捕获每个声音事件的特征。我们的主观实验表明,我们提出的方法实现了更高的多样性和自然度比传统的方法使用声音事件标签。
In this paper, we propose a framework for environmental sound synthesis from onomatopoeic words. As one way of expressing an environmental sound, we can use an onomatopoeic word, which is a character sequence for phonetically imitating a sound. An onomatopoeic word is effective for describing diverse sound features. Therefore, using onomatopoeic words for environmental sound synthesis will enable us to generate diverse environmental sounds. To generate diverse sounds, we propose a method based on a sequence-to-sequence framework for synthesizing environmental sounds from onomatopoeic words. We also propose a method of environmental sound synthesis using onomatopoeic words and sound event labels. The use of sound event labels in addition to onomatopoeic words enables us to capture each sound event's feature depending on the input sound event label. Our subjective experiments show that our proposed methods achieve higher diversity and naturalness than conventional methods using sound event labels.