Design and collection of acoustic sound data for hands-free speech recognition and sound scene understanding

Design and collection of acoustic sound data for hands-free speech recognition and sound scene understanding
复制标题

设计和收集声学声音数据,用于免提语音识别和声音场景理解

DOI:
10.1109/icme.2002.1035537
复制
发表时间:
2002
期刊:
Proceedings. IEEE International Conference on Multimedia and Expo
影响因子:
--
通讯作者:
H. Saruwatari
H. Saruwatari
中科院分区:
--
文献类型:
--
作者:
Satoshi Nakamura;K. Hiyane;F. Asano;Y. Kaneda;Takeshi Yamada;T. Nishiura;Tetsunori Kobayashi;S. Ise;H. Saruwatari

文献摘要

被引文献

相似文献

在真实声环境中,声源定位、声音检索、声音识别、免提语音识别等研究都需要开放评价的声音数据。本文报道了我们的声学数据收集项目。真实环境中的声音场景有很多种。声音场景由声源和房间声学指定。在真实的声学环境中,声源、声源位置和房间的组合数量是巨大的。我们假设环境中的声音可以通过孤立声源和脉冲响应的卷积来模拟。作为一个孤立的声源,收集了上百种环境声和语音。在不同的声环境中采集脉冲响应。此外,我们还收集了来自移动源的声音。本文介绍了我们的声音场景数据库采集项目及其在环境声音识别和免提语音识别中的应用进展。
The sound data for open evaluation is necessary for studies such as sound source localization, sound retrieval, sound recognition and hands-free speech recognition in real acoustic environments. This paper reports on our project for acoustic data collection. There are many kinds of sound scenes in real environments. The sound scene is specified by sound sources and room acoustics. The number of combinations of the sound sources, source positions and rooms is huge in real acoustic environments. We assumed that the sound in the environments can be simulated by convolution of the isolated sound sources and impulse responses. As an isolated sound source, hundred kinds of environment sounds and speech sounds are collected. The impulse responses are collected in various acoustic environments. Additionally we collected sounds from a moving source. In this paper, progress of our sound scene database collection project and application to environment sound recognition and hands-free speech recognition are described.