Speaker localization among multi-faces in noisy environment by audio-visual integration

Speaker localization among multi-faces in noisy environment by audio-visual integration
复制标题

通过视听集成在嘈杂环境中进行多面说话者定位

DOI:
10.1109/robot.2006.1641889
复制
发表时间:
2006
期刊:
Proceedings 2006 IEEE International Conference on Robotics and Automation, 2006. ICRA 2006.
影响因子:
--
通讯作者:
Munsang Kim
Munsang Kim
中科院分区:
--
文献类型:
--
作者:
Hyun;Jong;Munsang Kim

文献摘要

参考文献

被引文献

相似文献

在本文中,我们不仅开发了一个可靠的声音定位系统,包括VAD(语音活动检测)组件使用三个麦克风,但也是一个人脸跟踪系统使用视觉摄像头。此外,我们提出了一种方法,将这些系统集成在人机交互中,以补偿扬声器定位中的误差,并有效地拒绝不必要的语音或噪声信号从不希望的方向进入。为了验证我们的系统的性能,我们安装了建议的听觉和视觉系统的原型机器人,称为IROBAA(智能机器人主动Audition),并展示了如何整合视听系统
In this paper, we not only developed a reliable sound localization system including VAD (voice activity detection) component using three microphones but also a face tracking system using a vision camera. Moreover, we proposed a way to integrate these systems in the human-robot interaction to compensate the errors in the localization of a speaker and to reject unnecessary speech or noise signals entering from the undesired directions effectively. For the purpose of verifying our system's performances, we installed the proposed audition and vision system to the prototype robot, called IROBAA (Intelligent ROBot for Active Audition), and showed how to integrate an audio-visual system
DOI: 10.1109/jra.1987.1087109
发表时间: 1987-08-01
期刊: IEEE JOURNAL OF ROBOTICS AND AUTOMATION
影响因子: --
作者:
TSAI, RY
通讯作者: TSAI, RY