Speaker localization among multi-faces in noisy environment by audio-visual integration
Speaker localization among multi-faces in noisy environment by audio-visual integration
复制标题
通过视听集成在嘈杂环境中进行多面说话者定位
DOI:
10.1109/robot.2006.1641889
复制
发表时间:
2006
期刊:
影响因子:
--
通讯作者:
Munsang Kim
中科院分区:
文献类型:
--
作者:
Hyun;Jong;Munsang Kim
In this paper, we not only developed a reliable sound localization system including VAD (voice activity detection) component using three microphones but also a face tracking system using a vision camera. Moreover, we proposed a way to integrate these systems in the human-robot interaction to compensate the errors in the localization of a speaker and to reject unnecessary speech or noise signals entering from the undesired directions effectively. For the purpose of verifying our system's performances, we installed the proposed audition and vision system to the prototype robot, called IROBAA (Intelligent ROBot for Active Audition), and showed how to integrate an audio-visual system
DOI:
10.1109/jra.1987.1087109
发表时间:
1987-08-01
期刊:
IEEE JOURNAL OF ROBOTICS AND AUTOMATION
影响因子:
--
作者:
TSAI, RY
通讯作者:
TSAI, RY