Human-robot interaction through real-time auditory and visual multiple-talker tracking
Human-robot interaction through real-time auditory and visual multiple-talker tracking
复制标题
通过实时听觉和视觉多说话者跟踪进行人机交互
DOI:
10.1109/iros.2001.977177
复制
发表时间:
2001
期刊:
影响因子:
--
通讯作者:
H. Kitano
中科院分区:
文献类型:
--
作者:
HIroshi G. Okuno;K. Nakadai;K. Hidai;H. Mizoguchi;H. Kitano
Nakadai et al. (2001) have developed a real-time auditory and visual multiple-talker tracking technique. In this paper, this technique is applied to human-robot interaction including a receptionist robot and a companion robot at a party. The system includes face identification, speech recognition, focus-of-attention control, and sensorimotor task in tracking multiple talkers. The system is implemented on a upper-torso humanoid and the talker tracking is attained by distributed processing on three nodes connected by 100Base-TX network. The delay of tracking is 200 msec. Focus-of-attention is controlled by associating auditory and visual streams by using the sound source direction and talker position as a clue. Once an association is established, the humanoid keeps its face to the direction of the associated talker.