Integration of Multiple Sound Source Localization Results for Speaker Identification in Multiparty Dialogue System
Integration of Multiple Sound Source Localization Results for Speaker Identification in Multiparty Dialogue System
复制标题
集成多声源定位结果进行多方对话系统中说话人识别
DOI:
10.1007/978-1-4614-8280-2_14
复制
发表时间:
2012
期刊:
影响因子:
--
通讯作者:
Satoshi Sato
中科院分区:
文献类型:
--
作者:
Taichi Nakashima;Kazunori Komatani;Satoshi Sato
Humanoid robots need to head toward human participants when answering to their questions in multiparty dialogues. Some positions of participants are difficult to localize from robots in multiparty situations, especially when the robots can only use their own sensors. We present a method for identifying the speaker more accurately by integrating the multiple sound source localization results obtained from two robots: one talking mainly with participants and the other also joining the conversation when necessary. We place them so that they can compensate for each other’s localization capabilities and then integrate their two results. Our experimental evaluation revealed that using two robots improved speaker identification compared with using only one robot. We furthermore implemented our method into humanoid robots and constructed a demo system.