A sensor-fusion method for detecting a speaking student

A sensor-fusion method for detecting a speaking student
复制标题

一种检测说话学生的传感器融合方法

DOI:
10.1109/icme.2003.1220871
复制
发表时间:
2003
期刊:
2003 International Conference on Multimedia and Expo. ICME '03. Proceedings (Cat. No.03TH8698)
影响因子:
--
通讯作者:
M. Minoh
M. Minoh
中科院分区:
--
文献类型:
--
作者:
Satoshi Nishiguchi;Kazuhide Higashi;Y. Kameda;M. Minoh

文献摘要

被引文献

相似文献

在本文中,我们提出了一种检测演讲者位置的方法,演讲者是远程学习和讲座档案中自动视频拍摄的目标。要求在讲课视频中拍摄讲话学生的脸。为此,有必要检测扬声器的位置。诸如麦克风阵列之类的声传感器被广泛用于探测声源的位置。但是,在演讲厅等大空间中,由于噪声的存在,仅使用麦克风阵列很难准确地探测声源的位置。在本文中,我们提出了一种利用麦克风阵列和视觉传感器更精确地检测演讲室内演讲者位置的方法。结果表明,该方法可将说话人位置检测的精度提高20%左右。
In this paper, we propose a method for detecting the location of the speaker that is a target of automatic video filming in distance learning and lecture archive. It is required that a face of a speaking student is filmed in a lecture video. For this purpose, it is necessary to detect the location of a speaker. An acoustic sensor such as a microphone array is used widely to detect the location of a sound source. However, it is difficult to detect the location of a sound source precisely using only microphone array because of sound noise in a large space such as a lecture room. In this paper, we propose a method for detecting more precise location of a speaker in the lecture room using not only the microphone array but also visual sensors. The result shows that the precision ratio of detecting the location of a speaker was improved about 20% by our sensor-fusion method.