Man-Machine Interaction Using a Vision System with Dual Viewing Angles

Man-Machine Interaction Using a Vision System with Dual Viewing Angles
复制标题

使用双视角视觉系统的人机交互

DOI:
--
复制
发表时间:
1997
影响因子:
0.7
通讯作者:
Mitsurti Ishizuka
Mitsurti Ishizuka
中科院分区:
计算机科学4区
文献类型:
--
作者:
Ying;H. Dohi;Mitsurti Ishizuka

文献摘要

被引文献

相似文献

本文介绍了一种双视角视觉系统,宽视角和窄视角,以及基于视觉系统的人性化语音对话环境方案。宽视角为大范围运动跟踪提供了宽视场,窄视角能够在宽视场中跟踪目标,以足够的分辨率拍摄目标的图像。为了实现快速、鲁棒的运动跟踪,定义了修正运动能量(MME)和存在能量(EE),在检测目标运动的同时提取运动区域。代替使用诸如在语音对话系统中通常使用的脚踏开关的物理设备,在我们的系统中从用户的嘴的运动检测话语的开始/结束。在不直接识别嘴唇的运动的情况下,跟踪嘴唇之间的区域的形状变化,以更稳定地识别对话的跨度。当不执行识别时,跟踪速度约为10帧/秒,当不使用任何特殊硬件执行跟踪和识别时,跟踪速度约为5帧/秒。关键词:视觉系统,双视角,语音对话系统,运动跟踪,口型识别
This paper describes a vision system with dual viewing angles, i.e., wide and narrow viewing angles, and a scheme of user-friendly speech dialogue environment based on the vision system. The wide viewing angle provides a wide viewing field for wide range motion tracking, and the narrow viewing angle is capable of following a target in wide viewing field to take the image of the target with sufficient resolution. For a fast and robust motion tracking, modified motion energy (MME) and existence energy (EE) are defined to detect the motion of the target and extract the motion region at the same time. Instead of using a physical device such as a foot switch commonly used in speech dialogue systems, the begin/end of an utterance is detected from the movement of user’s mouth in our system. Without recognizing the movement of lips directly, the shape variation of the region between lips is tracked for more stable recognition of the span of a dialogue. The tracking speed is about 10 frames/sec when no recognition is performed and about 5 frames/sec when both tracking and recognition are performed without using any special hardware. key words: vision system, dual viewing angles, speech dialogue system, motion tracking, mouth pattern recognition