An Accurate Scene Segmentation Method Based on Graph Analysis Using Object Matching and Audio Feature

An Accurate Scene Segmentation Method Based on Graph Analysis Using Object Matching and Audio Feature
复制标题

基于对象匹配和音频特征的图分析的精确场景分割方法

DOI:
10.1587/transfun.e92.a.1883
复制
发表时间:
2009
期刊:
IEICE Trans. Fundam. Electron. Commun. Comput. Sci.
影响因子:
--
通讯作者:
M. Haseyama
M. Haseyama
中科院分区:
--
文献类型:
--
作者:
Makoto Yamamoto;M. Haseyama

文献摘要

参考文献

被引文献

相似文献

提出了一种利用目标匹配和音频特征得到的两种有向图进行精确场景分割的方法。通常,在诸如广播节目和电影的视听材料中,存在包括相同背景、对象或地点的帧的相似镜头的重复出现,并且这样的镜头被包括在单个场景中。基于这一思想已经提出了许多场景分割方法,然而,由于它们使用颜色信息作为视觉特征,如果颜色特征在不同的镜头中因缩放和平移等相机操作而包括相同对象的帧中发生变化,则它们不能提供准确的场景分割结果。为了解决这一问题,采用两种新的方法实现了基于该方法的场景分割。在第一种方法中,在分别包括在不同镜头中的两个帧之间执行对象匹配。通过使用这些匹配结果,可以成功地找到包括相同对象的帧的镜头的重复出现,并将其表示为有向图。在第二种方法中,该方法还生成另一个有向图来表示具有相似音频特征的镜头的重复出现。通过结合使用这两种有向图,可以避免由于只使用一种图而导致的场景分割精度的下降,从而实现准确的场景分割。将该方法应用于实际的广播节目,实验结果验证了该方法的有效性。
A method for accurate scene segmentation using two kinds of directed graph obtained by object matching and audio features is proposed. Generally, in audiovisual materials, such as broadcast programs and movies, there are repeated appearances of similar shots that include frames of the same background, object or place, and such shots are included in a single scene. Many scene segmentation methods based on this idea have been proposed; however, since they use color information as visual features, they cannot provide accurate scene segmentation results if the color features change in different shots for which frames include the same object due to camera operations such as zooming and panning. In order to solve this problem, scene segmentation by the proposed method is realized by using two novel approaches. In the first approach, object matching is performed between two frames that are each included in different shots. By using these matching results, repeated appearances of shots for which frames include the same object can be successfully found and represented as a directed graph. The proposed method also generates another directed graph that represents the repeated appearances of shots with similar audio features in the second approach. By combined use of these two directed graphs, degradation of scene segmentation accuracy, which results from using only one kind of graph, can be avoided in the proposed method and thereby accurate scene segmentation can be realized. Experimental results performed by applying the proposed method to actual broadcast programs are shown to verify the effectiveness of the proposed method.
使用 PCA、MGD 和模糊算法进行视听索引的基于音频的镜头分类
DOI: --
发表时间: 2007
期刊: IEICE Trans. Fundamentals E90-A(8)
影响因子: --
作者:
Nitanda;N. and Haseyama;M.
通讯作者: M.