Audio characterization for video indexing

Audio characterization for video indexing
复制标题

用于视频索引的音频表征

DOI:
10.1117/12.234776
复制
发表时间:
1996
期刊:
--
影响因子:
--
通讯作者:
I. Sethi
I. Sethi
中科院分区:
--
文献类型:
--
作者:
Nilesh V. Patel;I. Sethi

文献摘要

被引文献

相似文献

视频数据库面临的主要问题是,一旦剪切边界已被确定的视频剪辑的内容表征。目前在这个方向上的努力仅仅集中在图像信息的使用上,从而忽略了内容信息的重要补充源,即嵌入的音频或声轨。当前音频处理的研究可以很容易地应用于创建许多不同的视频索引,用于视频点播(VOD),教育视频索引,体育视频特征等。MPEG是一种新兴的视频和音频压缩标准,在多媒体行业中迅速普及。压缩比特流的处理已经得到了研究者们的认可。我们还演示了MPEG压缩视频中的特征提取,该特征提取实现了压缩视频上的大多数场景变化检测方案。在本文中,我们研究了音频信息的内容表征的潜力,演示了音频处理中广泛使用的功能,直接从压缩数据流和它们的应用程序的视频剪辑分类的提取。
The major problem facing video databases is that of content characterization of video clips once the cut boundaries have been determined. The current efforts in this direction are focussed exclusively on the use of pictorial information, thereby neglecting an important supplementary source of content information, i.e. the embedded audio or sound track. The current research in audio processing can be readily applied to create many different video indices for use in Video On Demand (VOD), educational video indexing, sports video characterization, etc. MPEG is an emerging video and audio compression standard with rapidly increasing popularity in multimedia industry. Compressed bit stream processing has gained good recognition among the researchers. We have also demonstrated feature extraction in MPEG compressed video which implements a majority of scene change detection schemes on compressed video. In this paper, we examine the potential of audio information for content characterization by demonstrating the extraction of widely used features in audio processing directly from compressed data stream and their application to video clip classification.