An Overview of Multimodal Video Representation for Semantic Analysis

An Overview of Multimodal Video Representation for Semantic Analysis
复制标题

用于语义分析的多模态视频表示概述

DOI:
10.1049/ic.2005.0708
复制
发表时间:
2006
期刊:
影响因子:
--
通讯作者:
Y. Kompatsiaris
Y. Kompatsiaris
中科院分区:
--
文献类型:
--
作者:
J. Calic;N. Campbell;S. Dasiopoulou;Y. Kompatsiaris

文献摘要

被引文献

相似文献

本文概述了针对基于内容的索引和检索的语义分析的视频表示方法。它突出了现有方法的主要成就,并对尚未解决的挑战提出了新的见解。对数字多媒体的自适应表示问题进行了批判性的评估,并提出了一些新的想法。此外,本文还对视频多模态的概念进行了重新评价和定义,以便引入编辑技术等模态。对所涉及的主题进行了广泛的文献调查。
This paper gives an overview of approaches to video representation targeting semantic analysis for content-based indexing and retrieval. It highlights the major achievements of the existing methodologies and sheds new light to the challenges that are still unsolved. The problem of adaptive representation of digital multimedia is critically assessed and some novel ideas are presented. In addition, the concept of video multimodality is reevaluated and redefined in order to introduce the modalities like editing technique. An extensive literature survey on the topics involved is given.