Semantic transcoding of video based on regions of interest

Semantic transcoding of video based on regions of interest
复制标题

基于感兴趣区域的视频语义转码

DOI:
10.1117/12.503081
复制
发表时间:
2003
期刊:
--
影响因子:
--
通讯作者:
Kyeong
Kyeong
中科院分区:
--
文献类型:
--
作者:
Jeongyeon Lim;Munchurl Kim;Jong;Kyeong

文献摘要

被引文献

相似文献

对多媒体的传统转码已经从用户终端能力(诸如显示器尺寸和解码处理能力)和网络资源(诸如可用网络带宽和服务质量(QoS)等)的角度来执行。多媒体内容的编码(或代码转换)到给定的这样的约束已经通过帧丢弃和视听的解码来实现,以及通过节省所得到的比特率来降低SNR(信噪比)值。不仅这样的传统转码是从用户的环境的角度来看,但我们也纳入了一种方法的语义转码的视听感兴趣的区域(ROI)从用户的角度来看。用户可以在图像或视频中指定他们感兴趣的部分,使得相应的视频内容可以被适配为聚焦于用户的ROI。我们将MPEG-21 DIA(数字项目自适应)的框架中,这样的语义信息的用户的ROI表示和交付给内容提供方作为XDI(上下文数字项目)。我们的用户感兴趣区域的语义信息的表示模式已被采用在MPEG-21 DIA适配模型。在本文中,我们提出了使用用户的ROI的语义信息进行转码,并显示我们的系统实现与实验结果。
Traditional transcoding on multimedia has been performed from the perspectives of user terminal capabilities such as display sizes and decoding processing power, and network resources such as available network bandwidth and quality of services (QoS) etc. The adaptation (or transcoding) of multimedia contents to given such constraints has been made by frame dropping and resizing of audiovisual, as well as reduction of SNR (Signal-to-Noise Ratio) values by saving the resulting bitrates. Not only such traditional transcoding is performed from the perspective of user’s environment, but also we incorporate a method of semantic transcoding of audiovisual based on region of interest (ROI) from user’s perspective. Users can designate their interested parts in images or video so that the corresponding video contents can be adapted focused on the user’s ROI. We incorporate the MPEG-21 DIA (Digital Item Adaptation) framework in which such semantic information of the user’s ROI is represented and delivered to the content provider side as XDI (context digital item). Representation schema of our semantic information of the user’s ROI has been adopted in MPEG-21 DIA Adaptation Model. In this paper, we present the usage of semantic information of user’s ROI for transcoding and show our system implementation with experimental results.