3-D Video Representation Using Depth Maps

3-D Video Representation Using Depth Maps
复制标题

DOI:
10.1109/jproc.2010.2091090
复制
发表时间:
2011-04
影响因子:
20.6
通讯作者:
K. Müller;P. Merkle;T. Wiegand
K. Müller;P. Merkle;T. Wiegand
中科院分区:
计算机科学1区
文献类型:
--
作者:
K. Müller;P. Merkle;T. Wiegand

文献摘要

被引文献

相似文献

目前的3d视频(3DV)技术是基于立体系统的。这些系统使用立体声视频编码,由两个输入摄像头传送图像。通常情况下,这种立体系统只能在接收器上再现这两个摄像头的视图,而要让多人观看立体显示器,则需要佩戴特殊的3-D眼镜。另一方面,新兴的自动立体多视图显示器可以发出大量的视图,使多个用户无需3d眼镜即可观看3d。为了表示大量的视图,使用立体视频编码的多视图扩展,通常需要与视图数成比例的比特率。然而,由于多视图显示器的质量提高将取决于发射视图的增加,因此需要一种格式,允许在传输比特率恒定的情况下生成任意数量的视图。这种格式是视频信号和相关深度图的组合。深度图提供与视频信号的每个样本相关的差异,可用于通过视图合成渲染任意数量的附加视图。本文介绍了视频和深度数据的有效编码方法。对于视图的生成,提出了一种综合方法,减轻了深度估计和编码带来的误差。
Current 3-D video (3DV) technology is based on stereo systems. These systems use stereo video coding for pictures delivered by two input cameras. Typically, such stereo systems only reproduce these two camera views at the receiver and stereoscopic displays for multiple viewers require wearing special 3-D glasses. On the other hand, emerging autostereoscopic multiview displays emit a large numbers of views to enable 3-D viewing for multiple users without requiring 3-D glasses. For representing a large number of views, a multiview extension of stereo video coding is used, typically requiring a bit rate that is proportional to the number of views. However, since the quality improvement of multiview displays will be governed by an increase of emitted views, a format is needed that allows the generation of arbitrary numbers of views with the transmission bit rate being constant. Such a format is the combination of video signals and associated depth maps. The depth maps provide disparities associated with every sample of the video signal that can be used to render arbitrary numbers of additional views via view synthesis. This paper describes efficient coding methods for video and depth data. For the generation of views, synthesis methods are presented, which mitigate errors from depth estimation and coding.