DCVGAN: Depth Conditional Video Generation
DCVGAN: Depth Conditional Video Generation
复制标题
DOI:
10.1109/icip.2019.8803764
复制
发表时间:
2019-09
期刊:
影响因子:
--
通讯作者:
Yuki Nakahira;K. Kawamoto
中科院分区:
文献类型:
--
作者:
Yuki Nakahira;K. Kawamoto
In the past few years, several generative adversarial networks (GANs) for video generation have been proposed although most of them only use color videos to train the generative model. However, to make the model understand scene dynamics more accurately, not only optical information but also three-dimensional geometrical information is important. In this paper, using depth video together with color video, we propose a GAN architecture for video generation. In the generator of our architecture, the depth video is generated in the first half and in the second half, the color video is generated by solving the domain translation from the depth to the color. By modeling the scene dynamics with a focus on the depth information, we were able to produce videos of higher quality than the conventional method. Furthermore, we show that our method produces better video samples than ones by conventional method in terms of both variety and quality when evaluating on facial expression and hand gesture datasets. The codes and generated sample videos are publicly available on Github1.