Deep Convolutional Neural Networks for Efficient Pose Estimation in Gesture Videos

Deep Convolutional Neural Networks for Efficient Pose Estimation in Gesture Videos
复制标题

DOI:
10.1007/978-3-319-16865-4_35
复制
发表时间:
2014-11
影响因子:
4
通讯作者:
Tomas Pfister;K. Simonyan;James Charles;Andrew Zisserman
Tomas Pfister;K. Simonyan;James Charles;Andrew Zisserman
中科院分区:
工程技术2区
文献类型:
--
作者:
Tomas Pfister;K. Simonyan;James Charles;Andrew Zisserman

文献摘要

被引文献

相似文献

Our objective is to efficiently and accurately estimate the upper body pose of humans in gesture videos. To this end, we build on the recent successful applications of deep convolutional neural networks (ConvNets). Our novelties are: (i) our method is the first to our knowledge to use ConvNets for estimating human pose in videos; (ii) a new network that exploits temporal information from multiple frames, leading to better performance; (iii) showing that pre-segmenting the foreground of the video improves performance; and (iv) demonstrating that even without foreground segmentations, the network learns to abstract away from the background and can estimate the pose even in the presence of a complex, varying background.We evaluate our method on the BBC TV Signing dataset and show that our pose predictions are significantly better, and an order of magnitude faster to compute, than the state of the art [3].