A Deep Reinforcement Learning Framework for Identifying Funny Scenes in Movies
A Deep Reinforcement Learning Framework for Identifying Funny Scenes in Movies
复制标题
DOI:
10.1109/icassp.2018.8462686
复制
发表时间:
2018-04
期刊:
影响因子:
--
通讯作者:
Haoqi Li;Naveen Kumar;Ruxin Chen;P. Georgiou
中科院分区:
文献类型:
--
作者:
Haoqi Li;Naveen Kumar;Ruxin Chen;P. Georgiou
This paper presents a novel deep Reinforcement Learning (RL) framework for classifying movie scenes based on affect using the face images detected in the video stream as input. Extracting affective information from the video is a challenging task modulating complex visual and temporal representations intertwined with the complex aspects of human perception and information integration. This also makes it difficult to collect a large annotated corpus restricting the use of supervised learning methods. We present an alternative learning framework based on RL that is tolerant to label sparsity and can easily make use of any available ground truth in an online fashion. We employ this modified RL model for the binary classification of whether a scene is funny or not on a dataset of movie scene clips. The results show that our model correctly predicts 72.95% of the time on the 2–3 minute long movie scenes while on shorter scenes the accuracy obtained is 84.13%.