Generalized framework for summarization of fixed-camera lecture videos by detecting and binarizing handwritten content
Generalized framework for summarization of fixed-camera lecture videos by detecting and binarizing handwritten content
复制标题
DOI:
10.1007/s10032-019-00327-y
复制
发表时间:
2019-06
期刊:
影响因子:
--
通讯作者:
B. Kota;Kenny Davila;Alexander Stone;S. Setlur;V. Govindaraju
中科院分区:
文献类型:
--
作者:
B. Kota;Kenny Davila;Alexander Stone;S. Setlur;V. Govindaraju
We propose a framework to extract and binarize handwritten content in lecture videos. The extracted content could potentially be used to index video collections powering content-based search and navigation within lecture videos helping students and educators across the world. A deep learning pipeline is used to detect handwritten text, formulae and sketches and then binarize the extracted content. We exploit the spatio-temporal structure of our binarized detections to compute associativity information of content across all video frames. This information is later used to segment the video. Experiments are conducted to compare the performance of key components of our framework in isolation, as well as the impact on overall performance, with respect to existing methods. We evaluate our framework on the publicly available AccessMath lecture video dataset obtaining anf-measure offor binary connected components. Code for the framework (including trained weights) and summarization will be released.