A robust video text extraction method for character recognition

A robust video text extraction method for character recognition
复制标题

一种用于字符识别的鲁棒视频文本提取方法

DOI:
10.1002/scj.10148
复制
发表时间:
2005
期刊:
Systems and Computers in Japan
影响因子:
--
通讯作者:
Takeshi Mita
Takeshi Mita
中科院分区:
--
文献类型:
--
作者:
O. Hori;Takeshi Mita

文献摘要

被引文献

相似文献

本文提出了一种高精度提取视频图像中的文字部分的方法,用于OCR的阅读。过去的研究已经产生了基于阈值的二值化从背景中提取视频图像中的文本的方法,利用文本的强度高于背景的强度的事实。用于确定阈值的一种方法是Shio应用大津方法,假设背景和字符的两个强度在局部块中的分布。然而,基于诸如视频图像的背景的各种强度的方法具有由于不一定有效的假设而不能产生好的阈值的问题。此外,在现实中,它们不能以足以OCR可读性的精度提取字符,因为由于阴影、边缘消除和信号转换处理的影响,字符周围的强度不一定高。因此,本文提出了一种通过稳健地估计文本部分的强度分布,初始提取高可靠性区域作为文本部分,并基于估计的分布扩展区域来仅提取文本部分的方法。实验结果表明,该方法比传统方法具有更高的准确率和更好的OCR可读性。© 2005威利期刊有限公司Syst Comp Jpn,36(9):87-96,2005;在线发表于Wiley InterScience(www.interscience.wiley.com)。DOI 10.1002/scj.10148
This paper proposes a method for extracting text portions occurring in video images with high accuracy for reading by OCR. Past studies have produced methods of extracting text in a video image from its background by binarization based on a threshold, utilizing the fact that the intensity of the text is higher than that of the background. One method for determining the threshold is Shio's application of Otsu's method, assuming the distribution of two intensities of the background and characters in local blocks. However, methods based on various intensities of the background such as those of video images have the problem of not yielding a good threshold due to assumptions that are not necessarily valid. In addition, in reality, they cannot extract characters with accuracy sufficient for OCR readability because the intensity around the characters is not necessarily high due to the effects of shadowing, edge elimination, and signal conversion processing. Thus, this paper proposes a method of extracting only the text portions by robustly estimating the intensity distribution of the text portions, initially extracting high‐reliability areas as text portions, and extending the areas based on the estimated distribution. Experimental results show that the proposed method extracts text portions with higher accuracy and better OCR readability than the conventional methods. © 2005 Wiley Periodicals, Inc. Syst Comp Jpn, 36(9): 87–96, 2005; Published online in Wiley InterScience (www.interscience.wiley.com). DOI 10.1002/scj.10148