Text line and word segmentation of handwritten documents

Text line and word segmentation of handwritten documents
复制标题

DOI:
10.1016/j.patcog.2008.12.016
复制
发表时间:
2009-12-01
影响因子:
8
通讯作者:
Halatsis, C.
Halatsis, C.
中科院分区:
计算机科学1区
文献类型:
--
作者:
Louloudis, G.;Gatos, B.;Halatsis, C.

文献摘要

被引文献

相似文献

在本文中,我们提出了一种手写文档在其不同实体中的分割方法,即文本行和单词。文本线分割是通过对文档图像连接组件的子集应用霍夫变换实现的。后处理步骤包括纠正可能的假警报,检测霍夫变换未能创建的文本行,最后使用基于骨架化的新方法有效分离垂直连接的字符。分词作为两类问题来解决,文本行中相邻重叠组件之间的距离使用两个距离度量的组合来计算,并且在高斯混合建模框架中,每个组件都被分类为词间或词内距离。所提出的方法的性能基于一致和具体的评估方法,该方法使用合适的性能度量,以便将文本行分割和词分割结果与相应的基础真值注释进行比较。在两个不同的数据集上进行的实验证明了所提出方法的效率:(a)在ICDAR2007手写分割竞赛的测试集上,(b)在一组历史手写文档上。2009爱思唯尔有限公司版权所有。
In this paper, we present a segmentation methodology of handwritten documents in their distinct entities, namely, text lines and words. Text line segmentation is achieved by applying Hough transform on a subset of the document image connected components. A post-processing step includes the correction of possible false alarms, the detection of text lines that Hough transform failed to create and finally the efficient separation of vertically connected characters using a novel method based on skeletonization. Word segmentation is addressed as a two class problem, The distances between adjacent overlapped components in a text line are calculated using the combination of two distance metrics and each of them is categorized either as an inter- or an intra-word distance in a Gaussian mixture modeling framework. The performance of the proposed methodology is based on a consistent and concrete evaluation methodology that uses suitable performance measures in order to compare the text line segmentation and word segmentation results against the corresponding ground truth annotation. The efficiency of the proposed methodology is demonstrated by experimentation conducted on two different datasets: (a) on the test set of the ICDAR2007 handwriting segmentation competition and (b) on a set of historical handwritten documents. (C) 2009 Elsevier Ltd. All rights reserved.