A Retrieval Method of Similar Strings Using Substrings

A Retrieval Method of Similar Strings Using Substrings
复制标题

一种利用子串的相似字符串检索方法

DOI:
10.1109/iccea.2010.210
复制
发表时间:
2010
期刊:
2010 Second International Conference on Computer Engineering and Applications
影响因子:
--
通讯作者:
J. Aoe
J. Aoe
中科院分区:
--
文献类型:
--
作者:
M. Fuketa;Nobuo Fujisawa;H. Bando;K. Morita;J. Aoe

文献摘要

被引文献

相似文献

随着互联网的快速发展,检索与输入字符串相似的字符串的需求不断增加。编辑距离有助于从大量数据中检索出必要的信息;编辑距离是两个字符串的相似度。而且,输入字符串必须与字典中的所有字符串进行比较,以便通过计算编辑距离来检索相似的字符串,这个过程是非常耗时的工作。本文提出了一种有效检索相似字符串的字典结构和一种高速输出字符串编辑距离降序的方法。该方法可以检索所有相似的字符串,并且该方法的速度比n-gram方法更快。
With the rapid growth of the Internet, demand for retrieving similar strings to an input string has been increasing. Edit distance is helpful to retrieve necessary information from the large amount of data; edit distance is the similarity of two strings. Moreover, an input string must be compared with all strings in a dictionary in order to retrieve similar strings by computing edit distance, and this process is very time-consuming work. This paper proposes a structure of a dictionary for retrieving similar strings effectively and a method to output edit distance of strings in descending order at high speed. This method can retrieve all similar strings and the speed of this method is faster than that of a n-gram method.