Fast Plagiarism Detection Based on Simple Document Similarity
Fast Plagiarism Detection Based on Simple Document Similarity
复制标题
基于简单文档相似度的快速抄袭检测
DOI:
10.1109/icdim.2017.8244662
复制
发表时间:
2017
期刊:
影响因子:
--
通讯作者:
Kensuke Baba
中科院分区:
文献类型:
--
作者:
Arief MAULANA;Kazumi SAITO;Tetsuo IKEDA;Hiroaki YUZE;Kensuke Baba
Plagiarism detection in a large number of documents requires efficient methods. This paper proposes a plagiarism detection algorithm based on approximate string matching to be specified in “copy and paste”-type plagiarisms, and a speed improvement to an implementation of the algorithm. Most of the computations required in the algorithm are omitted by two kinds of approximations of the output used for plagiarism detection, while the decrease of accuracy caused by the approximations is acceptable. The effect of the improvement on the processing time and accuracy of the algorithm is evaluated by conducting experiments with a data set. The experimental results show that the improvement can reduce the processing time to approximately one-twentieth for a 6.4% decrease of the accuracy from those for the normal implementation of the algorithm.