Improving record matching in imprecise and uncertain datasets

Improving record matching in imprecise and uncertain datasets
复制标题

DOI:
10.1093/llc/fqs028
复制
发表时间:
2012-12
期刊:
Lit. Linguistic Comput.
影响因子:
--
通讯作者:
David Croft
David Croft
中科院分区:
其他
文献类型:
--
作者:
David Croft

文献摘要

相似文献

博物馆藏品是一个极具挑战性的搜索空间。本文提出了一种新的方法,共指记录识别,这是适用于跨多个单独的集合。所提出的方法是为了适合使用,尽管高度不精确/不确定的属性值的记录。人们希望这可以通过从概率记录链接,文档分类和模糊聚类领域的方面的组合来实现。.................................................................................................................................................................................
Museum collections represent a highly challenging search space. This article proposes a novel approach for co-referent record identification which is suitable for use across multiple separate collections. The proposed approach is intended to be suitable for use despite highly imprecise/uncertain attribute values in the records. It is hoped that this can be achieved through a combination of aspects from the fields of probabillistic record linkage, document classification, and fuzzy clustering. .................................................................................................................................................................................