Automatic Metadata Generation in an Archaeological Digital Library: Semantic Annotation of Grey Literature

Automatic Metadata Generation in an Archaeological Digital Library: Semantic Annotation of Grey Literature
复制标题

考古数字图书馆中的自动元数据生成:灰色文献的语义注释

DOI:
10.1007/978-3-642-34399-5_10
复制
发表时间:
2013
影响因子:
1
通讯作者:
D. Tudhope
D. Tudhope
中科院分区:
农林科学4区
文献类型:
--
作者:
A. Vlachidis;C. Binding;K. May;D. Tudhope

文献摘要

被引文献

相似文献

本文讨论了从考古数据服务灰色文献库(OASIS)的发掘报告中自动生成丰富元数据的问题。这项工作是星星项目的一部分,与英国遗产。CIDOC CRM本体的考古领域的扩展作为一个核心本体。丰富的元数据自动提取灰色文献,指导CRM,通过三个阶段的语义丰富的过程中采用GATE工具包与定制的规则和知识资源。本文展示了潜在的结合知识为基础的资源(本体论和叙词表)在信息提取,并提供自动提取的元数据作为XML注释加上灰色文献报告和RDF图从内容解耦的技术。从两个消费应用程序的例子进行了讨论,Andronikos门户网站,它提供了注释的XML文件的目视检查和星星项目,研究示范,它提供了统一的搜索考古发掘数据和灰色文献通过核心本体CRM-EH。
This paper discusses the automatic generation of rich metadata from excavation reports from the Archaeological Data Service library of grey literature (OASIS). The work is part of the STAR project, in collaboration with English Heritage. An extension of the CIDOC CRM ontology for the archaeological domain acts as a core ontology. Rich metadata is automatically extracted from grey literature, directed by the CRM, via a three phase process of semantic enrichment employing the GATE toolkit augmented with bespoke rules and knowledge resources. The paper demonstrates the potential of combining knowledge based resources (ontologies and thesauri) in information extraction, and techniques for delivering the automatically extracted metadata as XML annotations coupled with the grey literature reports and as RDF graphs decoupled from content. Examples from two consuming applications are discussed, the Andronikos web portal which serves the annotated XML files for visual inspection and the STAR project, research demonstrator which offers unified search across of archaeological excavation data and grey literature via the core ontology CRM-EH.