Ontology-based information extraction: An introduction and a survey of current approaches

Ontology-based information extraction: An introduction and a survey of current approaches
复制标题

DOI:
10.1177/0165551509360123
复制
发表时间:
2010-06-01
影响因子:
2.4
通讯作者:
Dou, Dejing
Dou, Dejing
中科院分区:
计算机科学3区
文献类型:
--
作者:
Wimalasuriya, Daya C.;Dou, Dejing

文献摘要

被引文献

相似文献

信息提取(IE)旨在通过自动处理从自然语言文本中检索某些类型的信息。例如,IE 系统可能会从一组网页中检索有关国家地缘政治指标的信息,而忽略其他类型的信息。基于本体的信息提取(OBIE)最近作为信息提取的一个子领域出现。在这里,本体——提供概念化的正式和明确的规范——在 IE 过程中发挥着至关重要的作用。由于本体论的使用,该领域与知识表示相关,并且有潜力协助语义网的发展。在本文中,我们介绍了基于本体的信息提取,并回顾了迄今为止开发的不同 OBIE 系统的细节。我们试图确定这些系统之间的共同架构,并根据不同因素对它们进行分类,从而更好地理解它们的操作。我们还讨论了这些系统的实现细节,包括它们使用的工具以及用于衡量其性能的指标。此外,我们尝试确定该领域未来可能的方向。
Information extraction (IE) aims to retrieve certain types of information from natural language text by processing them automatically. For example, an IE system might retrieve information about geopolitical indicators of countries from a set of web pages while ignoring other types of information. Ontology-based information extraction (OBIE) has recently emerged as a subfield of information extraction. Here, ontologies - which provide formal and explicit specifications of conceptualizations - play a crucial role in the IE process. Because of the use of ontologies, this field is related to knowledge representation and has the potential to assist the development of the Semantic Web. In this paper, we provide an introduction to ontology-based information extraction and review the details of different OBIE systems developed so far. We attempt to identify a common architecture among these systems and classify them based on different factors, which leads to a better understanding on their operation. We also discuss the implementation details of these systems including the tools used by them and the metrics used to measure their performance. In addition, we attempt to identify the possible future directions for this field.