XSemantic: An Extension of LCA Based XML Semantic Search

XSemantic: An Extension of LCA Based XML Semantic Search
复制标题

DOI:
10.1587/transinf.e92.d.1079
复制
发表时间:
2009-05
期刊:
IEICE Trans. Inf. Syst.
影响因子:
--
通讯作者:
Umaporn Supasitthimethee;Toshiyuki Shimizu;Masatoshi Yoshikawa;Kriengkrai Porkaew
Umaporn Supasitthimethee;Toshiyuki Shimizu;Masatoshi Yoshikawa;Kriengkrai Porkaew
中科院分区:
其他
文献类型:
--
作者:
Umaporn Supasitthimethee;Toshiyuki Shimizu;Masatoshi Yoshikawa;Kriengkrai Porkaew

文献摘要

相似文献

查询XML数据最方便的方法之一是关键字搜索,因为它不需要任何XML结构知识或学习新的用户界面。关键词搜索是模糊的。用户可以使用不同的术语来搜索相同的信息。此外,系统难以决定哪个节点可能被选为返回节点以及在结果中应包括多少信息。为了解决这些挑战,我们提出了一个XML语义搜索的基础上称为XSemantic的关键字。一方面,我们给出了三个定义,以完成在语义上。首先,在语义词扩展方面,利用领域本体对歧义关键词进行了扩展,提高了系统的鲁棒性。其次,为了返回语义上有意义的答案,我们从用户查询中自动推断返回信息,并利用最短路径返回关键字之间有意义的连接。第三,我们提出的语义排名,反映了程度的相似性,以及语义关系,使搜索结果具有较高的相关性首先呈现给用户。另一方面,在LCA和邻近搜索方法中,我们研究了搜索结果中包含的信息的问题。因此,我们引入了最低公共元素祖先(LCEA)的概念,并定义了我们的简单规则,而不需要任何模式信息,如DTD或XML模式。第一个实验表明,XSemantic不仅正确地推断返回信息,而且产生紧凑的有意义的结果。此外,我们提出的语义的好处证明了第二个实验。
One of the most convenient ways to query XML data is a keyword search because it does not require any knowledge of XML structure or learning a new user interface. However, the keyword search is ambiguous. The users may use different terms to search for the same information. Furthermore, it is difficult for a system to decide which node is likely to be chosen as a return node and how much information should be included in the result. To address these challenges, we propose an XML semantic search based on keywords called XSemantic. On the one hand, we give three definitions to complete in terms of semantics. Firstly, the semantic term expansion, our system is robust from the ambiguous keywords by using the domain ontology. Secondly, to return semantic meaningful answers, we automatically infer the return information from the user queries and take advantage of the shortest path to return meaningful connections between keywords. Thirdly, we present the semantic ranking that reflects the degree of similarity as well as the semantic relationship so that the search results with the higher relevance are presented to the users first. On the other hand, in the LCA and the proximity search approaches, we investigated the problem of information included in the search results. Therefore, we introduce the notion of the Lowest Common Element Ancestor (LCEA) and define our simple rule without any requirement on the schema information such as the DTD or XML Schema. The first experiment indicated that XSemantic not only properly infers the return information but also generates compact meaningful results. Additionally, the benefits of our proposed semantics are demonstrated by the second experiment.